bert-large-japanese-v2 is a large BERT model pretrained on a combination of CC-100 and jawiki data using whole word masking and unidic-lite tokenization. It serves as a foundation model for various Japanese natural language processing tasks. Available through Hugging Face Transformers, it supports direct loading for fine-tuning or feature extraction in research and production environments.
In the Other AI space, Bert Large Japanese takes a focused approach. It focuses on providing a strong pretrained Japanese language model for downstream NLP tasks. Bert Large Japanese is an open-source project aimed at Japanese NLP researchers and developers. Bert Large Japanese is open source under the Apache-2.0 license. Bert Large Japanese is available on the web and API.
It is developed by Tohoku NLP (Japan), and it first shipped in 2018. The GitHub repository has been archived. Among its 3 catalogued features are Pretrained BERT, Japanese Tokenization, and Whole Word Masking. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do