This is a large DeBERTa V2 model pre-trained on Japanese text using character-level whole word masking. It is designed for Japanese natural language processing tasks including fill-mask, sequence classification, and token classification. The model is hosted on Hugging Face and can be used with the Transformers library.
Deberta V2 Large Japanese Char Wwm is a Foundation models & chat project. It focuses on performing masked language modeling and downstream NLP tasks on Japanese text with character-aware modeling. It is built as an open-source project for developers. Deberta V2 Large Japanese Char Wwm is open source under the Apache-2.0 license. Deberta V2 Large Japanese Char Wwm is available on the web and API.
It is developed by Kyoto University NLP Lab (Japan), and it first shipped in 2017. Development happens publicly on GitHub with 12k stars and 172 commits in the last 90 days. Key capabilities include fill-Mask, Japanese NLP, and Character-level Tokenization.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do