FinBERT2 Large is a finance-domain Chinese language model hosted on Hugging Face. It addresses the need for specialized language understanding in Chinese financial texts by providing a pretrained transformer suitable for downstream natural language processing tasks in that domain.
The model contains 355 million parameters and is based on Chinese RoBERTa-Large. It was further pretrained on 32 billion tokens drawn from Chinese financial corpora that include research reports, news, and announcements. The model type is a transformer-based language model, with Chinese as its primary language. Developers can load it directly using the transformers library with AutoModel and AutoTokenizer from the repository valuesimplex-ai-lab/FinBERT2-large. It also supports continual pre-training or fine-tuning through resources available in an associated GitHub repository.
Value Simplex Technology Co. Ltd developed the model. It carries an Apache-2.0 license and is distributed in Safetensors format. Additional details appear in a linked research paper and GitHub repository. The parent model information is available under the chinese-roberta reference.
In the Other AI space, FinBERT2 Large takes a focused approach. It focuses on creating high-quality embeddings and representations for Chinese financial text including research reports, news, and announcements. FinBERT2 Large is an open-source project aimed at financial analysts and Chinese NLP researchers. The project is open source (MIT). It ships for the web and API.
Behind FinBERT2 Large is Value Simplex Technology Co. Ltd, and it first shipped in 2020. Development happens publicly on GitHub with 930 stars.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do