FinBERT2 Large is a finance-domain Chinese language model hosted on Hugging Face. It addresses the need for specialized language understanding in Chinese financial texts by providing a pretrained transformer suitable for downstream natural language processing tasks in that domain.
The model contains 355 million parameters and is based on Chinese RoBERTa-Large. It was further pretrained on 32 billion tokens drawn from Chinese financial corpora that include research reports, news, and announcements. The model type is a transformer-based language model, with Chinese as its primary language. Developers can load it directly using the transformers library with AutoModel and AutoTokenizer from the repository valuesimplex-ai-lab/FinBERT2-large. It also supports continual pre-training or fine-tuning through resources available in an associated GitHub repository.
Value Simplex Technology Co. Ltd developed the model. It carries an Apache-2.0 license and is distributed in Safetensors format. Additional details appear in a linked research paper and GitHub repository. The parent model information is available under the chinese-roberta reference.
FinBERT2 Large sits in PulseGate's Other AI category. It focuses on creating high-quality embeddings and representations for Chinese financial text including research reports, news, and announcements. FinBERT2 Large is an open-source project aimed at financial analysts and Chinese NLP researchers. The project is open source (MIT). The product ships for the web and API.
It is developed by Value Simplex Technology Co. Ltd, and the product first shipped in 2020. The project is developed in the open on GitHub with 930 stars.
Latest indexed changes and source events
valuesimplex-ai-lab/FinBERT2-large verified by the PulseGate indexer
Other apps tracked under the same category.