This is a Japanese-specific Sentence-BERT model (version 2) based on the sentence-transformers framework. It generates high-quality embeddings for Japanese text that can be used for semantic similarity, clustering, and retrieval tasks. The model improves upon the first version with better loss functions and training data.
In the Embeddings & retrieval space, Sentence Bert Base Ja Mean Tokens takes a focused approach. It focuses on creating high-quality vector embeddings for Japanese sentences to enable semantic search and similarity tasks. Sentence Bert Base Ja Mean Tokens is an open-source project aimed at developers. The project is open source (Open Source). Sentence Bert Base Ja Mean Tokens is available on the web, the command line, and API.
Behind Sentence Bert Base Ja Mean Tokens is sonoisa, based in Japan, and it first shipped in 2022. Among its 3 catalogued features are Sentence Embeddings, Japanese Support, and Semantic Similarity.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do