This is a 4-bit quantized version of the Qwen3 embedding model optimized for generating dense vector representations of text. It is compatible with the sentence-transformers library and supports tasks such as semantic similarity, retrieval, and clustering. The model is available on Hugging Face for easy integration into RAG and search systems.
Qwen3 Embedding 4B W4A16 G128 is a Foundation models & chat project. It focuses on generating high-quality dense vector embeddings for semantic search, clustering, and retrieval applications. It is built as an open-source project for developers. The project is open source (Open Source). It ships for the web, the command line, and API.
Behind Qwen3 Embedding 4B W4A16 G128 is boboliu, and it first shipped in 2025. The project is developed in the open on GitHub with 2k stars. PulseGate's similarity index places it among 9 comparable projects. Among its 4 catalogued features are Embedding Generation, Sentence Similarity, and Quantized Model.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do