This is a 4-bit quantized version of the Qwen2.5-3B-Instruct model created by Unsloth. It enables efficient local inference of a capable instruction-tuned LLM. The model is suitable for developers seeking to deploy language models with reduced memory requirements while maintaining strong performance.
In the Quantised & converted weights space, Qwen2.5 3B Instruct Bnb takes a focused approach. It focuses on running large language models efficiently on consumer hardware with minimal performance loss. Qwen2.5 3B Instruct Bnb is an open-source project aimed at developers. The project is open source (Apache-2.0). Qwen2.5 3B Instruct Bnb is available on the web, the command line, and API.
Behind Qwen2.5 3B Instruct Bnb is unsloth, and it first shipped in 2023. The project is developed in the open on GitHub with 68.7k stars and 1.2k commits in the last 90 days. It competes in a saturated segment with 25 similar projects in PulseGate's index. Among its 3 catalogued features are Quantized Model, Instruction Tuning, and Efficient Inference.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do