This repository hosts GGUF quantized files for the Mistral-Small-24B-Instruct-2501 model. It offers various quantization formats (Q2_K to Q8_0 and FP16) enabling efficient local inference of this 24-billion-parameter instruction-tuned model. Targeted at developers and AI practitioners who prefer running capable open models on their own hardware or in self-hosted setups without relying on paid cloud APIs.
In the Foundation models & chat space, Mistral Small 24B Instruct 2501 takes a focused approach. It focuses on running the 24B-parameter Mistral Small Instruct model locally with manageable hardware requirements. It is built as an open-source project for developers. The project is open source (MIT). It runs on the web and the command line, and it can be self-hosted.
It is developed by MaziyarPanahi, and it first shipped in 2023. Development happens publicly on GitHub with 122.1k stars and 1.2k commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 11 similar projects. Key capabilities include GGUF Quantization, Instruct Model, and Multiple Quant Levels.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do