This is a 4-bit quantized version of Google's Gemma 3 4B instruction-tuned (it) model, prepared by the MLX community. It is optimized specifically for Apple's MLX machine learning framework, enabling efficient inference on Mac computers. The model includes full chat templates and is suitable for local deployment of a capable instruction-following language model.
In the Quantised & converted weights space, Gemma 3 4b It Qat takes a focused approach. It focuses on running Google's Gemma 3 model efficiently on Apple silicon hardware using the MLX framework. It is built as an open-source project for developers. Gemma 3 4b It Qat is open source under the Open Source license. It ships for the web, the command line, and API.
It is developed by mlx-community. It operates in a well-populated space: PulseGate tracks 17 similar projects. Among its 4 catalogued features are Instruction Model, 4-bit Quantization, and MLX Optimized.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do