This repository provides a 4-bit quantized MLX version of Gemma-4-12B-Instruct optimized for Apple silicon. It supports efficient local inference with the MLX framework while preserving instruction-following capabilities.
Gemma 4 12B It is a Foundation models & chat project. It focuses on running a capable 12B instruction model efficiently on Apple silicon hardware. Gemma 4 12B It is an open-source project aimed at developers. The project is open source (MIT). It ships for the web, the command line, and API.
lmstudio-community builds and maintains Gemma 4 12B It, and it first shipped in 2024. The project is developed in the open on GitHub with 5.3k stars and 524 commits in the last 90 days. The category is crowded — PulseGate's index counts 25 comparable apps. Among its 3 catalogued features are instruction tuning, 4-bit quantization, and MLX optimization.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do