This is a community-quantized 8-bit version of the Gemma 4 instruction-tuned model optimized for the MLX framework. It enables efficient local inference on Apple devices. The model includes a custom chat template and is distributed via Hugging Face for use with Python and MLX tooling.
Gemma 4 E2B It sits in PulseGate's Foundation models & chat category. It focuses on running the Gemma 4 model efficiently on Apple silicon hardware with reduced memory usage. It is built as an open-source project for developers. Gemma 4 E2B It is open source under the MIT license. Gemma 4 E2B It is available on the web, the command line, and API.
It is developed by lmstudio-community, and the product first shipped in 2024. Development happens publicly on GitHub with 5.2k stars and 419 commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 14 similar tools. Key capabilities include Quantized Inference, Instruction Tuned, and MLX Optimization.
Latest indexed changes and source events
lmstudio-community/gemma-4-E2B-it-MLX-8bit verified by the PulseGate indexer
Other apps tracked under the same category.