This is a 6-bit quantized GGUF variant of the Gemma 4 E2B instruction-tuned model, optimized for the MLX framework on Apple hardware. It allows efficient local inference of a capable language model on Macs and other Apple Silicon devices. The model includes a chat template and is suitable for local AI application development.
Gemma 4 E2B It sits in PulseGate's Quantised & converted weights category. It focuses on running large language models efficiently on Apple Silicon hardware with reduced memory usage. It is built as an open-source project for developers and researchers using Apple MLX. Gemma 4 E2B It is open source under the MIT license. Gemma 4 E2B It is available on the web, the command line, and API.
It is developed by lmstudio-community, and it first shipped in 2024. The project is developed in the open on GitHub with 5.2k stars and 419 commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 11 similar projects. Key capabilities include Quantized Model, Instruction Tuned, and MLX Compatibility.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do