This is a community-quantized version of an open-source 20 billion parameter GPT model using MLX format (MXFP4-Q8). It is designed for efficient inference on Apple silicon devices. The model supports chat templates, tool calling, and can be used via Hugging Face Transformers or the MLX framework for local execution.
Gpt Oss 20b is a Foundation models & chat product. It focuses on running large language models efficiently on Apple silicon hardware with reduced memory usage. Gpt Oss 20b is an open-source project aimed at developers. The project is open source (Open Source). The product ships for the web, the command line, and API.
Behind Gpt Oss 20b is mlx-community, and the product first shipped in 2024. Among its 4 catalogued features are Model Quantization, MXFP4 Format, and Chat Template.
Latest indexed changes and source events
mlx-community/gpt-oss-20b-MXFP4-Q8 verified by the PulseGate indexer
Other apps tracked under the same category.