This is a GGUF quantized variant of the 20B parameter gpt-oss model hosted on Hugging Face. It enables efficient local inference using tools like llama.cpp or Ollama. The model is designed for developers who want to run capable open-weight LLMs on standard hardware with reduced memory requirements.
Gpt Oss 20b sits in PulseGate's Quantised & converted weights category. It focuses on running large language models efficiently on consumer hardware without high-end GPUs. It is built as an open-source project for developers. The project is open source (Apache-2.0). It runs on the web and the command line, and it can be self-hosted.
Behind Gpt Oss 20b is Unsloth, and it first shipped in 2023. Development happens publicly on GitHub with 68.7k stars and 1.2k commits in the last 90 days. PulseGate's similarity index places it among 5 comparable projects. Key capabilities include GGUF Quantization, Local Inference, and Model Conversion.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do