This repository hosts GGUF quantized weights of Alibaba's Qwen3 8B model, optimized for use with llama.cpp, LM Studio, and other local LLM runners. It supports tool calling and follows the latest Qwen3 chat template. The model is intended for developers and enthusiasts who want to run a capable 8B-parameter LLM locally without relying on cloud APIs.
Qwen3 8B is a Foundation models & chat product. It focuses on running the Qwen3 8B model efficiently on consumer hardware using quantized GGUF files compatible with llama.cpp and similar engines. It is built as an open-source project for developers. Qwen3 8B is open source under the Open Source license. Qwen3 8B is available on the web, the command line, and API.
Maziyar Panahi builds and maintains Qwen3 8B, and the product first shipped in 2025. It operates in a well-populated space: PulseGate tracks 14 similar tools. Key capabilities include Quantized Model, GGUF Format, and Local Inference. It exposes integrations via a public API.
Latest indexed changes and source events
MaziyarPanahi/Qwen3-8B-GGUF verified by the PulseGate indexer
Other apps tracked under the same category.