This repository hosts GGUF quantized weights of Alibaba's Qwen3 8B model, optimized for use with llama.cpp, LM Studio, and other local LLM runners. It supports tool calling and follows the latest Qwen3 chat template. The model is intended for developers and enthusiasts who want to run a capable 8B-parameter LLM locally without relying on cloud APIs.
In the Foundation models & chat space, Qwen3 8B takes a focused approach. It focuses on running the Qwen3 8B model efficiently on consumer hardware using quantized GGUF files compatible with llama.cpp and similar engines. Qwen3 8B is an open-source project aimed at developers. The project is open source (Open Source). It runs on the web, the command line, and API.
Behind Qwen3 8B is Maziyar Panahi, and it first shipped in 2025. PulseGate's similarity index places it among 14 comparable projects. Among its 3 catalogued features are Quantized Model, GGUF Format, and Local Inference. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do