GGUF quantized versions of the Qwen3.5-2B model, optimized by Unsloth for fast local inference. Compatible with llama.cpp, Ollama, and other GGUF runtimes. Suitable for edge devices or low-memory environments while retaining strong language modeling performance.
Qwen3.5 2B is a Foundation models & chat product. It focuses on running a capable 2B-parameter Qwen3.5 model locally with minimal resource usage via quantized GGUF weights. It is built as an open-source project for developers. Qwen3.5 2B is open source under the Open Source license. Qwen3.5 2B is available on the web, the command line, and API.
Behind Qwen3.5 2B is Unsloth, and the product first shipped in 2025. The category is crowded — PulseGate's index counts 25 comparable apps. Key capabilities include Small LLM, GGUF Format, and quantized. It exposes integrations via a public API.
Latest indexed changes and source events
unsloth/Qwen3.5-2B-GGUF verified by the PulseGate indexer
Other apps tracked under the same category.