This repository contains GGUF quantized files for the Qwen3.5-4B model, making it easy to run locally using tools like LM Studio, llama.cpp, or Ollama. The 4B parameter model offers a good balance between performance and resource requirements for local inference. It supports standard chat templates and is suitable for various text generation and assistant applications.
Qwen3.5 4B is a Foundation models & chat project. It focuses on running a capable 4B parameter language model efficiently on consumer CPUs and GPUs using the GGUF format. Qwen3.5 4B is an open-source project aimed at developers. The project is open source (MIT). It runs on the web and the command line.
lmstudio-community builds and maintains Qwen3.5 4B, and it first shipped in 2023. Development happens publicly on GitHub with 121.8k stars and 1.2k commits in the last 90 days. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 4 catalogued features are GGUF Format, quantized, and 4B Parameters.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do