Qwen3.5-9B-NVFP4 is a community-quantized version of the Qwen 3.5 9B large language model, optimized for efficient local and on-device inference. It provides GGUF files suitable for tools like llama.cpp and Ollama, supporting text generation, multimodal inputs, and tool-calling capabilities. Primarily used by developers and AI researchers who self-host open-weight models.
Qwen3.5 9B is a Foundation models & chat project. It focuses on running high-performance quantized LLMs locally or on custom inference hardware without relying on proprietary cloud APIs. It is built as an open-source project for developers. The project is open source (Apache-2.0). It ships for the web, the command line, and API, and it can be self-hosted.
It is developed by ig1, and it first shipped in 2019. The project is developed in the open on GitHub with 3.6k stars and 168 commits in the last 90 days. It competes in a saturated segment with 25 similar apps in PulseGate's index. Key capabilities include Quantized GGUF, chat templates, and multimodal support. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do