This is a 4B parameter instruction-tuned version of the Qwen3 model, optimized with NVFP4A16 quantization. It supports advanced chat templates, tool calling, and multi-step reasoning. Available on Hugging Face for use with Transformers or inference providers, it targets developers needing lightweight yet capable language models.
Qwen3 4B Instruct 2507 NVFP4A16 is a Foundation models & chat project. It focuses on running efficient quantized large language models locally or via API for instruction and tool-use tasks. It is built as an open-source project for developers. The project is open source (Open Source). It runs on the web and API.
Behind Qwen3 4B Instruct 2507 NVFP4A16 is apolloparty, and it first shipped in 2025. The category is crowded — PulseGate's index counts 25 comparable apps. Key capabilities include Instruction Following, Tool Calling, and Chat Template. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do