Qwen3-4B-Instruct-2507-NVFP4 is a 4-billion parameter instruction-tuned model from the Qwen series, provided in an optimized NVFP4 quantized format. It supports advanced features such as tool calling and follows a chat template compatible with many inference engines. It is designed for developers seeking efficient local or cloud deployment of a capable language model.
Qwen3 4B Instruct 2507 sits in PulseGate's Foundation models & chat category. It focuses on running a capable 4B parameter instruction-tuned LLM with reduced memory and compute requirements. Qwen3 4B Instruct 2507 is an open-source project aimed at developers. The project is open source (Open Source). Qwen3 4B Instruct 2507 is available on the web and API.
llmat builds and maintains Qwen3 4B Instruct 2507, and the product first shipped in 2025. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 3 catalogued features are Instruction Following, Tool Calling, and Quantized Weights.
Latest indexed changes and source events
llmat/Qwen3-4B-Instruct-2507-NVFP4 verified by the PulseGate indexer