Qwen2.5-32B-Instruct-GPTQ-Int4 is a quantized variant of the Qwen2.5 32B instruction-tuned language model hosted on Hugging Face. It is provided as a GPTQ-Int4 model file intended for inference on compatible hardware. The model follows a system prompt that identifies it as Qwen created by Alibaba Cloud and positions it as a helpful assistant.
The page supplies a chat template that defines how the model processes messages. When a system message is present it uses that content; otherwise it defaults to stating that the model is Qwen created by Alibaba Cloud and is a helpful assistant. The template also includes explicit support for tool use. It instructs the model that it may call one or more functions to assist with a user query, supplies function signatures inside XML-style tools tags, and requires each function call to be returned as a JSON object wrapped in tool_call XML tags. This structure enables the model to handle tool calling and function calling formats during generation.
The model is distributed through the Hugging Face repository at Qwen/Qwen2.5-32B-Instruct-GPTQ-Int4. It belongs to the class of foundation models made available for download and local or hosted inference.
Qwen2.5 32B Instruct is a Foundation models & chat product. It focuses on running a high-performance 32-billion-parameter instruction model efficiently on consumer or enterprise hardware using 4-bit quantization. It is built as an open-source project for developers. Qwen2.5 32B Instruct is open source under the Open Source license. It runs on the web, the command line, and API.
Behind Qwen2.5 32B Instruct is Qwen, and the product first shipped in 2024. Development happens publicly on GitHub with 27.4k stars. The category is crowded — PulseGate's index counts 25 comparable apps. Key capabilities include Instruction Following, Tool Calling, and Quantized LLM.
Latest indexed changes and source events
Qwen/Qwen2.5-32B-Instruct-GPTQ-Int4 verified by the PulseGate indexer
Other apps tracked under the same category.