Qwen2.5-32B-Instruct-GPTQ-Int8 is an INT8-quantized version of Alibaba's Qwen2.5 32B Instruct model. It supports advanced features including tool calling and follows a detailed chat template. The model is designed for efficient inference while retaining strong reasoning and instruction-following capabilities.
Qwen2.5 32B Instruct sits in PulseGate's Foundation models & chat category. It focuses on running a high-performance 32B instruction model with reduced memory requirements via quantization. It is built as an open-source project for developers and researchers. Qwen2.5 32B Instruct is open source under the Open Source license. Qwen2.5 32B Instruct is available on the web, the command line, and API.
Behind Qwen2.5 32B Instruct is Alibaba Cloud, based in China, and the product first shipped in 2024. Development happens publicly on GitHub with 27.4k stars. The category is crowded — PulseGate's index counts 25 comparable apps. Key capabilities include Instruction Following, quantized, and Tool Calling.
Latest indexed changes and source events
Qwen/Qwen2.5-32B-Instruct-GPTQ-Int8 verified by the PulseGate indexer
Other apps tracked under the same category.