Qwen2.5-32B-Instruct-GPTQ-Int8 is an INT8-quantized version of Alibaba's Qwen2.5 32B Instruct model. It supports advanced features including tool calling and follows a detailed chat template. The model is designed for efficient inference while retaining strong reasoning and instruction-following capabilities.
Qwen2.5 32B Instruct sits in PulseGate's Foundation models & chat category. It focuses on running a high-performance 32B instruction model with reduced memory requirements via quantization. It is built as an open-source project for developers and researchers. Qwen2.5 32B Instruct is open source under the Open Source license. It runs on the web, the command line, and API.
It is developed by Alibaba Cloud (China), and it first shipped in 2024. Development happens publicly on GitHub with 27.4k stars. The category is crowded — PulseGate's index counts 25 comparable projects. Key capabilities include Instruction Following, quantized, and Tool Calling.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do