Qwen2.5-72B-Instruct-GGUF is a quantized version of the Qwen2.5-72B-Instruct large language model provided in GGUF format on Hugging Face. It is distributed by bartowski and supports a specific chat template for instruction following and tool calling. The template defines behavior for system prompts, defaulting to the identity of Qwen created by Alibaba Cloud as a helpful assistant, and includes structured XML-based handling for function calls with JSON arguments when tools are supplied.
The repository contains the necessary prompt formatting logic to enable the model to process messages, insert tool definitions within designated XML tags, and generate tool calls in a precise format enclosed in tool_call tags. This implementation allows the model to operate with one or more functions during inference. The GGUF format itself is intended to facilitate local execution through compatible engines.
It belongs to the class of foundation models released for open use. No pricing, licensing terms, or specific hardware requirements are stated in the page content.
Qwen2.5 72B Instruct is a Text generation project. It focuses on running large language models locally without relying on cloud APIs. Qwen2.5 72B Instruct is an open-source project aimed at developers. The project is open source (MIT). It ships for the web and the command line, and it can be self-hosted.
bartowski builds and maintains Qwen2.5 72B Instruct, and it first shipped in 2023. Development happens publicly on GitHub with 121.2k stars and 1.2k commits in the last 90 days. It competes in a saturated segment with 25 similar projects in PulseGate's index. Among its 4 catalogued features are GGUF Format, Quantized Weights, and Tool Calling.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do