This repository offers a full range of GGUF quantized files for Alibaba's Qwen2-7B-Instruct model. Created by MaziyarPanahi, it supports common local LLM runtimes and provides options from low-bit IQ quantizations to higher-precision formats. The model is known for strong multilingual and reasoning performance in a compact size.
Qwen2 7B Instruct is a Foundation models & chat project. It focuses on running the Qwen2 7B Instruct model locally with flexible quantization levels for different performance and memory tradeoffs. It is built as an open-source project for developers. Qwen2 7B Instruct is open source under the MIT license. It ships for the web, the command line, and API, and it can be self-hosted.
It is developed by MaziyarPanahi, and it first shipped in 2023. The project is developed in the open on GitHub with 122.3k stars and 1.2k commits in the last 90 days. The category is crowded — PulseGate's index counts 25 comparable apps. Among its 4 catalogued features are chat templates, quantized variants, and instruction tuned. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do