Qwen3-1.7B-FP8 is a quantized variant of Alibaba's Qwen3 series of large language models. It supports advanced features including tool calling and follows a chat template optimized for instruction following. The FP8 format enables faster inference with reduced memory requirements while maintaining strong performance.
In the Tool calling space, Qwen3 1.7B takes a focused approach. It focuses on running a capable 1.7B parameter LLM efficiently on consumer hardware using FP8 quantization. Qwen3 1.7B is an open-source project aimed at developers. The project is open source (Open Source). It ships for the web, the command line, and API.
Qwen builds and maintains Qwen3 1.7B, and it first shipped in 2024. The project is developed in the open on GitHub with 27.4k stars. The category is crowded — PulseGate's index counts 25 comparable projects. Key capabilities include Text Generation, Tool Calling, and Quantized Inference.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do