Qwen3-14B-FP8 is a quantized version of Alibaba's Qwen3 14B model optimized for efficient inference. It supports advanced features such as tool calling and follows an instruction-tuned chat template. The model is distributed on Hugging Face for developers who need high-performance open language models that can run on consumer or enterprise hardware.
In the Tool calling space, Qwen3 14B takes a focused approach. It focuses on running a powerful 14B language model with reduced memory requirements using FP8 quantization. It is built as an open-source project for AI developers and researchers. The project is open source (Open Source). It runs on the web, the command line, and API.
Qwen builds and maintains Qwen3 14B, and it first shipped in 2024. Development happens publicly on GitHub with 27.4k stars. It competes in a saturated segment with 25 similar projects in PulseGate's index. Among its 3 catalogued features are Large Language Model, Tool Calling, and Quantized Inference.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do