This is an FP8 quantized checkpoint of NVIDIA's Llama-4 Scout 17B model with 16 experts. It is released under the NVIDIA Open Model License and supports commercial use and derivative works. The model is designed for instruction following and general language tasks.
Llama 4 Scout 17B 16E Instruct is a Text generation project. It focuses on running large-scale instruction-tuned language models efficiently with reduced precision on compatible hardware. It is built as an open-source project for developers. The project is open source (Open Source). It runs on the web and API.
NVIDIA builds and maintains Llama 4 Scout 17B 16E Instruct, and it first shipped in 2023. The project is developed in the open on GitHub with 14.2k stars and 1.9k commits in the last 90 days. Among its 3 catalogued features are mixture of Experts, Instruction Tuning, and Quantized Inference.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do