This repository contains the FP8 quantized version of Meta's Llama-4-Maverick-17B-128E-Instruct model. It is an open-weights large language model optimized for instruction following, tool calling, and long-context tasks. The model can be used with the Hugging Face Transformers library, local inference engines, or cloud providers.
Llama 4 Maverick 17B 128E Instruct is a Foundation models & chat product. It focuses on deploying a capable, open-weight instruction-tuned language model with efficient memory usage for local or cloud inference. It is built as an open-source project for developers. Llama 4 Maverick 17B 128E Instruct is open source under the Open Source license. The product ships for the web, the command line, and API, and it can be self-hosted.
Meta builds and maintains Llama 4 Maverick 17B 128E Instruct, and the product first shipped in 2024. Development happens publicly on GitHub with 7.7k stars. Key capabilities include Instruction Following, Tool Use, and Long Context. It exposes integrations via a public API.
Latest indexed changes and source events
meta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8 verified by the PulseGate indexer
Other apps tracked under the same category.