This is an NVIDIA-optimized variant of Meta's Llama 3.1 8B Instruct model using NVFP4 quantization for improved performance on compatible hardware. It supports chat and instruction-following use cases while maintaining a smaller memory footprint. The model is hosted on Hugging Face and intended for advanced AI development and deployment.
Llama 3.1 8B Instruct sits in PulseGate's Foundation models & chat category. It focuses on running efficient large language model inference with reduced precision. Llama 3.1 8B Instruct is an open-source project aimed at developers. The project is open source (Open Source). It runs on the web and API.
NVIDIA builds and maintains Llama 3.1 8B Instruct, and it first shipped in 2023. The project is developed in the open on GitHub with 14.2k stars and 1.9k commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 9 similar projects. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do