Llama 3.1 70B Instruct FP8 is an open-weight large language model checkpoint optimized for instruction-following text generation. Developers can download and run it locally or deploy it with Python and Docker-based inference tooling.
Llama 3.1 70B Instruct sits in PulseGate's Foundation models & chat category. It focuses on running an instruction-following large language model locally or in self-hosted inference environments. It is built as an open-source project for AI developers and ML engineers. The project is open source (Apache-2.0). Llama 3.1 70B Instruct is available on the web, the command line, and API, and it can be self-hosted.
It is developed by NVIDIA, and it first shipped in 2024. Development happens publicly on GitHub with 3.4k stars and 349 commits in the last 90 days. PulseGate's similarity index places it among 16 comparable projects. Among its 6 catalogued features are instruction tuning, chat template, and tool calling.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do