Llama-3.1-8B-Instruct-4bit is a 4-bit MLX conversion of Meta's Llama 3.1 8B Instruct model for local inference on Apple silicon. Developers can download and run the quantised weights with MLX-compatible tooling.
Llama 3.1 8B Instruct is a Foundation models & chat project. It focuses on running an open-weight Llama language model locally with reduced memory requirements. Llama 3.1 8B Instruct is an open-source project aimed at developers and machine-learning practitioners. The project is open source (Open Source). Llama 3.1 8B Instruct is available on the command line, and it can be self-hosted.
It is developed by mlx-community. PulseGate's similarity index places it among 14 comparable projects. Key capabilities include 4-bit quantization, instruction tuning, and local inference.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do