This is an AWQ (INT4) quantized version of Meta's Llama 3.1 70B Instruct model, optimized for reduced memory usage while maintaining performance. It includes a detailed chat template and is suitable for local inference. The model is provided by the hugging-quants organization on Hugging Face.
Meta Llama 3.1 70B Instruct sits in PulseGate's Foundation models & chat category. It focuses on running large 70B language models efficiently on consumer or enterprise hardware with reduced memory requirements. Meta Llama 3.1 70B Instruct is an open-source project aimed at AI developers and researchers. The project is open source (MIT). Meta Llama 3.1 70B Instruct is available on the command line and API.
It is developed by hugging-quants, and the product first shipped in 2023. The GitHub repository has been archived. PulseGate's similarity index places it among 13 comparable tools. Among its 3 catalogued features are Quantized Weights, Instruction Tuned, and Chat Template.
Latest indexed changes and source events
hugging-quants/Meta-Llama-3.1-70B-Instruct-AWQ-INT4 verified by the PulseGate indexer
⚠ Archived
Other apps tracked under the same category.