llama-3.3-70b-instruct-awq is a quantized version of Meta's Llama 3.3 70B Instruct model using the AWQ quantization method. This allows the large model to run with reduced memory requirements while maintaining performance. It includes a chat template optimized for instruction following and is compatible with Transformers and other inference frameworks.
Llama 3.3 70b Instruct sits in PulseGate's Foundation models & chat category. It focuses on running the large 70B Llama 3.3 model efficiently on consumer or enterprise hardware through quantization. Llama 3.3 70b Instruct is an open-source project aimed at developers and researchers running large language models locally. The project is open source (MIT). The product ships for the web, the command line, and API.
It is developed by casperhansen, and the product first shipped in 2023. The GitHub repository has been archived. PulseGate's similarity index places it among 6 comparable tools. Among its 3 catalogued features are quantized weights, 70B parameters, and instruction tuned.
Latest indexed changes and source events
casperhansen/llama-3.3-70b-instruct-awq verified by the PulseGate indexer
⚠ Archived
Other apps tracked under the same category.