This is an AWQ-quantized version of NVIDIA's Nemotron-3 Nano 30B (A3B) model. It is a large language model optimized for efficient inference while preserving performance. The model is suitable for various text generation and instruction-following tasks and is compatible with the Hugging Face ecosystem.
In the Foundation models & chat space, NVIDIA Nemotron 3 Nano 30B A3B takes a focused approach. It focuses on running large language models efficiently on hardware with limited VRAM. NVIDIA Nemotron 3 Nano 30B A3B is an open-source project aimed at developers and AI researchers. The project is open source (Apache-2.0). It runs on the command line and API.
NVIDIA builds and maintains NVIDIA Nemotron 3 Nano 30B A3B, and the product first shipped in 2019. The project is developed in the open on GitHub with 3.6k stars and 168 commits in the last 90 days. Among its 4 catalogued features are AWQ Quantization, Large Language Model, and NVIDIA Nemotron.
Latest indexed changes and source events
stelterlab/NVIDIA-Nemotron-3-Nano-30B-A3B-AWQ verified by the PulseGate indexer
Other apps tracked under the same category.