This is an AWQ-quantized version of NVIDIA's Nemotron-3 Nano 30B (A3B) model. It is a large language model optimized for efficient inference while preserving performance. The model is suitable for various text generation and instruction-following tasks and is compatible with the Hugging Face ecosystem.
In the Foundation models & chat space, NVIDIA Nemotron 3 Nano 30B A3B takes a focused approach. It focuses on running large language models efficiently on hardware with limited VRAM. It is built as an open-source project for developers and AI researchers. The project is open source (Apache-2.0). It runs on the command line and API.
Behind NVIDIA Nemotron 3 Nano 30B A3B is NVIDIA, based in the United States, and it first shipped in 2019. Development happens publicly on GitHub with 3.6k stars and 168 commits in the last 90 days. Among its 4 catalogued features are AWQ Quantization, Large Language Model, and NVIDIA Nemotron.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do