NVIDIA Nemotron-3 Super is a 120B parameter language model with an A12B active parameter count. The FP8 version provides efficient inference while maintaining model quality. It uses a mixture-of-experts architecture and is designed for high-performance text generation tasks. The model is available on Hugging Face for developers and researchers.
NVIDIA Nemotron 3 Super 120B A12B is a Foundation models & chat product. It focuses on delivering high-performance language model capabilities with efficient inference through quantization and mixture-of-experts architecture. It is built as an open-source project for developers. NVIDIA Nemotron 3 Super 120B A12B is open source under the Apache-2.0 license. It runs on the command line and API.
NVIDIA builds and maintains NVIDIA Nemotron 3 Super 120B A12B, and the product first shipped in 2026. Development happens publicly on GitHub with 314 stars and 72 commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 5 similar tools. Key capabilities include Text Generation, mixture of Experts, and FP8 Quantization. It exposes integrations via a public API.
Latest indexed changes and source events
nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP8 verified by the PulseGate indexer
Other apps tracked under the same category.