NVIDIA Nemotron Nano 12B v2 GGUF is a quantized language model distributed through Hugging Face for local inference. It can be downloaded and run with compatible tools such as llama.cpp and other GGUF runtimes by developers and AI practitioners.
NVIDIA Nemotron Nano 12B sits in PulseGate's Foundation models & chat category. It focuses on running a capable language model locally without relying on a hosted inference service. It is built as an open-source project for developers and AI practitioners. The project is open source (MIT). It runs on the command line, Linux, macOS, and Windows, and it can be self-hosted.
Behind NVIDIA Nemotron Nano 12B is MaziyarPanahi, and it first shipped in 2023. The project is developed in the open on GitHub with 124.5k stars and 1.2k commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 13 similar projects. Key capabilities include GGUF quantization, local inference, and text generation.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do