NVIDIA-Nemotron-Nano-9B-v2-FP8 is a 9 billion parameter language model published on Hugging Face. It provides optimized FP8 weights for efficient inference while supporting advanced features such as tool calling and custom chat templates. The model is intended for developers who want to run or fine-tune capable open-weight LLMs either locally or through inference providers.
NVIDIA Nemotron Nano 9B sits in PulseGate's Tool calling category. It focuses on accessing and running a high-performance, quantized open-weight language model locally or via API. NVIDIA Nemotron Nano 9B is an open-source project aimed at developers. NVIDIA Nemotron Nano 9B is open source under the Open Source license. It ships for the command line and API.
NVIDIA builds and maintains NVIDIA Nemotron Nano 9B, and it first shipped in 2025. Key capabilities include model weights, chat template, and tool calling. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do