Nemotron-Mini-4B-Instruct is a small yet powerful instruction-tuned model developed by NVIDIA. It is designed for low-latency inference on edge devices while maintaining strong reasoning and tool-use capabilities. The model uses a custom chat template and supports function calling.
Nemotron Mini 4B Instruct sits in PulseGate's Foundation models & chat category. It focuses on running capable language models efficiently on resource-constrained devices and edge hardware. It is built as an open-source project for developers. Nemotron Mini 4B Instruct is open source under the Apache-2.0 license. The product ships for the web, the command line, and API.
It is developed by NVIDIA (United States), and the product first shipped in 2023. Development happens publicly on GitHub with 8.5k stars and 254 commits in the last 90 days. Key capabilities include Instruction Following, On-device Inference, and Tool Integration.
Latest indexed changes and source events
nvidia/Nemotron-Mini-4B-Instruct verified by the PulseGate indexer
Other apps tracked under the same category.