NVIDIA Nemotron 3.5 Lightning 30B A3B GGUF is an open-weight, quantized language model distributed in GGUF format for local inference. Developers can run it with compatible tools such as llama.cpp and other GGUF runtimes.
NVIDIA Nemotron 3.5 Lightning 30B A3B is a Foundation models & chat project. It focuses on running an open-weight language model locally without relying on a hosted inference service. NVIDIA Nemotron 3.5 Lightning 30B A3B is an open-source project aimed at developers running local language models. The project is open source (Open Source). NVIDIA Nemotron 3.5 Lightning 30B A3B is available on the command line, and it can be self-hosted.
It is developed by ggml-org, and it first shipped in 2026. Development happens publicly on GitHub with 15 stars and 124 commits in the last 90 days. PulseGate's similarity index places it among 6 comparable projects. Key capabilities include text generation, chat templates, and quantized weights.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do