An FP4 quantized and optimized version of the DeepSeek-R1 model provided by NVIDIA. It is designed for efficient inference on NVIDIA GPUs while maintaining reasoning capabilities. The model uses a custom chat template and is suitable for developers building AI applications that require strong reasoning performance with lower memory footprint.
DeepSeek R1 0528 NVFP4 sits in PulseGate's Foundation models & chat category. It focuses on running large reasoning models efficiently on NVIDIA hardware with reduced precision. DeepSeek R1 0528 NVFP4 is an open-source project aimed at developers. The project is open source (Apache-2.0). DeepSeek R1 0528 NVFP4 is available on the web, and it can be self-hosted.
Behind DeepSeek R1 0528 NVFP4 is NVIDIA, based in the United States, and the product first shipped in 2024. The project is developed in the open on GitHub with 3.3k stars and 354 commits in the last 90 days. Among its 3 catalogued features are Quantized Inference, Reasoning Model, and NVIDIA Optimized.
Latest indexed changes and source events
nvidia/DeepSeek-R1-0528-NVFP4-v2 verified by the PulseGate indexer
Other apps tracked under the same category.