DeepSeek-V4-Flash-GGUF provides GGUF quantized files for the DeepSeek-V4 model optimized for fast inference. It includes advanced reasoning capabilities with explicit thinking tokens and tool-calling support. The model is distributed by Unsloth and is intended for local or self-hosted use with GGUF-compatible inference engines.
DeepSeek V4 Flash sits in PulseGate's Foundation models & chat category. Efficiently running the large DeepSeek-V4 reasoning model locally using quantized weights. DeepSeek V4 Flash is an open-source project aimed at developers and AI practitioners. The project is open source (Apache-2.0). The product ships for the web, the command line, and API.
It is developed by Unsloth, and the product first shipped in 2023. The project is developed in the open on GitHub with 68.4k stars and 1.2k commits in the last 90 days. PulseGate's similarity index places it among 11 comparable tools. Among its 3 catalogued features are GGUF Format, Reasoning Model, and Tool Use.
Latest indexed changes and source events
unsloth/DeepSeek-V4-Flash-GGUF verified by the PulseGate indexer
Other apps tracked under the same category.