DeepSeek-V4-Flash-GGUF provides GGUF quantized files for the DeepSeek-V4 model optimized for fast inference. It includes advanced reasoning capabilities with explicit thinking tokens and tool-calling support. The model is distributed by Unsloth and is intended for local or self-hosted use with GGUF-compatible inference engines.
DeepSeek V4 Flash sits in PulseGate's Quantised & converted weights category. Efficiently running the large DeepSeek-V4 reasoning model locally using quantized weights. DeepSeek V4 Flash is an open-source project aimed at developers and AI practitioners. The project is open source (Apache-2.0). DeepSeek V4 Flash is available on the web, the command line, and API.
It is developed by Unsloth, and it first shipped in 2023. The project is developed in the open on GitHub with 68.4k stars and 1.2k commits in the last 90 days. PulseGate's similarity index places it among 11 comparable projects. Key capabilities include GGUF Format, Reasoning Model, and Tool Use.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do