This is a quantized and distilled version of the DeepSeek-R1 reasoning model based on Llama architecture. It is hosted on Hugging Face for download and local inference using libraries such as transformers or llama.cpp. The model supports advanced features including tool calling and is optimized for on-device or local deployment with significantly reduced memory footprint while retaining strong reasoning capabilities.
Deepseek R1 Distill Llama 70b sits in PulseGate's Foundation models & chat category. It focuses on running high-performance reasoning models locally with reduced memory requirements. Deepseek R1 Distill Llama 70b is an open-source project aimed at developers. The project is open source (Open Source). It runs on the web and API.
casperhansen builds and maintains Deepseek R1 Distill Llama 70b, and the product first shipped in 2025. Among its 3 catalogued features are quantized weights, chat template, and tool calling support.
Latest indexed changes and source events
casperhansen/deepseek-r1-distill-llama-70b-awq verified by the PulseGate indexer
Other apps tracked under the same category.