Skip to content
Alternatives
Software like DeepSeek R1
What else does this job. Matched on what each project does, not on who links to whom.
Closest first
- DeepSeek R1huggingface.coDeepSeek-R1 is an open-source large language model developed by DeepSeek AI for text generation and conversational AI. It is suitable for developers and researchers building chatbots, virtual assistants, and other NLP applications.
- DeepSeek R1 0528huggingface.coDeepSeek-R1-0528 is an open-weight reasoning model released by DeepSeek. It excels at step-by-step reasoning and complex problem solving. The model is available for local download and cloud inference on Hugging Face.
- DeepSeekhuggingface.coDeepSeek-V3.1 is the latest iteration of DeepSeek's open foundation model series. It features a sophisticated chat template, support for tool calling, and strong performance on reasoning tasks. The model is fully open and can be run locally or deployed via standard inference frameworks.
- DeepSeek R1 Distill Qwen 32Bhuggingface.coDeepSeek-R1-Distill-Qwen-32B is an open-weight language model distilled from DeepSeek-R1 and based on Qwen. Developers can run it locally or self-host it for reasoning, text generation, and code-related workloads.
- DeepSeek R1 0528 Qwen3 8Bhuggingface.coDeepSeek-R1-0528-Qwen3-8B is a large open-source language model designed for advanced text generation and understanding. It is suitable for developers and researchers seeking a robust model for NLP applications, with support for local deployment and Docker integration.
- DeepSeek R1 NextNhuggingface.coDeepSeek-R1-NextN is a model repository hosted on Hugging Face under the lmsys organization. It provides a tokenizer configuration that defines specific tokens for conversation handling, including a beginning-of-sentence token, an end-of-sentence token used also for padding, and a custom chat template written in Jinja2 syntax. The template processes lists of messages with roles such as system, and conditionally builds a system prompt while managing generation prompts. The repository page supplies the exact token definitions and template logic required to format inputs for the underlying model. This supports consistent encoding and decoding when the model is loaded through the Hugging Face ecosystem. The page itself forms part of the broader Hugging Face platform, which hosts models, datasets, and related resources for open-source machine learning work. No additional capabilities, training details, performance metrics, licensing terms, or intended user groups are stated on the page. The content focuses on the structural elements of the tokenizer and chat template rather than describing downstream applications or deployment methods beyond the hosting location.
- DeepSeek R1 Distill Qwen 14Bhuggingface.coDeepSeek-R1-Distill-Qwen-14B is a 14-billion parameter language model created by distilling reasoning capabilities from the DeepSeek-R1 series into the Qwen architecture. It is provided as open weights on Hugging Face for local or hosted inference using standard Transformers pipelines. The model supports chat-based interactions and is intended for developers building reasoning, coding, or agentic applications.
- DeepSeek V2 Litehuggingface.coDeepSeek-V2-Lite is a smaller, more efficient variant of the DeepSeek-V2 large language model. It uses a mixture-of-experts architecture to deliver strong performance while requiring fewer resources for inference. The model supports chat interactions and is available on Hugging Face for developers to integrate into applications.
- DeepSeekhuggingface.coDeepSeek-V3 is an open-weight large language model distributed through Hugging Face for local or hosted inference. Developers can download, run, and integrate the model using Python, Docker, or compatible inference providers.
- DeepSeek V2 Lite Chathuggingface.coDeepSeek-V2-Lite-Chat is an open-weight language model for conversational text generation. Developers can download and run it locally using compatible Python tooling or Docker-based deployments.
- Deepseek r1 0528 APIhuggingface.coThis Hugging Face Space provides a simple chat interface to the DeepSeek V3 model. Users can ask questions and optionally enable "Deep Research" which extracts key terms, searches the web, and incorporates results into the response. It is intended for users wanting quick access to a capable reasoning model.
- DeepSeek V4 Flash 0731huggingface.coDeepSeek-V4-Flash-0731 is an open-weight large language model published by deepseek-ai on Hugging Face. It provides high-performance text generation capabilities that can be downloaded, fine-tuned, or self-hosted by developers. The model includes tokenizer configurations and supports both local inference and cloud deployment, making advanced AI accessible without vendor lock-in.
- DeepSeek V4 Prohuggingface.coDeepSeek-V4-Pro is an open-source large language model released by deepseek-ai. It supports text generation, fine-tuning, and inference via API or CLI, making it suitable for AI research, development, and custom NLP applications.
- DeepSeekhuggingface.coDeepSeek-V3.2 is an open-source large language model designed for local inference and text generation. It provides AI researchers and developers with the ability to run advanced language tasks on their own infrastructure, supporting customization and privacy.
- DeepSeek V3 0324huggingface.coDeepSeek-V3-0324 is an open-weight large language model distributed through Hugging Face for local inference and developer integration. It supports text and code generation, conversational prompting, and tool-oriented workflows.
- DeepSeek R1 Distill Llama 70Bhuggingface.codeepseek-ai/DeepSeek-R1-Distill-Llama-70B is a distilled version of the DeepSeek-R1 reasoning model based on Llama architecture. It offers strong performance on complex reasoning tasks while being smaller than the original. The model is distributed as open weights on Hugging Face and can be run locally or served via inference frameworks.
- DeepSeek V3.2 Exphuggingface.coDeepSeek-V3.2-Exp is an experimental release in the DeepSeek series of large language models. It features an advanced tokenizer and chat template designed for complex reasoning and instruction following. The model is hosted on Hugging Face and intended for local or hosted inference.
- DeepSeek R1huggingface.coDeepSeek-R1-MXFP4 is a quantized version of the DeepSeek-R1 reasoning model, optimized for AMD hardware. It supports efficient inference with reduced memory requirements while maintaining strong performance on reasoning, coding, and general language tasks. Hosted on Hugging Face, it provides model weights, chat templates, and integration instructions for developers building local AI applications.
- DeepSeek R1 0528 NVFP4huggingface.coAn FP4 quantized and optimized version of the DeepSeek-R1 model provided by NVIDIA. It is designed for efficient inference on NVIDIA GPUs while maintaining reasoning capabilities. The model uses a custom chat template and is suitable for developers building AI applications that require strong reasoning performance with lower memory footprint.
- DeepSeek R1 Distill Qwen 32Bhuggingface.coDeepSeek-R1-Distill-Qwen-32B-GGUF is an open-weight reasoning model distributed in GGUF format for local inference. Developers can download quantized variants and run the model with compatible tools such as llama.cpp or Ollama.
- DeepSeek V4 Flash Basehuggingface.coDeepSeek-V4-Flash-Base is the base version of DeepSeek's V4 large language model series, optimized for speed and efficiency. It is designed for text generation and other language tasks and can be run locally or integrated via the Transformers library. The model is part of the open-weight DeepSeek model family.
- DeepSeek R1 Distill Qwen 1.5Bhuggingface.coDeepSeek-R1-Distill-Qwen-1.5B is a distilled version of the DeepSeek-R1 reasoning model based on the Qwen architecture. It offers strong reasoning performance in a much smaller 1.5 billion parameter footprint, making it suitable for local deployment and fine-tuning.
- DeepSeek R1visualstudio.comA Visual Studio Code extension that provides DeepSeek R1-powered coding assistance. It supports code completion, explanation, review, refactoring, testing, chat, multiple programming languages, and customizable prompts for developers.
- DeepSeek OCRhuggingface.codeepseek-ai/DeepSeek-OCR is an open-source model for optical character recognition, enabling extraction of text from images. It supports integration with PyTorch and Docker, making it suitable for document digitization and data extraction workflows. Designed for data scientists and developers working with scanned documents and images.
- DeepSeek V4 Pro 0813huggingface.coDeepSeek-V4-Pro-0813 is an open-weight AI model distributed through Hugging Face for local and hosted inference. Developers can run it with Python, Docker, or compatible inference providers for text, code, and multimodal workloads.
- DeepSeek V4 Flashhuggingface.coDeepSeek V4 Flash is an open-source large language model optimized for fast and efficient text generation. It is suitable for developers and researchers building AI-powered applications that require high-speed inference and supports integration via API and CLI.
- Deepseek Vl2 Tinyhuggingface.coDeepSeek-VL2 Tiny is an open multimodal vision-language model that processes images and text for image understanding, visual question answering, and image-grounded conversations. Developers can run it locally using Transformers and its published model weights.
- DeepSeek OCR 2huggingface.coDeepSeek-OCR-2 is an open-source AI model for optical character recognition, enabling accurate extraction of text from images. It is designed for integration into applications requiring OCR, supporting high accuracy and open deployment.
- DeepSeek R1 Distill Llama 8Bhuggingface.coThis is a distilled 8B parameter model from the DeepSeek-R1 series, based on the Llama architecture. It retains strong reasoning and problem-solving abilities from the larger teacher model while being significantly more efficient. The model is suitable for local inference, coding assistance, and complex reasoning tasks.
- DeepSeek R1 Distill Qwen 7Bhuggingface.coDeepSeek-R1-Distill-Qwen-7B is a distilled version of the DeepSeek-R1 reasoning model based on the Qwen architecture. It supports advanced chat templates, tool calling, and instruction following while being significantly smaller than the full model.
- Deepseek ai.DeepSeekhuggingface.coDeepSeek-V3.2-GGUF is a downloadable GGUF-format quantization of the DeepSeek-V3.2 language model for local inference. It provides multiple quantization levels and sharded files for running text and code generation with compatible tools.
- DeepSeek Toolboxdeepseektoolbox.comDeepSeek Toolbox is a browser extension suite designed to improve the DeepSeek AI experience. It offers features like one-click copying of AI-generated content and auto-folding for cleaner chat interfaces. Ideal for users seeking to streamline their AI workflows.
- DeepSeek R1 0528 Qwen3 8Bhuggingface.coThis model is an 8-bit quantized version of DeepSeek-R1-0528 based on the Qwen3 8B architecture, optimized for the MLX framework used on Apple Silicon. It includes a custom chat template and is distributed by the lmstudio-community on Hugging Face for local inference.
- DeepSeek OCRdeepseekocr.appDeepSeek OCR is a free online tool that uses a vision-language AI model to extract text from images and documents with high accuracy. Users can convert documents to Markdown, parse charts, and process complex layouts. It is designed for professionals, researchers, and anyone needing fast, reliable OCR capabilities.
- DeepSeek - AI Assistantgoogle.comDeepSeek - AI Assistant is an Android app offering an intelligent AI chatbot for personal assistance, information retrieval, and productivity. It is designed for users seeking conversational AI support on their mobile devices.
- DeepSeek-Prover-V2-671Bhuggingface.coThis Hugging Face Space offers a straightforward chat UI for the DeepSeek-Prover V2 671B model. After signing in, users submit prompts or questions and receive generated responses. It is primarily intended for users exploring advanced mathematical reasoning and proof generation capabilities.
- DeepSeek V4 Flash 162B REAPhuggingface.coDeepSeek-V4-Flash-162B-REAP-GGUF is an open-source large language model available on Hugging Face, designed for advanced text generation and inference. It supports GGUF format, multi-GPU setups, and quantized weights, making it suitable for developers and researchers working on natural language processing projects in English and Chinese.
Ranked by how close each one sits to DeepSeek R1 in the index, not by popularity. Back to DeepSeek R1 →