Skip to content
Alternatives
Software like Qwen3.8-Max
What else does this job. Matched on what each project does, not on who links to whom.
Closest first
- QwenCloudqwencloud.comQwenCloud Token Plan is a subscription plan for access to QwenCloud’s AI models and related tooling. The page presents it as a way to use multiple models under one plan and to start building with AI, with an emphasis on upgraded individual access and a more affordable team option. The plan offers unified access to leading text, vision, speech, and image generation models. The page also highlights higher quota and better price, and it names model examples including glm-5.2, deepseek-v4-pro, wan2.7-image-pro, and Qwen3.8-Max-Preview. In addition, the page refers to a seamless AI toolchain integration and shows an example of using Qwen Cloud with Qwen Code by installing Qwen Code, opening Settings, choosing the provider section, selecting Qwen Cloud Coding Plan, signing in with a Qwen Cloud account, confirming, and saving the settings before coding with Qwen Cloud in Qwen Code. The audience is described in general terms as individual and team users. The pricing model is a token plan, with the page stating that an upgraded individual option is available and that Team is more affordable. The page is titled Token Plan and invites subscription to QwenCloud Token Plan.
- Qwen1.5 MoE A2.7Bhuggingface.coQwen1.5-MoE-A2.7B is a Mixture-of-Experts (MoE) model from Alibaba's Qwen series. Despite having more total parameters, it activates only 2.7 billion parameters per token, offering a strong balance between performance and efficiency. It uses the ChatML format and is suitable for text generation, dialogue, and further fine-tuning.
- Qwen3 1.7Bhuggingface.coQwen3-1.7B is a foundation model hosted on Hugging Face. The model page provides a chat template that defines how the model processes conversation history, system prompts, and tool calls. This template supports a multi-step tool-use format in which the model receives function signatures inside XML-style tags and must respond with structured JSON objects wrapped in tool_call tags when invoking external functions. The template includes conditional logic for handling an initial system message and for formatting tool definitions as JSON objects. It also specifies an instruction prefix that tells the model it may call one or more functions to assist with a user query. The page itself carries the standard Hugging Face interface elements for models, including tabs for files, discussions, and community features. Qwen3-1.7B belongs to the class of foundation models. It is delivered as a downloadable model repository on the Hugging Face platform. No pricing, licensing terms, or target user roles are stated on the page.
- Qwen3 0.6Bhuggingface.coQwen/Qwen3-0.6B is an open-source large language model designed for advanced text generation and understanding tasks. It is transformer-based, supports Python integration, and is suitable for NLP developers and researchers seeking open weights for customization.
- Qwen2.5 3Bhuggingface.coQwen2.5-3B is part of Alibaba's Qwen2.5 family of open foundation models. The 3B variant offers a balance of performance and efficiency, supporting chat, reasoning, coding, and tool-calling use cases. It includes an advanced chat template with tool integration and is distributed as open weights on Hugging Face for flexible deployment.
- Qwen 3.6 35B A3B VRAP 4 Bit AWQ 21.2GBhuggingface.coA heavily quantized (4-bit AWQ) version of a Qwen 3.6 model with 35B+3B parameters and vision capabilities (VRAP). The 21.2GB model supports both text and image inputs. It is designed for local inference using GGUF-compatible tools or optimized runtimes.
- Qwen3.6 35B A3Bhuggingface.coQwen3.6-35B-A3B-MLX-8bit is a community-quantized version of Alibaba's Qwen model using 8-bit precision and optimized for Apple's MLX framework. It supports text and vision inputs and can run locally on compatible hardware. The model is provided on Hugging Face for developers building local AI applications or experimenting with large language models without heavy cloud dependency.
- Qwen3.6 35B A3Bhuggingface.coQwen3.6-35B-A3B-NVFP4 is a foundation model hosted on Hugging Face under the nvidia organization. The model processes multimodal inputs that include text, images, and video through a specialized chat template. Its template defines distinct handling for each content type, inserting vision-specific tokens such as vision_start, image_pad, video_pad, and vision_end while counting vision elements and raising exceptions for unsupported cases like videos in system messages. The template also supports tool calling by generating a system prompt that lists available functions when tools are supplied. It iterates over conversation messages, applies conditional formatting based on content type, and enforces rules such as requiring at least one message. The implementation appears as Jinja2-style macros that output structured strings compatible with the model's expected input format. This model is listed alongside standard Hugging Face infrastructure for models, datasets, spaces, and community resources. The associated organization page promotes open source and open science initiatives in artificial intelligence.
- Qwen3.5 0.8Bhuggingface.coQwen3.5-0.8B is a compact, open-source large language model for text generation and conversational AI. It is suitable for developers building chatbots, virtual assistants, and other NLP applications, with support for CLI, API, and Docker deployment.
- Qwen3.5 9Bhuggingface.coQwen3.5-9B-NVFP4 is a quantized (NVFP4) version of Alibaba's Qwen3.5-9B large language model. It is optimized for reduced memory usage and faster inference while maintaining strong performance. The model supports multimodal inputs and can be run locally using standard Hugging Face tools.
- Qwen3 32Bhuggingface.coQwen3-32B-AWQ is an open-source large language model designed for advanced text generation and tool calling. It is suitable for AI developers and researchers seeking a powerful LLM for integration into applications or research workflows.
- Qwen3 1.7Bhuggingface.coQwen3-1.7B-FP8 is a quantized variant of Alibaba's Qwen3 series of large language models. It supports advanced features including tool calling and follows a chat template optimized for instruction following. The FP8 format enables faster inference with reduced memory requirements while maintaining strong performance.
- Qwen3.5 122B A10Bhuggingface.coThis is a GPTQ-Int4 quantized variant of Alibaba's Qwen3.5-122B-A10B model. It supports multimodal inputs (including vision and video) and advanced features such as tool calling. The model is distributed on Hugging Face and is intended for developers who want to run a high-performance open model locally or on modest GPU hardware.
- Qwen3.5 9Bhuggingface.coQwen3.5-9B is an open-source large language model from the Qwen family, designed for text generation and conversational AI. It is intended for developers and researchers building advanced NLP and AI applications.
- Qwen2.5 14Bhuggingface.coQwen2.5-14B is part of the Qwen2.5 series of large language models from Alibaba's Qwen team. It features strong performance on reasoning, coding, and multilingual tasks. The model includes a chat template supporting tool/function calling and is provided with full weights on Hugging Face for local inference or fine-tuning.
- Qwen3 0.6Bhuggingface.coQwen3-0.6B-FP8 is an open-source, compact language model for text generation and tool calling. It is suitable for developers and researchers seeking a lightweight LLM for integration and experimentation.
- Qwen3.6 35B A3Bhuggingface.coQwen3.6-35B-A3B-MLX-6bit is a 6-bit quantized version of the Qwen 3.6 model hosted by the lmstudio-community on Hugging Face. It belongs to the class of foundation models and is distributed as a repository containing model weights, tokenizer configuration, and a chat template that supports multimodal inputs. The repository includes a tokenizer definition with specific pad, end-of-text, and unknown tokens. Its chat template contains logic for processing mixed content, including separate handling for text strings and iterable content. The template detects image and video items, increments internal counters for each, prepends numbered labels such as "Picture 1:" when a flag is set, and inserts special vision markers. It explicitly prevents images or videos from appearing in system messages by raising an exception. The model is made available through the Hugging Face platform, which provides access to models, datasets, and related resources. No information is given on licensing, pricing, intended user roles, or specific hardware optimizations beyond what is encoded in the repository name and template.
- Qwen3.6 35B A3Bhuggingface.coQwen3.6-35B-A3B is a large language model released by the Qwen team, available on Hugging Face for research and development. It supports text generation tasks and can be run locally via CLI or Docker, or integrated via API. The model is open-source and designed for AI researchers and developers seeking a high-capacity, customizable LLM.
- Qwen3 Next Moehuggingface.coThis is a minimal 'tiny-random' test model for the Qwen3-Next MoE (Mixture of Experts) architecture. It includes support for tool calling via a specialized chat template and is intended for developers testing integration with the Qwen3 model family and its function-calling features.
- Qwen3.6 35B A3B NVFP4 Fasthuggingface.coThis is a highly optimized, quantized version of the Qwen 3.6 35B model using NVFP4 precision for accelerated inference. It supports both text and vision inputs while maintaining strong performance. Distributed via Hugging Face, it is designed for developers needing fast multimodal inference on consumer or enterprise GPUs with Unsloth and Transformers compatibility.
- Qwen3.5 122B A10Bhuggingface.coQwen3.5-122B-A10B is a 122 billion parameter multimodal model from Alibaba's Qwen series. It processes text, images, and video inputs within a unified architecture and supports advanced capabilities such as tool calling. The model is distributed on Hugging Face and can be used through inference providers, pip packages, and Docker containers for research and application development.
- Qwen3.5 122B A10Bhuggingface.coQwen3.5-122B-A10B-NVFP4 is a highly quantized version of the large Qwen 3.5 model (122B parameters). It supports vision inputs alongside text and includes advanced tool-calling features. The model is distributed on Hugging Face for use in high-performance inference environments.
- Qwen3 30B A3B Thinking 2507huggingface.coThis repository hosts a quantized FP8 version of the Qwen3 30B-A3B Thinking model. It is designed for complex reasoning tasks, multi-step tool use, and agentic workflows. The model can be used locally with Docker or via pip with the Transformers library.
- Qwen3.6 35B A3Bhuggingface.coQwen/Qwen3.6-35B-A3B-FP8 is an open-source large language model designed for advanced text generation and understanding. It supports instruction following and multilingual capabilities, making it suitable for developers and researchers building AI-powered solutions.
- Qwen3.6 35B A3B MLX VQ 3.4bpwhuggingface.coQwen3.6-35B-A3B-MLX-VQ-3.4bpw is an open-source large language model available on Hugging Face, designed for natural language processing tasks. It supports local inference and can be integrated via API or CLI for research and development purposes. The model is suitable for AI researchers and developers seeking customizable LLMs.
- Qwen3.5 35B A3Bhuggingface.coQwen3.5-35B-A3B is an open-source large language model supporting both text and multimodal inputs. It is designed for advanced AI applications, including chatbots and multimodal assistants, and is suitable for developers and researchers in AI.
- Qwen3.6 35B A3Bhuggingface.coQwen3.6-35B-A3B is an open-source large language model checkpoint available on Hugging Face, designed for advanced natural language processing tasks. It enables AI researchers and developers to run, fine-tune, or integrate a state-of-the-art LLM in their own environments. The model supports local inference and experimentation for a wide range of text generation applications.
- Qwen3.6 35B A3B PrismaQuant 4.75bit Vllmhuggingface.coThis is a 4.75-bit PrismaQuant version of a Qwen3.6 35B-A3B model optimized for use with the vLLM inference engine. It supports multimodal inputs including vision and includes custom chat templates for tool use. The model is distributed on Hugging Face for efficient local or server-based deployment.
- Qwen3 8Bhuggingface.coQwen3-8B is an open-source large language model designed for text generation and conversational AI tasks. It supports fine-tuning and can be deployed locally or via API, making it suitable for machine learning engineers building advanced NLP applications.
- Qwen3.5 122B A10Bhuggingface.coThis is a quantized (NVFP4) variant of the Qwen3.5-122B model optimized by NVIDIA for efficient inference. It supports advanced features including tool calling, multimodal inputs, and is designed to run on NVIDIA GPUs with significantly lower memory requirements than the original model. The model is distributed on Hugging Face and can be used with standard Transformers pipelines or custom inference servers.
- Qwen3 1.7Bhuggingface.coThis is an optimized version of the Qwen3-1.7B language model provided by Unsloth. It includes support for advanced chat templates, tool calling, and function calling capabilities. The model is designed for efficient training and inference, allowing developers to fine-tune large models on consumer hardware with significantly reduced memory requirements compared to standard implementations.
- Qwen3 32Bhuggingface.coQwen/Qwen3-32B is an open-source large language model designed for advanced natural language understanding and generation tasks. It supports instruction following and multilingual capabilities, making it suitable for developers and researchers building AI-powered applications.
- Qwen3.5 27Bhuggingface.coQwen3.5-27B is a 27B parameter open-source large language model for advanced text generation and conversational AI. It is designed for AI researchers and developers building sophisticated NLP applications, with support for CLI, API, and Docker deployment.
- Qwen3 8B Basehuggingface.coQwen3-8B-Base is the foundational 8B parameter model of the Qwen3 family. It features an enhanced architecture supporting long contexts, multilingual performance, and native tool-calling abilities. It is designed as a base for further fine-tuning and is compatible with the Transformers library.
- Qwen3.6 35B A3Bhuggingface.coThis is a community-quantized 8-bit version of Qwen3.6-35B (with A3B MoE architecture) prepared for the MLX framework on Apple devices. It enables high-performance local inference on Macs with reduced memory footprint while retaining strong reasoning capabilities. The model uses standard MLX conversion and loading patterns.
- Qwen3.5 9B Basehuggingface.coQwen3.5-9B-Base is a 9B parameter foundational model supporting both textual and visual (image/video) inputs. It features an advanced tokenizer, long-context handling, and native support for tool calling. The model is intended as a strong starting point for fine-tuning multimodal and agentic applications.
- Qwen3 4Bhuggingface.coA 4 billion parameter version of the Qwen3 large language model in FP8 precision. It supports instruction following, tool calling, and general text generation. The model is hosted on Hugging Face and can be used locally or via inference providers.
- Qwen AgentWorld 35B A3Bhuggingface.coThis is Alibaba's Qwen-AgentWorld 35B (A3B) model, a large multimodal foundation model designed for agentic workflows. It supports text, image, video, and audio inputs along with advanced tool-calling capabilities. The weights are hosted on Hugging Face for developers building autonomous AI agents.
- Qwen3.5 397B A17Bhuggingface.coQwen3.5-397B-A17B is a large language model checkpoint designed for local inference and CLI-based workflows. It enables developers and researchers to run advanced language models on their own hardware for experimentation and application development.
Ranked by how close each one sits to Qwen3.8-Max in the index, not by popularity. Back to Qwen3.8-Max →