This is a GPTQ Int4 quantized version of the Qwen3.6-35B-A3B model optimized for reduced memory usage while maintaining performance. It supports vision inputs, complex chat templates, and multimodal content. The model is hosted on Hugging Face for use with Transformers and inference engines.
In the Foundation models & chat space, Qwen3.6 35B A3B takes a focused approach. It focuses on deploying a large 35B parameter multimodal model efficiently on consumer or edge hardware using 4-bit quantization. It is built as an open-source project for developers needing efficient LLMs. Qwen3.6 35B A3B is open source under the Open Source license. Qwen3.6 35B A3B is available on the web and the command line, and it can be self-hosted.
Palmfuture builds and maintains Qwen3.6 35B A3B, and the product first shipped in 2025. It operates in a well-populated space: PulseGate tracks 10 similar tools. Key capabilities include Quantized Inference, Vision Support, and Advanced Chat Templates. It exposes integrations via a public API.
Latest indexed changes and source events
palmfuture/Qwen3.6-35B-A3B-GPTQ-Int4 verified by the PulseGate indexer
Other apps tracked under the same category.