This repository contains a w8a8 quantized version of the Qwen3.5-9B model from RedHatAI. It supports both text and vision inputs and is optimized for lower memory usage and faster inference. The model includes vision tower components and is compatible with the Transformers library.
Qwen3.5 9B Quantized.w8a8 is a Foundation models & chat product. It focuses on running large language and vision models with reduced memory footprint using 8-bit quantization. Qwen3.5 9B Quantized.w8a8 is an open-source project aimed at developers building AI applications. The project is open source (Apache-2.0). The product ships for the web, the command line, and API.
Behind Qwen3.5 9B Quantized.w8a8 is RedHatAI, and the product first shipped in 2019. The project is developed in the open on GitHub with 3.6k stars and 165 commits in the last 90 days. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 4 catalogued features are Quantized Weights, Vision Support, and Chat Template. It exposes integrations via a public API.
Latest indexed changes and source events
RedHatAI/Qwen3.5-9B-quantized.w8a8 verified by the PulseGate indexer
Other apps tracked under the same category.