This is a quantized version of the Qwen2.5-VL-3B vision-language model optimized with FP8 dynamic quantization. It can understand and reason about images, videos, and text together. The model is provided by RedHatAI and is suitable for multimodal AI applications.
In the Multimodal & vision space, Qwen2.5 VL 3B Instruct FP8 Dynamic takes a focused approach. It focuses on processing both visual and textual inputs with a single efficient multimodal model. Qwen2.5 VL 3B Instruct FP8 Dynamic is an open-source project aimed at developers. The project is open source (Apache-2.0). Qwen2.5 VL 3B Instruct FP8 Dynamic is available on the web, the command line, and API.
RedHatAI builds and maintains Qwen2.5 VL 3B Instruct FP8 Dynamic, and it first shipped in 2019. The project is developed in the open on GitHub with 3.6k stars and 165 commits in the last 90 days. PulseGate's similarity index places it among 14 comparable projects. Key capabilities include Vision Language, Image Understanding, and Video Processing. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do