Qwen2.5-VL-72B-Instruct is a large multimodal model capable of understanding both images and video alongside text. The AWQ quantized version allows efficient deployment. It supports advanced vision-language tasks and follows a chat-based instruction format, making it suitable for complex multimodal applications.
Qwen2.5 VL 72B Instruct sits in PulseGate's Foundation models & chat category. It focuses on enabling high-performance multimodal understanding of images, video, and text in a single open model. Qwen2.5 VL 72B Instruct is an open-source project aimed at Multimodal AI developers. The project is open source (Apache-2.0). The product ships for the web, the command line, and API.
Qwen builds and maintains Qwen2.5 VL 72B Instruct, and the product first shipped in 2024. The project is developed in the open on GitHub with 19.6k stars. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 4 catalogued features are Vision Understanding, Video Understanding, and Instruction Following.
Latest indexed changes and source events
Qwen/Qwen2.5-VL-72B-Instruct-AWQ verified by the PulseGate indexer