LLaVA 1.5 13B is an open-weight vision-language model that processes images and text for tasks such as visual question answering and image understanding. It is distributed through Hugging Face for developers and researchers to run with compatible machine learning tooling.
Llava 1.5 13b Hf sits in PulseGate's Multimodal & vision category. It focuses on understanding and generating text from combined image and language inputs. Llava 1.5 13b Hf is an open-source project aimed at machine learning developers and researchers. Llava 1.5 13b Hf is open source under the BSD-3-Clause license. It ships for the web, the command line, and API, and it can be self-hosted.
LLaVA-HF builds and maintains Llava 1.5 13b Hf, and it first shipped in 2022. Development happens publicly on GitHub with 24.8k stars and 58 commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 5 similar projects. Key capabilities include image understanding, visual question answering, and text generation.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do