Salesforce/blip-vqa-base is a Bootstrapping Language-Image Pre-training (BLIP) model fine-tuned specifically for visual question answering. It accepts an image and a text question and returns a natural language answer. The model is available through the Hugging Face Transformers library with both pipeline and direct model loading options, making it easy to integrate into computer vision applications.
Blip Vqa Base sits in PulseGate's Foundation models & chat category. It focuses on answering natural language questions about the content of images using a unified vision-language model. It is built as an open-source project for AI developers and researchers. Blip Vqa Base is open source under the Open Source license. It runs on the web and API.
It is developed by Salesforce AI Research, and the product first shipped in 2022. Key capabilities include Visual Question Answering, Transformers Integration, and processor & Model.
Latest indexed changes and source events
Salesforce/blip-vqa-base verified by the PulseGate indexer
Other apps tracked under the same category.