fancyfeast/llama-joycaption-beta-one-hf-llava is a fine-tuned Llama model using the Llava architecture for high-quality image captioning. It is designed to produce rich, descriptive captions and is distributed in Hugging Face format for easy integration with Transformers and inference tools.
is a Foundation models & chat product. It focuses on generating high-quality, detailed captions for images using an open-source vision-language model. It is built as an open-source project for developers. is open source under the Apache-2.0 license. The product ships for the web, the command line, and API.
It is developed by FancyFeast, and the product first shipped in 2024. Development happens publicly on GitHub with 1.2k stars. Key capabilities include Image Captioning, vision-Language, and Llava Architecture.
Latest indexed changes and source events
fancyfeast/llama-joycaption-beta-one-hf-llava verified by the PulseGate indexer
Other apps tracked under the same category.