This is a CLIP ViT-B/16 model trained on the DataComp XL dataset (s13B-b90K). It is designed for vision-language tasks including zero-shot image classification and retrieval. The model is available on Hugging Face and uses a standard CLIP architecture for aligning image and text embeddings in a shared space.
CLIP ViT B 16 DataComp.XL s13B b90K sits in PulseGate's Other AI category. It focuses on enabling zero-shot image classification and vision-language alignment using contrastive pretraining. CLIP ViT B 16 DataComp.XL s13B b90K is an open-source project aimed at developers. The project is open source (Open Source). It runs on the web and API.
flavour builds and maintains CLIP ViT B 16 DataComp.XL s13B b90K, and the product first shipped in 2023. The project is developed in the open on GitHub with 787 stars. Among its 3 catalogued features are Contrastive Learning, Vision Transformer, and Zero-shot Classification.
Latest indexed changes and source events
flavour/CLIP-ViT-B-16-DataComp.XL-s13B-b90K verified by the PulseGate indexer
Other apps tracked under the same category.