This LAION model combines a ViT-B/32 image encoder with an XLM-RoBERTa-base text encoder, trained on the LAION-5B dataset. It is part of the OpenCLIP ecosystem and supports cross-lingual image-text similarity, retrieval, and zero-shot classification. The model is distributed as open weights and integrates directly with the OpenCLIP library.
CLIP ViT B 32 Xlm Roberta Base laion5B s13B B90k sits in PulseGate's Other AI category. It focuses on enabling zero-shot image classification and retrieval across many languages using contrastive vision-language embeddings. CLIP ViT B 32 Xlm Roberta Base laion5B s13B B90k is an open-source project aimed at developers. The project is open source (Open Source). The product ships for the web and API.
LAION builds and maintains CLIP ViT B 32 Xlm Roberta Base laion5B s13B B90k, and the product first shipped in 2021. The project is developed in the open on GitHub with 14k stars and 110 commits in the last 90 days. Among its 3 catalogued features are contrastive image-text model, multilingual text encoder, and openCLIP compatible.
Latest indexed changes and source events
laion/CLIP-ViT-B-32-xlm-roberta-base-laion5B-s13B-b90k verified by the PulseGate indexer
Other apps tracked under the same category.