CLIP-ViT-B-16-laion2B-s34B-b88K is a vision transformer model trained with the CLIP objective on the large-scale LAION-2B dataset. It produces aligned embeddings for images and text that support zero-shot classification, retrieval, and similarity tasks. The model weights are publicly available on Hugging Face for research and commercial applications.
CLIP ViT B 16 laion2B s34B b88K sits in PulseGate's Other AI category. It focuses on creating robust open-source multimodal embeddings for image and text without relying on proprietary training data. It is built as an open-source project for AI researchers and developers. CLIP ViT B 16 laion2B s34B b88K is open source under the Open Source license. It runs on the web and API.
LAION builds and maintains CLIP ViT B 16 laion2B s34B b88K, and the product first shipped in 2021. Development happens publicly on GitHub with 14k stars and 115 commits in the last 90 days. Key capabilities include Image-Text Alignment, Zero-Shot Classification, and Contrastive Learning.
Latest indexed changes and source events
laion/CLIP-ViT-B-16-laion2B-s34B-b88K verified by the PulseGate indexer