This is a large ViT-bigG/14 CLIP model trained by LAION on the LAION-2B dataset. It produces powerful joint embeddings for images and text, enabling zero-shot image classification, retrieval, and other multimodal tasks. The model is widely used in the open-source AI community and available on Hugging Face.
In the Multimodal & vision space, CLIP ViT bigG 14 laion2B 39B B160k takes a focused approach. It focuses on providing high-quality open-weight CLIP embeddings for image-text similarity and zero-shot classification tasks. It is built as an open-source project for developers. The project is open source (Open Source). CLIP ViT bigG 14 laion2B 39B B160k is available on the web, API, and the command line.
LAION builds and maintains CLIP ViT bigG 14 laion2B 39B B160k, and it first shipped in 2021. The project is developed in the open on GitHub with 14k stars and 110 commits in the last 90 days. Key capabilities include Contrastive Learning, Image-Text Alignment, and Zero-Shot Classification.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do