This model is a large ConvNeXt-based CLIP variant trained on the LAION-2B dataset with 29 billion samples and further fine-tuned. It maps images and text into a shared embedding space for tasks such as zero-shot classification, retrieval, and similarity computation. It is distributed on Hugging Face for use by machine learning developers building multimodal AI applications.
CLIP Convnext Large D 320.laion2B s29B b131K Ft Soup is an Other AI product. Lack of strong open-source CLIP models for multimodal image-text understanding. It is built as an open-source project for developers. CLIP Convnext Large D 320.laion2B s29B b131K Ft Soup is open source under the Open Source license. The product ships for the web and API.
Behind CLIP Convnext Large D 320.laion2B s29B b131K Ft Soup is LAION, and the product first shipped in 2021. Development happens publicly on GitHub with 14k stars and 110 commits in the last 90 days.
Latest indexed changes and source events
laion/CLIP-convnext_large_d_320.laion2B-s29B-b131K-ft-soup verified by the PulseGate indexer
Other apps tracked under the same category.