This model is a large ConvNeXt-based CLIP variant trained on the LAION-2B dataset with 29 billion samples and further fine-tuned. It maps images and text into a shared embedding space for tasks such as zero-shot classification, retrieval, and similarity computation. It is distributed on Hugging Face for use by machine learning developers building multimodal AI applications.
CLIP Convnext Large D 320.laion2B s29B b131K Ft Soup is a Multimodal & vision project. Lack of strong open-source CLIP models for multimodal image-text understanding. It is built as an open-source project for developers. The project is open source (Open Source). It ships for the web and API.
LAION builds and maintains CLIP Convnext Large D 320.laion2B s29B b131K Ft Soup, and it first shipped in 2021. The project is developed in the open on GitHub with 14k stars and 110 commits in the last 90 days.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do