This model is a ConvNeXt Base image encoder trained with CLIP on the LAION-2B dataset and provided through the timm library. It generates rich feature embeddings suitable for image-text similarity, zero-shot classification, and retrieval. It integrates with both timm and Hugging Face Transformers for easy use in computer vision pipelines.
Convnext Base.clip Laion2b sits in PulseGate's Multimodal & vision category. It focuses on extracting high-quality visual embeddings for image understanding and retrieval tasks. It is built as an open-source project for developers. Convnext Base.clip Laion2b is open source under the Open Source license. Convnext Base.clip Laion2b is available on the web and API.
It is developed by timm. Among its 3 catalogued features are Image Feature Extraction, CLIP Encoder, and Zero-shot Classification.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do