This is a Vision Transformer (ViT) model trained with the CLIP objective on the LAION-400M dataset. It enables zero-shot image classification and image-text similarity computation. The model is provided through the timm library and OpenCLIP on Hugging Face.
Vit Base Patch32 Clip 224.laion400m E32 sits in PulseGate's Other AI category. It focuses on classifying images into arbitrary categories without task-specific training. It is built as an open-source project for machine learning researchers. The project is open source (Open Source). Vit Base Patch32 Clip 224.laion400m E32 is available on the web and API.
Behind Vit Base Patch32 Clip 224.laion400m E32 is timm. Key capabilities include Zero-Shot Image Classification, CLIP Embeddings, and Vision Transformer.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do