A port of OpenAI's CLIP ViT-B/32 model by Xenova, optimized for execution in JavaScript environments via ONNX Runtime or WebNN. It enables zero-shot image classification, image-text similarity, and multimodal embeddings directly in the browser or on the server with Node.js. The model is widely used for client-side computer vision tasks.
Clip Vit Base Patch32 is an Other AI project. It focuses on running CLIP vision-language models in browser or Node.js environments without heavy Python dependencies. It is built as an open-source project for web developers. Clip Vit Base Patch32 is open source under the Open Source license. It ships for the web, the command line, and API.
It is developed by Xenova (United States), and it first shipped in 2023. Among its 3 catalogued features are Image-Text Embedding, ONNX Runtime, and Vision Transformer.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do