Xenova/clip-vit-base-patch16 is a model repository on Hugging Face that hosts a version of the CLIP ViT-B/16 architecture. It is provided for use in JavaScript-based machine learning environments through compatibility with Transformers.js and ONNX.js. The repository was created on 19 May 2023 and last modified on 8 October 2024. It has recorded over 868000 downloads to date.
The entry contains configuration details for a tokenizer that includes specific added tokens for beginning-of-sequence, end-of-sequence, padding, and unknown tokens, all set with defined stripping and normalization behaviors. No further capabilities, tasks, or technical specifications are described on the page itself. The model is listed without any assigned inference providers and with discussions enabled under recent sorting.
It forms part of the broader collection of models hosted on the Hugging Face platform, which focuses on open source and open science initiatives in artificial intelligence. The page provides no information on licensing, pricing, target audience, or deployment methods beyond the repository context.
Clip Vit Base Patch16 sits in PulseGate's Other AI category. It focuses on running CLIP vision-language models directly in the browser or Node.js without Python dependencies. Clip Vit Base Patch16 is an open-source project aimed at developers. The project is open source (Open Source). Clip Vit Base Patch16 is available on the web, the command line, and API.
It is developed by Xenova, and the product first shipped in 2023. Among its 3 catalogued features are Image Embeddings, Text Embeddings, and Zero-shot Classification.
Latest indexed changes and source events
Xenova/clip-vit-base-patch16 verified by the PulseGate indexer
Other apps tracked under the same category.