Skip to content
Back to the index

Video LLaMA

huggingface.coMultimodal & vision

PulseGate's liveness check found it on 8 Oct 2026; it is registered on Hugging Face Spaces and has been in the index since 10 Jul 2026. How this is checked

Video LLaMA is a web app that allows users to upload videos or images and interact with them by asking questions. The AI provides detailed explanations about the content, making it useful for educators, researchers, and analysts.

Inferred · not functionally tested

FreeWebSelf-hostedCloud-managed
Video LLaMA preview
Visit huggingface.co

Overview

4 features

Purpose: Understanding and extracting information from videos or images through conversational AI.

Inferred · not functionally tested

Audience: educators, researchers, content analysts

Inferred · not functionally tested

Functions: Unknown

Interfaces: API: unknown · MCP: unknown · CLI: unknown · Self-hosting: indicated (inferred, not tested)

Recorded constraints: pricing: free · license: Proprietary · platforms: WEB · deployment: browser, self_hosted, cloud_managed

Constraint provenance is unknown; confirm requirements with the publisher.

Record sources: huggingface.co. These links do not verify the individual claims.

Video LLaMA is a Multimodal & vision project. Inferred · not functionally tested: It focuses on understanding and extracting information from videos or images through conversational AI. Inferred · not functionally tested: It is built as a B2B2C product for educators, researchers, content analysts. Basis unknown · not verified: It is available for free. Basis unknown · not verified: It runs on the web, and it can be self-hosted.

It is developed by DAMO-NLP-SG, and it first shipped in 2024. Inferred · not functionally tested: Among its 4 catalogued features are Video Q&A, Image Q&A, and content analysis.

Summary written by a language model from the project’s public pages.

Tasks: Inferred · not functionally tested

  • Video Q&A
  • Image Q&A
  • Content analysis
  • Detailed explanations

Topics: Inferred · not functionally tested

Tags
video-analysisvisual-qacontent-explanationai-conversationmedia-understanding
AI capabilities
VideoImageText
Inference: Cloud API

JSON profile · Text profile · Access guide

Built with & integrations

Hosting
AWS
AI providers
local_oss
Runs on
BrowserSelf-hostedCloud-managed
Detected from
AWS
x-amz-cf-id header · x-amz-cf-pop header · via header
local_oss
huggingface in the HTML

Trust & compliance

Public signals
HTTPSFree tier

Indexing history

1

What PulseGate has recorded for this listing

  1. Indexed10 Jul · 19:16 UTC
    DAMO-NLP-SG/Video-LLaMA seen via Hugging Face Enumerator
    Source: Hugging Face Enumerator · Open

Frequently asked questions about Video LLaMA

What does Video LLaMA do?
Inferred · not functionally tested: Video LLaMA focuses on understanding and extracting information from videos or images through conversational AI. It is catalogued under Multimodal & vision on PulseGate.
Who should use Video LLaMA?
Inferred · not functionally tested: Video LLaMA is a B2B2C product built for educators, researchers, content analysts.
Is Video LLaMA free?
Basis unknown · not verified: Yes — Video LLaMA is free to use.
What platforms does Video LLaMA run on?
Basis unknown · not verified: Video LLaMA runs on the web. It can also be self-hosted.
Is Video LLaMA still active?
PulseGate's liveness check found it on 8 Oct 2026.
What are alternatives to Video LLaMA?
Similar projects tracked by PulseGate include VideoLLaMA3, VideoLLaMA2, and Video LLaVA.VideoLLaMA3VideoLLaMA2Video LLaVA
Who makes Video LLaMA?
Video LLaMA is developed by DAMO-NLP-SG.
When did Video LLaMA launch?
Video LLaMA first shipped in 2024.

Similar projects

Closest matches by what these projects do