PulseGateLive intelligence on AI-era software
Coverage
180,820

Software tracked

 

Freshness
38 min ago

Last update

 

Cadence
511/day

7-day average

Indexed today: 237

PulseGate

Live intelligence on the software shipping in the AI era — apps, models, agents, and infra.

Software is shipping faster than ever, and a growing share of it lives outside the official app stores. PulseGate tracks it live — free, for builders, analysts, and everyone keeping up.

Follow
GitHubX (Twitter)LinkedIn

Platform

  • All Apps
  • Full Index
  • Categories
  • Industry Updates
  • Data Sources
  • Coverage Rules
  • Glossary
  • Embed Widget

Support

  • Help Center
  • Suggest a URL
  • Report an Issue

Company

  • About
  • Dimaxia
  • Press & Data
  • Contact
  • Platform Status

Legal

  • Privacy
  • Terms
  • Disclaimer

© 2026 PulseGate. Operated by Dimaxia, a brand of Dymaxio s.r.o., Prague, Czech Republic.·

All systems operational
PulseGate
MV
Mage VL
Visit ↗
  1. Home/
  2. Foundation models & chat/
  3. Mage VL
←Back to results
MV

Mage VL

huggingface.co·Infrastructure·🇺🇸

Mage-VL is a multimodal foundation model hosted on Hugging Face. It accepts mixed visual inputs of images and video alongside text.

The model employs a chat template that counts image and video occurrences in a conversation. For each visual element it inserts dedicated vision tokens such as vision_start, image_pad or video_pad, and vision_end. A system prompt is added automatically when the first message is not from the system role. The template supports both string content and structured content lists that distinguish text from image or video objects.

It is delivered as a model repository under the microsoft organization on the Hugging Face platform. The repository supplies tokenizer configuration that defines special tokens including an image pad token, a video pad token, a vision start token, a vision end token, and an end-of-text token. Researchers and developers can download the model files and use the provided chat template for multimodal conversations.

No pricing, licensing terms, or additional capabilities are stated in the repository page.

Open SourceMIT
WebCLIAPI
M
Mage VL preview
Visit huggingface.co↗
⭐673
stars
🍴52
forks
✓5
features
📅2026
since

Overview

5 features

Mage VL is a Foundation models & chat product. It focuses on understanding and reasoning over both images and video content using a single unified model. Mage VL is an open-source project aimed at AI researchers and developers. The project is open source (MIT). It runs on the web, the command line, and API.

Behind Mage VL is Microsoft, based in the United States, and the product first shipped in 2026. The project is developed in the open on GitHub with 673 stars and 32 commits in the last 90 days. Among its 5 catalogued features are multimodal understanding, image analysis, and video analysis.

  • ✓Multimodal understanding
  • ✓Image analysis
  • ✓Video analysis
  • ✓Chat template
  • ✓Vision tokens

Tags

multimodal-modelvision-languagevideo-understandingmicrosoft-research

AI capabilities

MultimodalImageVideoInference: Cloud APIWeights: Open

Built with & integrations

Hosting
aws
AI providers
openailocal_oss
Runs on
BrowserCLIAPI-only

Trust & compliance

LicenseMIT
Verified signals
✓ HTTPS✓ Privacy Policy✓ Terms of Service✓ Open Source✓ Free tier✓ GitHub · ★ 673✓ Active maintenance
Legal
Privacy Policy →Terms of Service →

Recent events

Latest indexed changes and source events

  1. IndexedJul 27, 4:14 AM

    Multimodal foundation model for image and video understanding from Microsoft verified by the PulseGate indexer

    Source: PulseGate indexerOpen ↗

Frequently asked questions about Mage VL

What is Mage VL?
Mage VL focuses on understanding and reasoning over both images and video content using a single unified model. It is catalogued under Foundation models & chat on PulseGate.
Who should use Mage VL?
Mage VL is an open-source project built for AI researchers and developers.
Does Mage VL have a free plan?
Yes — Mage VL is open source under the MIT license and free to use.
What platforms does Mage VL run on?
Mage VL runs on the web, the command line, and API.
Is Mage VL still active?
PulseGate's liveness checks currently classify Mage VL as active. The GitHub repository shows 32 commits in the last 90 days.
What tools are similar to Mage VL?
Similar tools tracked by PulseGate include Multimodal VLM Thinking, InternVL3 1B, and MOVA 360p.Multimodal VLM ThinkingInternVL3 1BMOVA 360p
Who develops Mage VL?
Mage VL is developed by Microsoft, based in the United States.
How long has Mage VL been around?
Mage VL first shipped in 2026.

At a glance

Pricing
Open Source
Platforms
Cli · Web
Languages
English
Open source
Yes · ★ 673
License
MIT
First seen
Jul 27, 2026
Activity
🟢 Active
Status
🟢 Active
Built for
AI researchers and developers
Model
Open source
Solves
Understanding and reasoning over both images and video content using a single unified model.

Developer

🇺🇸Microsoft
Small team
↗ GitHub

Open source

View on GitHub →
⭐ Stars
673
🍴 Forks
52
Open issues
13
Last commit
2d ago
Commits 90d
32
Contributors
3
Authorship
Small team
Default branch
main

Live coverage

Confidence
Medium · 70
Indexed
Jul 27, 2026
Lifecycle
Alive
Activity
Active
First seen
Jul 2026
Last seen
2d ago
Identity audit (10)
Entity ID
cms2ppx4o02i6d161gbnyb147
Slug
multimodal-foundation-model-for-image-and-video-underst-huggingface-co
Lifecycle last checked
Jul 27, 2026
Verification state
Indexed for public listing
Claim / listing state
Unclaimed · listed: yes
Index status
Included in index
Latest evidence snapshot
Jul 27, 2026
Timeline basis
Indexed-at chronology (no inferred launch/funding milestones).
Last updated
Jul 27, 2026
Canonical URL
https://huggingface.co/microsoft/Mage-VL

Similar apps

Other apps tracked under the same category.

  • Multimodal VLM Thinking
    huggingface.co
  • InternVL3 1B
    huggingface.co
  • MOVA 360p
    huggingface.co
  • Videomae Base
    huggingface.co
  • LFM2.5 VL 1.6B
    huggingface.co
  • VLM Object Understanding
    huggingface.co