Phi-4-multimodal-instruct-NVFP4 is an NVIDIA-optimized, quantized version of the Phi-4 multimodal model. It supports text, vision, and other modalities under an instruction-tuned framework. Released under a permissive license that allows commercial use and derivative works, it is designed for local inference on compatible hardware.
In the Other AI space, Phi 4 Multimodal Instruct takes a focused approach. It focuses on running efficient multimodal AI models locally with reduced memory and compute requirements. It is built as an open-source project for AI developers and researchers. Phi 4 Multimodal Instruct is open source under the Apache-2.0 license. It runs on the web and API.
It is developed by NVIDIA (United States), and it first shipped in 2024. Development happens publicly on GitHub with 3.3k stars and 350 commits in the last 90 days. Among its 3 catalogued features are multimodal understanding, instruction following, and NVFP4 quantization.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do