gemma-4-E4B is Google's open multimodal model capable of any-to-any tasks involving text and images. Released in 2026, it supports image-text-to-text pipelines and runs locally via the Transformers library. The model is provided as open weights on Hugging Face for researchers and developers exploring unified multimodal AI capabilities.
In the Foundation models & chat space, Gemma 4 E4B takes a focused approach. It focuses on processing and generating across text and image modalities in a single unified open model. Gemma 4 E4B is an open-source project aimed at developers. The project is open source (Open Source). Gemma 4 E4B is available on the web, the command line, and API.
It is developed by Google, and the product first shipped in 2026. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 3 catalogued features are Image-Text Understanding, any-to-Any, and Multimodal Generation.
Latest indexed changes and source events
google/gemma-4-E4B verified by the PulseGate indexer
Other apps tracked under the same category.