Voxtral-Small-24B-2507 is a multimodal audio-text model developed by Mistral AI. It belongs to the Voxtral family and supports direct audio input and output in addition to text. The model is suitable for voice-based AI applications, speech recognition, and conversational agents that work with spoken language. It is distributed on Hugging Face and supports vLLM inference.
Voxtral Small 24B 2507 sits in PulseGate's Other AI category. It focuses on building applications that require direct understanding and generation of spoken audio alongside text. Voxtral Small 24B 2507 is an open-source project aimed at developers. The project is open source (Apache-2.0). Voxtral Small 24B 2507 is available on the web, the command line, and API.
Mistral AI builds and maintains Voxtral Small 24B 2507, and the product first shipped in 2023. The project is developed in the open on GitHub with 86.8k stars and 3k commits in the last 90 days. Among its 3 catalogued features are Audio Understanding, Speech Generation, and Multimodal Chat.
Latest indexed changes and source events
mistralai/Voxtral-Small-24B-2507 verified by the PulseGate indexer
Other apps tracked under the same category.