Whisper Large v3 is an open-source automatic speech recognition model developed by OpenAI. It transcribes spoken audio into text, supports multiple languages, and is robust to noise, making it suitable for developers and researchers building speech-to-text applications.
Whisper Large is a Voice, TTS & speech product. It focuses on automating the transcription of spoken audio into accurate text for various applications. It is built as an open-source project for developers and researchers needing speech-to-text capabilities. Whisper Large is open source under the BSD-3-Clause license. Whisper Large is available on the web, the command line, and API.
OpenAI builds and maintains Whisper Large, and the product first shipped in 2022. Development happens publicly on GitHub with 24.5k stars and 83 commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 8 similar tools. Key capabilities include speech recognition, multilingual support, and transcription.
Latest indexed changes and source events
openai/whisper-large-v3 verified by the PulseGate indexer
Other apps tracked under the same category.