Whisper Large v3 is an open-source automatic speech recognition model developed by OpenAI. It transcribes spoken audio into text, supports multiple languages, and is robust to noise, making it suitable for developers and researchers building speech-to-text applications.
Whisper Large is a Speech to text project. It focuses on automating the transcription of spoken audio into accurate text for various applications. It is built as an open-source project for developers and researchers needing speech-to-text capabilities. Whisper Large is open source under the BSD-3-Clause license. Whisper Large is available on the web, the command line, and API.
It is developed by OpenAI, and it first shipped in 2022. Development happens publicly on GitHub with 24.5k stars and 83 commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 8 similar projects. Among its 5 catalogued features are speech recognition, multilingual support, and transcription.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do