openai/whisper-base is an open-source automatic speech recognition (ASR) model that transcribes audio files into text. It supports multiple languages and is designed for developers and researchers working on speech-to-text applications. The model is easy to integrate into Python workflows.
In the Voice, TTS & speech space, Whisper Base takes a focused approach. It focuses on transcribing spoken audio into written text automatically and accurately. It is built as an open-source project for speech researchers and developers. Whisper Base is open source under the MIT license. The product ships for the web and the command line.
Behind Whisper Base is OpenAI, and the product first shipped in 2022. Development happens publicly on GitHub with 105.3k stars. Key capabilities include speech recognition, multilingual support, and transcription.
Latest indexed changes and source events
openai/whisper-base verified by the PulseGate indexer
Other apps tracked under the same category.