Whisper Large V2 is a web application that transcribes audio files into text using a large-scale speech recognition model. Users can upload audio and receive accurate transcriptions, making it useful for transcription tasks and accessibility.
In the Speech to text space, Whisper Large V2 takes a focused approach. It focuses on transcribing spoken audio into accurate text using advanced speech recognition models. It is built as a consumer product for users needing audio transcription and speech recognition. It is available for free. Whisper Large V2 is available on the web, and it can be self-hosted.
Behind Whisper Large V2 is sanchit-gandhi, and it first shipped in 2023. Key capabilities include speech-to-text, audio upload, and large model inference.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do