Whisper-medium is the 769M parameter version of OpenAI's Whisper speech recognition model family. It performs automatic speech recognition, language identification, and translation across many languages. The model is provided as open weights on Hugging Face and is widely used via the Transformers library for transcription and related audio tasks.
Whisper Medium is a Speech to text project. It focuses on converting spoken audio in multiple languages into accurate text transcriptions. It is built as an open-source project for developers. The project is open source (MIT). It runs on the web and API.
It is developed by OpenAI, and it first shipped in 2022. The project is developed in the open on GitHub with 105.3k stars. Among its 3 catalogued features are speech recognition, multilingual support, and transcription.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do