audiototext-optimizer is an open-source command-line tool that optimizes audio and video files to improve the accuracy of AI speech-to-text transcription. It is designed for developers and data engineers preparing datasets for machine learning models.
In the CLI tools & terminal space, audiototext-optimizer takes a focused approach. It focuses on preparing and optimizing audio and video files for accurate AI speech-to-text transcription. audiototext-optimizer is an open-source project aimed at developers and data engineers working with speech-to-text AI. audiototext-optimizer is open source under the MIT license. It runs on the command line, and it can be self-hosted.
double2dev builds and maintains audiototext-optimizer, and it first shipped in 2026. Development happens publicly on GitHub with 2 commits in the last 90 days. Among its 4 catalogued features are audio optimization, video optimization, and speech-to-text prep.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do