audio-prep-pipeline is a command-line tool that transforms raw MP3 audio files into clean, pretraining-ready WAV or FLAC format. It performs voice activity detection, validation, and generates dataset manifests suitable for training speech and audio models. Built for ML practitioners and researchers creating high-quality audio datasets from unprocessed recordings.
audio-prep-pipeline sits in PulseGate's AI & ML category. It focuses on preparing raw audio recordings for speech model pretraining and dataset creation. audio-prep-pipeline is an open-source project aimed at machine learning engineers. The project is open source (MIT). It runs on the command line.
Behind audio-prep-pipeline is nattkorat, and the product first shipped in 2026. The project is developed in the open on GitHub with 20 commits in the last 90 days. Among its 5 catalogued features are Audio Conversion, VAD Processing, and Dataset Manifest.
Latest indexed changes and source events
audio-prep-pipeline verified by the PulseGate indexer
Other apps tracked under the same category.