MOSS-Audio-Tokenizer is an open neural audio codec designed for converting speech audio into compact discrete tokens. It serves as a foundational component for the MOSS TTS family and other speech AI models. The model is distributed on Hugging Face with full weights under an Apache-2.0 license and integrates directly with the Transformers library.
In the Foundation models & chat space, MOSS Audio Tokenizer takes a focused approach. It focuses on converting raw audio into discrete tokens for efficient speech modeling and text-to-speech systems. MOSS Audio Tokenizer is an open-source project aimed at AI researchers and developers. The project is open source (Open Source). It ships for the web and API.
It is developed by OpenMOSS-Team. Among its 4 catalogued features are Feature Extraction, Neural Codec, and Speech Tokenization.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do