Rare find

MOSS-Audio-Tokenizer. MOSS-Audio-Tokenizer is a Causal Transformer-based audio tokenizer built on the CAT architecture. Trained on 3M hours of diverse audio, it supports streaming and variable bitrates, delivering SOTA reconstruction and strong performance in generation and understanding—serving as a unified interface for next-generation native audio language models.

github.com/OpenMOSS/MOSS-Audio-Tokenizer

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.