Rare find

MOSS-TTSD. MOSS-TTSD is a spoken dialogue generation model designed for expressive multi-speaker synthesis. It features long-context modeling, flexible speaker control, and multilingual support, while enabling zero-shot voice cloning from short audio references.

github.com/OpenMOSS/MOSS-TTSD

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.