ASR-LLM-TTS. This is a speech interaction system built on an open-source model, integrating ASR, LLM, and TTS in sequence. The ASR model is SenceVoice, the LLM models are QWen2.5-0.5B/1.5B, and there are three TTS models: CosyVoice, Edge-TTS, and pyttsx3

github.com/HonglinChu/ASR-LLM-TTS

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.