Rare find

mini-omni. open-source multimodal large language model that can hear, talk while thinking. Featuring real-time end-to-end speech input and streaming audio output conversational capabilities.

github.com/gpt-omni/mini-omni

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.