This is your work, valued
Deep Learning Scientist working on TTS @translated
machine-learning-of-pdes. Code of my master's thesis on "Physics Informed Machine Learning of Nonlinear Partial Differential Equations"
ai-audio-datasets. AI Audio Datasets (AI-ADS) 🎵, including Speech, Music, and Sound Effects, which can provide training data for Generative AI, AIGC, AI model training, intelligent audio tool development, and audio applications.
Orpheus-TTS. Towards Human-Sounding Speech
csm. A Conversational Speech Generation Model
speech-trident. Awesome speech/audio LLMs, representation learning, and codec models
StyleTTS-ZS_Scripts_AT. StyleTTS-ZS: Acoustic_Synth_training
StyleTTS-ZS. StyleTTS-ZS: Efficient High-Quality Zero-Shot Text-to-Speech Synthesis with Distilled Time-Varying Style Diffusion
TTS-arxiv-daily. Automatically Update Text-to-speech (TTS) Papers Daily using Github Actions (Update Every 12th hours)
tortoise-tts. A multi-voice TTS system trained with an emphasis on quality
Automatic-Audio-Dataset-Maker. Automatically cleaning, enhancing, segmenting, filtering, and formatting a dataset to fine tune or train a voice model.
awesome-audio-plaza. Daily tracking of awesome audio papers, including music generation, zero-shot tts, asr, audio generation
DEX-TTS. DEX-TTS: Diffusion-based EXpressive TTS with Style Modeling on Time Variability
audio-diffusion-pytorch. Audio generation using diffusion models, in PyTorch.
Amphion. Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.
audio-ai-timeline. A timeline of the latest AI models for audio generation, starting in 2023!
open-tts-tracker.
rl_project_spring_2021. DNQ+Pong