Interested in sound generation. Let's make AI more powerful.
FastSpeech. The Implementation of FastSpeech based on pytorch.
885FastVocoder. Include Basis-MelGAN, MelGAN, HifiGAN and Multiband-HifiGAN, maybe NHV in the future.
157Transformer-TTS. TTS model based on Transformer.
57FastSpeech2. The Implementation of FastSpeech2 Based on Pytorch.
52ConvTasNet4BasisMelGAN. This repo contains conv-tasnet for basis-melgan. If you want to get code of basis-melgan, please refer to FastVocoder.
21CLONE.
20Tacotron2-Pytorch. follow NVIDIA, simplify it and support data parallel.
13Lifelong-Learning-Tacotron2. MultiSpeaker Tacotron2 using LifeLong Learning.
13tacotron2.xcmyz. new version of tacotron2 (old version: https://github.com/xcmyz/Tacotron2-Pytorch)
8Hackathon-EnglishLearning. Voice Scoring System.
8LM-Tacotron2. Tacotron2 Combine with Language Model (BERT).
7SpeakerVerification. Speaker Verification (GE2E Loss)
7Gobang-AI. A C++ Implementation of Gobang AI.
6Speech-Resources. 语音方向实验室/公司/资源/实习等,欢迎推荐或自荐(排名不分先后)
3Forced-Alignment. using montreal-forced-aligner.
2bert-race. BERT/ALBERT based model for RACE dataset, support multi-worker, multi-GPU, FP16 and bind CPU.
2AVX-programming. CPU acceleration using AVX (Advanced Vector Extensions)
1xcmyz.
1ExpressionTransformation. prefix expression, infix expression, postfix expression.
1diffwave. DiffWave is a fast, high-quality neural vocoder and waveform synthesizer.
1Large-Audio-Models. Keep track of big models in audio domain, including speech, singing, music etc.
1Calculator. A Calculator implemented in Python.
1VAE-Tacotron. A Pytorch Implementation of Tacotron Combined with VAE
1FaceDetection. Python
1Polynomial-Calculator. 基于Python实现的带有图形界面的多项式计算器
1