Speech Synthesis, Voice Conversion, Machine Learning, Singing Voice Synthesis
wavenet_vocoder. WaveNet vocoder
2.4kdeepvoice3_pytorch. PyTorch implementation of convolutional neural networks-based text-to-speech synthesis models
2kgantts. PyTorch implementation of GAN-based text-to-speech synthesis and voice conversion (VC)
518pysptk. A python wrapper for Speech Signal Processing Toolkit (SPTK).
451nnmnkwii. Library to build speech synthesis systems designed for easy and fast prototyping.
399tacotron_pytorch. PyTorch implementation of Tacotron speech synthesis model.
310ttslearn. ttslearn: Library for Pythonで学ぶ音声合成 (Text-to-speech with Python)
269pyopenjtalk. Python wrapper for OpenJTalk
255pylibfreenect2. A python interface for libfreenect2
138SPTK. A modified version of Speech Signal Processing Toolkit (SPTK)
89pyreaper. A python wrapper for REAPER
81sinsy. A fork of sinsy: HMM/DNN-based singing voice synthesis system
74open_jtalk. A fork of open_jtalk
71nnmnkwii_gallery. A collection of examples demonstrating how we can build speech synthesis systems using nnmnkwii.
70gossp. Speech Signal Processing for Go (not maintained)
68pysinsy. Python wrapper for Sinsy
53jsut-lab. HTS-style full-context labels for JSUT v1.1
51VoiceConversion.jl. [Deprecated] Statistical Voice Conversion in Julia. See the website link for new library
38icassp2020-espnet-tts-merlin-baseline. ICASSP 2020 ESPnet-TTS: Merlin baseline system
37Tacotron-2. DeepMind's Tacotron-2 Tensorflow implementation
34WORLD.jl. A lightweight julia wrapper for WORLD - a high-quality speech analysis, modification and synthesis system
30nnet. A small collection of neural network algorithms in Go (no longer maintained)
29kiritan_singing. Labels for kiritan_singing data with extra resources for DNN-based singing voice synthesis (SVS) systems.
28hts_engine_API. A fork of hts_engine_API
23bayesian-kalmanfilter. Variational Baysian Kalman Filter
21nlp100. Assignments for NLP 100
19MelGeneralizedCepstrums.jl. Mel-Generalized Cepstrum analysis
19segmentation-kit. Speech Segmentation Toolkit using Julius
18WORLD. A modified version of WORLD (original: http://ml.cs.yamanashi.ac.jp/world/english/index.html)
13Colaboratory. Colaboratory notebooks
13SynthesisFilters.jl. Speech waveform synthesis filters
13SPTK.jl. A thin Julia wrapper for Speech Signal Processing Toolkit (SPTK) API
12ConstantQ.jl. A fast constant-q transform in Julia
11go-dsp. Digital Signal Processing for Go
10robust_pca. Robust Principal Component Analysis
10RobustPCA.jl. Robust Principal Component Analysis in Julia
7kiritan_singing_extra. Extra resources derived from https://github.com/mmorise/kiritan_singing for DNN-based singing voice synthesis
7VCTK-lab. Full context labels for VCTK corpus extracted by Merlin & speech tools
7demos. Deprecated. See https://github.com/r9y9/website
6BNMF.jl. Bayesian Non-negative Matrix Factorization
6World-cmake. WORLD with CMake support
6stav. Statistical voice conversion written in Go for signal processing backend, Python for model training and parameter conversions
6jvs_r9y9. JVS (Japanese versatile speech) コーパスの自作のラベル
5naive_bayes. Naive Bayes implementation with digit recognition sample
5julia-nmf-ss-toy. NMF-based Music Source Separation Demo in Julia
5Libfreenect2.jl. A Julia wrapper for libfreenect2
4svdd2024seg. Python
4dotfiles. Dotfiles
4REAPER. C-interface for REAPER (see cwrap/ for details)
3SiFiGAN. Official implementation of the source-filter HiFiGAN vocoder
2fastdtw. A Python implementation of FastDTW
2commonvoice-lab. HTS style full-context labels for common voice
2