This is your work, valued

Nagoya, Japan

Ryuichi Yamamoto

Elite
@r9y9

Speech Synthesis, Voice Conversion, Machine Learning, Singing Voice Synthesis

wavenet_vocoder. WaveNet vocoder

2.4k

deepvoice3_pytorch. PyTorch implementation of convolutional neural networks-based text-to-speech synthesis models

2k

gantts. PyTorch implementation of GAN-based text-to-speech synthesis and voice conversion (VC)

518

pysptk. A python wrapper for Speech Signal Processing Toolkit (SPTK).

451

nnmnkwii. Library to build speech synthesis systems designed for easy and fast prototyping.

399

tacotron_pytorch. PyTorch implementation of Tacotron speech synthesis model.

310

ttslearn. ttslearn: Library for Pythonで学ぶ音声合成 (Text-to-speech with Python)

269

pyopenjtalk. Python wrapper for OpenJTalk

255

pylibfreenect2. A python interface for libfreenect2

138

SPTK. A modified version of Speech Signal Processing Toolkit (SPTK)

89

pyreaper. A python wrapper for REAPER

81

sinsy. A fork of sinsy: HMM/DNN-based singing voice synthesis system

74

open_jtalk. A fork of open_jtalk

71

nnmnkwii_gallery. A collection of examples demonstrating how we can build speech synthesis systems using nnmnkwii.

70

gossp. Speech Signal Processing for Go (not maintained)

68

pysinsy. Python wrapper for Sinsy

53

jsut-lab. HTS-style full-context labels for JSUT v1.1

51

VoiceConversion.jl. [Deprecated] Statistical Voice Conversion in Julia. See the website link for new library

38

icassp2020-espnet-tts-merlin-baseline. ICASSP 2020 ESPnet-TTS: Merlin baseline system

37

Tacotron-2. DeepMind's Tacotron-2 Tensorflow implementation

34

WORLD.jl. A lightweight julia wrapper for WORLD - a high-quality speech analysis, modification and synthesis system

30

nnet. A small collection of neural network algorithms in Go (no longer maintained)

29

kiritan_singing. Labels for kiritan_singing data with extra resources for DNN-based singing voice synthesis (SVS) systems.

28

hts_engine_API. A fork of hts_engine_API

23

bayesian-kalmanfilter. Variational Baysian Kalman Filter

21

nlp100. Assignments for NLP 100

19

MelGeneralizedCepstrums.jl. Mel-Generalized Cepstrum analysis

19

segmentation-kit. Speech Segmentation Toolkit using Julius

18

WORLD. A modified version of WORLD (original: http://ml.cs.yamanashi.ac.jp/world/english/index.html)

13

Colaboratory. Colaboratory notebooks

13

SynthesisFilters.jl. Speech waveform synthesis filters

13

SPTK.jl. A thin Julia wrapper for Speech Signal Processing Toolkit (SPTK) API

12

ConstantQ.jl. A fast constant-q transform in Julia

11

go-dsp. Digital Signal Processing for Go

10

robust_pca. Robust Principal Component Analysis

10

RobustPCA.jl. Robust Principal Component Analysis in Julia

7

kiritan_singing_extra. Extra resources derived from https://github.com/mmorise/kiritan_singing for DNN-based singing voice synthesis

7

VCTK-lab. Full context labels for VCTK corpus extracted by Merlin & speech tools

7

demos. Deprecated. See https://github.com/r9y9/website

6

BNMF.jl. Bayesian Non-negative Matrix Factorization

6

World-cmake. WORLD with CMake support

6

stav. Statistical voice conversion written in Go for signal processing backend, Python for model training and parameter conversions

6

jvs_r9y9. JVS (Japanese versatile speech) コーパスの自作のラベル

5

naive_bayes. Naive Bayes implementation with digit recognition sample

5

julia-nmf-ss-toy. NMF-based Music Source Separation Demo in Julia

5

Libfreenect2.jl. A Julia wrapper for libfreenect2

4

svdd2024seg. Python

4

dotfiles. Dotfiles

4

REAPER. C-interface for REAPER (see cwrap/ for details)

3

SiFiGAN. Official implementation of the source-filter HiFiGAN vocoder

2

fastdtw. A Python implementation of FastDTW

2

commonvoice-lab. HTS style full-context labels for common voice

2