Deep Learning, Text To Speech and Singing Voice Synthesis.
ColorSplitter. A cli tool for split vocal timbre.
293R3MOE. [RecurrentNN × Regression × Regularized]-base Mouth Opening Estimation via SSL(Semi-supervised Learning).
26CODEY_Dataset. 一个第三方的泠鸢yousa歌声数据集
19ds2svc. Convert the "DiffSinger" model to an SVC (such as RVC) model
12Asaritsu4Diffsinger. A mult-languages(CN/JP) singing database for Diffsinger(OpenVPI).
6GPT-SoVITS-musa. 1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
6DiffSinger. An advanced singing voice synthesis system with high fidelity, expressiveness, controllability and flexibility based on DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism
5rectified-flow. 从零手搓Flow Matching(Rectified Flow)
4LaBei4DiffSinger. UTAU音库LaBei Rasubeku的DiffSinger移植
3auorange-torch. Audio LPC (linear prediction code) using mel spectorgram, compatible for LPCNet, torch version
2yousa-ling-diffsinger-v1. 泠鸢yousa的Diffsinger模型v1版
2libmininsf. C implementation of the MiniNSF source generator used by OpenVPI DiffSinger's NSF-HiFiGAN vocoder.
2game_ggml_cli. C++
2audio-preprocessing-scripts. 数据集自动化制作脚本
1mel2lpc_torch. Using Mel spectrum to calculate linear predictive coding coefficients (LPC), implemented by pytorch
1onedrive-vercel-index. OneDrive public directory listing, powered by Vercel and Next.js
1dataset-tools-mod4ass. A mod for dataset-tools(OpenVPI) for support .ass file(Advanced SubStation Alpha).
1ROSA-FPGA. A native ROSA (Rapid Online Suffix Automaton) implementation on Cyclone IV EP4CE10F17.
1