This is your work, valued
whisper. Robust Speech Recognition via Large-Scale Weak Supervision
★ 106kVoiceFlow-TTS. [ICASSP 2024] This is the official code for "VoiceFlow: Efficient Text-to-Speech with Rectified Flow Matching"
★ 376RectifiedFlow. Official Implementation of Rectified Flow (ICLR2023 Spotlight)
★ 1.6kSingFake. Official Repository for "SingFake: Singing Voice Deepfake Detection"
★ 64FSD-Dataset. This repository presents FSD dataset for song deepfake detection.
★ 24iNSFC. An awesome LaTeX template for NSFC proposal.
★ 514tortoise-tts. A multi-voice TTS system trained with an emphasis on quality
★ 15kvector-quantize-pytorch. Vector (and Scalar) Quantization, in Pytorch
★ 4kCross-Speaker-Emotion-Transfer. PyTorch Implementation of ByteDance's Cross-speaker Emotion Transfer Based on Speaker Condition Layer Normalization and Semi-Supervised Training in Text-To-Speech
★ 194EmotionControllableTextToSpeech. HTML
★ 21VAENAR-TTS. The official implementation of VAENAR-TTS, a VAE based non-autoregressive TTS model.
★ 144SpeechSplit. Unsupervised Speech Decomposition Via Triple Information Bottleneck
★ 697TextClassifier_Transformer. 个人基于谷歌开源的BERT编写的文本分类器(基于微调方式),可自由加载NLP领域知名的预训练语言模型BERT、Bert-wwm、Roberta、ALBert以及ERNIE1.0
★ 42bert. TensorFlow code and pre-trained models for BERT
★ 40kclip-as-service. 🏄 Scalable embedding, reasoning, ranking for images and sentences with CLIP
★ 13ka_journey_into_math_of_ml. 汉语自然语言处理视频教程-开源学习资料
★ 1.7knlp-materials-share. 自然语言处理NLP(自然语言生成NLG、自然语言理解NLU)、自然语言学术会议大盘点、自然语言大佬介绍、NLP研究机构、NLP资料分享、NLP学习资源分享、NLP学术论文介绍
★ 186MelNet. Implementation of "MelNet: A Generative Model for Audio in the Frequency Domain"
★ 209