This is your work, valued
active_learning. The active learning algorithm, mismatch-first farthest-traversal. Implementation and visualization.
★ 12childrenize. Signal processing method to convert adult speech into child-like
★ 7audio_embedding_segmentation. Extracting audio embeddings based on a gated CNN trained using Audioset. A change point detection algorithm based on embeddings.
★ 3ergodox_nordic_colemak. C
★ 1audioset_tagging_cnn. Python
★ 1.8kMaskSpec. The Pytorch implementation of paper: Masked Spectrogram Prediction For Self-Supervised Audio Pre-Training
★ 51DeepComplexCRN. HTML
★ 482ergodox_nordic_colemak. C
★ 1Python-WORLD. Python
★ 153vits_chinese. Best practice TTS based on BERT and VITS with some Natural Speech Features Of Microsoft; Support ONNX streaming out!
★ 1.2kchildrenize. Signal processing method to convert adult speech into child-like
★ 9AEC-Challenge. AEC Challenge
★ 496android-demo-app. PyTorch android examples of usage in applications
★ 1.6kTCN. Sequence modeling benchmarks and temporal convolutional networks
★ 4.5kDeepXi. Deep Xi: A deep learning approach to a priori SNR estimation implemented in TensorFlow 2/Keras. For speech enhancement and robust ASR.
★ 523IRM-based-Speech-Enhancement-using-LSTM. Ideal Ratio Mask (IRM) Estimation based Speech Enhancement using LSTM
★ 122ru_norm_kaggle. Python
★ 8Speech-Separation-Paper-Tutorial. A must-read paper for speech separation based on neural networks
★ 953Conv-TasNet. Python
★ 337vits. VITS: Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech
★ 7.9kdcase2022_task1_baseline. Baseline system for DCASE 2022 task 1
★ 13traditional-speech-enhancement. Spectral Subtraction, Wiener Filtering, MMSE
★ 129Python-Wrapper-for-World-Vocoder. A Python wrapper for the high-quality vocoder "World"
★ 790python-speech-enhancement. a python library for speech enhancement
★ 82AutoEq. Automatic headphone equalization from frequency responses
★ 16kawesome-speech-enhancement. speech enhancement\speech seperation\sound source localization
★ 1.2kactive_learning. The active learning algorithm, mismatch-first farthest-traversal. Implementation and visualization.
★ 12mongol_code. Unicode conversion library for traditional Mongolian script
★ 10NISQA. NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment
★ 965wav2vec2-live. A live speech recognition using Facebooks wav2vec 2.0 model.
★ 379Real-Time-Voice-Cloning. Clone a voice in 5 seconds to generate arbitrary speech in real-time
★ 60ks3prl. Self-Supervised Speech Pre-training and Representation Learning Toolkit
★ 2.6kpytorch-wavenet. An implementation of WaveNet with fast generation
★ 1kemacs-from-scratch. An example of a fully custom Emacs configuration developed live on YouTube!
★ 1.9kformant_edit. software for editing dynamic formant measurements
★ 15jackaudio.github.com. jackaudio's Website and Wiki hosted at github
★ 480audio_embedding_segmentation. Extracting audio embeddings based on a gated CNN trained using Audioset. A change point detection algorithm based on embeddings.
★ 3torch_audioset. PyTorch transcribed audioset classifier, including VGGish and YAMNet, along with utils to manipulate autioset category ontology.
★ 106matrix-creator-hal. Hardware Abstraction Layer for MATRIX Creator & MATRIX Voice
★ 73dcase2019-task5-urban-sound-tagging. 1st place solution to the DCASE 2019 - Task 5 - Urban Sound Tagging
★ 30onnx. Open standard for machine learning interoperability
★ 21kawesome-emacs. A community driven list of useful Emacs packages, libraries and other items.
★ 9.3kyoutube-dl. Command-line program to download videos from YouTube.com and other video sites
★ 141kweak_feature_extractor. Python
★ 59ESC-50. ESC-50: Dataset for Environmental Sound Classification
★ 1.9kmodels. Models and examples built with TensorFlow
★ 78k