Whisper-Finetune. Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inference and support Web deployment, Windows desktop deployment, and Android deployment
318Ke-Omni-R. Ke-Omni-R is an advanced audio reasoning model and achieved SOTA on MMAU
60STRAIGHT. This is a speech analysis, modification and synthesis system
54DeepNerualNetwork. Deep learning tools including RBM and DBM
8DTW_based_Isolated_Word_Recognition. MATLAB
7KNN. K近邻法
3shuaijiang.github.io. Record what I learn and study, including something useful and enjoyful
2doubanAPI. Can be looked as a Crawler, which crawling information from douban.com through douban API, written in Python.(通过豆瓣API获取电影信息)
2typhoon-blade. Automatically exported from code.google.com/p/typhoon-blade
2SmartWay. 智能交通查询系统
2ke-data-juicer. A one-stop data processing system to make data higher-quality, juicier, and more digestible for LLMs! 🍎 🍋 🌽 ➡️ ➡️🍸 🍹 🍷为大语言模型提供更高质量、更丰富、更易”消化“的数据!
1books. books which are interesting
1rnn-speech-denoising. Recurrent neural network training for noise reduction in robust automatic speech recognition
1HTKModelConvert. The Perl scripts are used for converting the binary model to text model in HTK and HTS
1LeetCode. LeetCode
1HTK_HTS. All versions of htk and hts, checkout the git tag to the htk/hts version you want
1athena-2. Python
1FemaleMaleDatabase. This database includes height and weight of females and males, the first column is the height(cm) and the second column is the weight(kg)
1speech-resynthesis. An official reimplementation of the method described in the INTERSPEECH 2021 paper - Speech Resynthesis from Discrete Disentangled Self-Supervised Representations.
1keras. Deep Learning library for Python. Convnets, recurrent neural networks, and more. Runs on Theano and TensorFlow.
1