Beijing, China

shuaijiang

Advanced
@shuaijiang

Whisper-Finetune. Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inference and support Web deployment, Windows desktop deployment, and Android deployment

318

Ke-Omni-R. Ke-Omni-R is an advanced audio reasoning model and achieved SOTA on MMAU

60

STRAIGHT. This is a speech analysis, modification and synthesis system

54

DeepNerualNetwork. Deep learning tools including RBM and DBM

8

DTW_based_Isolated_Word_Recognition. MATLAB

7

KNN. K近邻法

3

shuaijiang.github.io. Record what I learn and study, including something useful and enjoyful

2

doubanAPI. Can be looked as a Crawler, which crawling information from douban.com through douban API, written in Python.(通过豆瓣API获取电影信息)

2

typhoon-blade. Automatically exported from code.google.com/p/typhoon-blade

2

SmartWay. 智能交通查询系统

2

ke-data-juicer. A one-stop data processing system to make data higher-quality, juicier, and more digestible for LLMs! 🍎 🍋 🌽 ➡️ ➡️🍸 🍹 🍷为大语言模型提供更高质量、更丰富、更易”消化“的数据!

1

books. books which are interesting

1

rnn-speech-denoising. Recurrent neural network training for noise reduction in robust automatic speech recognition

1

HTKModelConvert. The Perl scripts are used for converting the binary model to text model in HTK and HTS

1

LeetCode. LeetCode

1

HTK_HTS. All versions of htk and hts, checkout the git tag to the htk/hts version you want

1

athena-2. Python

1

FemaleMaleDatabase. This database includes height and weight of females and males, the first column is the height(cm) and the second column is the weight(kg)

1

speech-resynthesis. An official reimplementation of the method described in the INTERSPEECH 2021 paper - Speech Resynthesis from Discrete Disentangled Self-Supervised Representations.

1

keras. Deep Learning library for Python. Convnets, recurrent neural networks, and more. Runs on Theano and TensorFlow.

1
20
Apply