streaming-sensevoice. Pseudo Streaming SenseVoice with Hotwords
467g2p-mix. Grapheme-to-Phoneme for Mixed Chinese (Mandarin or Cantonese) and English.
115pyannote-onnx. ONNX Inference of Pyannote Segmentation
99pyrnnoise. Python Wrapper for RnNoise v0.2
78pysilero. Python Wrapper of Silero VAD
63torchfa. Torch Audio Forced Aligner for Mixed Chinese (Mandarin or Cantonese) and English.
61wetext. Python runtime for WeTextProcessing (does not depend on Pynini)
53audiolab. A streaming audio reader, processor, and writer built on top of soundfile, and PyAV (bindings for FFmpeg)
39asr-decoder. CTC decoder with hotwords for ASR.
38streaming-vocos. Streaming Vocos
31welm. One command to build TLG.fst for WeNet.
30compute-wer. Compute WER and SER for speech recognition evaluation
27streaming-ChatTTS. Jupyter Notebook
23audio-pipeline. Python
23streaming-tts-webui. Streaming Text to Speech Web UI
22ngram-punctuator. An N-gram punctuator for Chinese and English.
20wavesurfer. For audio visualization and playback in Jupyter notebooks.
18speaker-diarization. Offline Speaker Diarization with SenseVoice by Sherpa ONNX.
15streaming-asr. One command to start a streaming ASR server.
12ipyaudio. A Jupyter Widget for Web Audio Playing and Recording.
11pyvisqol. Python Wrapper of visqol
11streaming-dvae. Python
8wenet-openfst-android. C++
6online-fbank. Python
6openfst-android. Build OpenFst with gflags and glog for Android.
5snr. SNR example audios
5leetcode. Java
5asr-hotword. C++
5fireredasr. Python
5ABFL. Python
4tts-pqe. Perceptual Quality Estimator for Speech
4audio-codec. Python
2dotfiles. Shell
2play-remote-wavs. PowerShell
2kenlm. KenLM: Faster and Smaller Language Model Queries
2simd. SIMD demo
2asr. Python
2datasets-pyannote. Automatically setup the AISHELL-4 and MSDWild dataset for usage with pyannote-database (and pyannote-audio)
2audio-mask. Python
1docker-pusher. 使用 Github Action 将国外的 Docker 镜像转存到阿里云私有仓库,供国内服务器使用,免费易用
1shenlan-asr. C
1bert-g2p. BERT with G2P
1