This is your work, valued
seamless_communication. Foundational Models for State-of-the-Art Speech and Text Translation
★ 12kEdge-Punct-Casing. Python
★ 33kaldi-model-server. Simple Kaldi model server for chain (nnet3) models in online recognition mode directly from a local microphone
★ 35greek_podcasts_asr. Shell
★ 9mirascope. The LLM Anti-Framework
★ 1.5kfixwav. Quick utility to fix WAV files with incorrect lengths
★ 19cylimiter. A C++/Cython audio limiter for Python.
★ 25audiomentations. A Python library for audio data augmentation. Useful for making audio ML models work well in the real world, not just in the lab.
★ 2.3kmistral-inference. Official inference library for Mistral models
★ 11kgpt4all. GPT4All: Run Local LLMs on Any Device. Open-source and available for commercial use.
★ 77kCQT_toolbox_python. Constant-Q Transform Toolbox for Python/MATLAB
★ 39vocode-core. 🤖 Build voice-based LLM agents. Modular + open source.
★ 3.8kGigaSpeech. Large, modern dataset for speech recognition
★ 731kaldilm. Python wrapper for kaldi's arpa2fst
★ 38mitlm. MIT Language Modeling Toolkit
★ 120pocolm. Small language toolkit for creation, interpolation and pruning of ARPA language models
★ 92libriheavy. Libriheavy: a 50,000 hours ASR corpus with punctuation casing and context
★ 220json. JSON for Modern C++
★ 50kX-Punctuator. A PyTorch implementation of a punctuation prediction system using (B)LSTM, which automatically adds suitable punctuation into text without punctuation.
★ 63k2. FSA/FST algorithms, differentiable, with PyTorch compatibility.
★ 1.3kwikiextractor. A tool for extracting plain text from Wikipedia dumps
★ 4kstk. STK, the Brno speech recognition toolkit.
★ 5fdndlp. A speech dereverberation algorithm, also called wpe
★ 158keyczar. Easy-to-use crypto toolkit
★ 1.1k