This is your work, valued
Bachelor's of Mechatronics || Working on AI powered projects || Dev Rel @ Fish Audio || YouTuber @Jarods_Journey
ai-voice-cloning. Python
778audiobook_maker. Python
581Vivy. Python
153rvc-tts-pipeline. TTS pipeline that uses RVC to enhance audio quality and cloning
148index-tts. An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
146audiosplitter_whisper. Python
101StyleTTS-WebUI. Python
74F5-TTS. Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"
54beatrice_trainer_webui. Python
47VibeVoice. Frontier Open-Source Text-to-Speech
47open-neruosama. Python
43StyleTTS2. StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
36dataset-maker. Python
34audiosplitter. Python
33GPT-SoVITS-Package. 1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
29chatterbox. SoTA open-source TTS
26tortoise_dataset_tools. Misc. tools/scripts that I made to use for tortoise
19ChatGPT-and-Whiper-with-TTS. Python
18DMOSpeech2. Jupyter Notebook
16tortoise_api. A simple module for making a request to the tortoise gradio page.
16tts-pipeline-example. Python
12rvc-python. Using RVC via console or python scripts
12seed-vc. zero-shot voice conversion & singing voice conversion, with real-time support
11SmartGPT. Python
11voice_recorder. Python
10StyleTTS-ZS. StyleTTS-ZS: Efficient High-Quality Zero-Shot Text-to-Speech Synthesis with Distilled Time-Varying Style Diffusion
10rvc. Installable package for rvc voice inferencing
9youtube-transcriber. Python
7openai_tts_example. Python
5whisper_real_time. Real time transcription with OpenAI Whisper.
5tortoise-tts. Python
4VibeVoice-finetuning. Unofficial WIP LoRa Finetuning repository for VibeVoice
4VoiceStar. VoiceStar: Robust, Duration-controllable TTS that can Extrapolate
3kits-api-test. Python
3tortoise_tts_api. Python
3Midjourney-Prompts.
2bitsandbytes-windows. Python
2audiocraft. Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor / tokenizer, along with MusicGen, a simple and controllable music generation LM with textual and melodic conditioning.
2beatrice_trainer. Python
2higgs-audio. Text-audio foundation model from Boson AI
2flash-attention. Fast and memory-efficient exact attention
2ai_workbench_example. Python
2transcriber_app. Real time speech to text transcription app.
2styletts-api. Python
1gpt4all_voice_wrapper. Python
1youtube_miner. Python
1styletts2_phonemizer. Python
1fish-speech. SOTA Open Source TTS
1DL-Art-School. Python
1eSpeak-Files.
1MECA-470-UBR.
1uv_pyproject_windows_template. Helper for making sure uv works with pytorch
1Meca470_UBR_Project. Moving joints of a UBR robot via a ROS Virtual machine
1cuda-samples. Samples for CUDA Developers which demonstrates features in CUDA Toolkit
1Meca-470. Moving joints of a UBR robot via a ROS Virtual machine
1