Jarod Mica

Elite
@JarodMica

Bachelor's of Mechatronics || Working on AI powered projects || Dev Rel @ Fish Audio || YouTuber @Jarods_Journey

ai-voice-cloning. Python

778

audiobook_maker. Python

581

Vivy. Python

153

rvc-tts-pipeline. TTS pipeline that uses RVC to enhance audio quality and cloning

148

index-tts. An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System

146

audiosplitter_whisper. Python

101

StyleTTS-WebUI. Python

74

F5-TTS. Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"

54

beatrice_trainer_webui. Python

47

VibeVoice. Frontier Open-Source Text-to-Speech

47

open-neruosama. Python

43

StyleTTS2. StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models

36

dataset-maker. Python

34

audiosplitter. Python

33

GPT-SoVITS-Package. 1 min voice data can also be used to train a good TTS model! (few shot voice cloning)

29

chatterbox. SoTA open-source TTS

26

tortoise_dataset_tools. Misc. tools/scripts that I made to use for tortoise

19

ChatGPT-and-Whiper-with-TTS. Python

18

DMOSpeech2. Jupyter Notebook

16

tortoise_api. A simple module for making a request to the tortoise gradio page.

16

tts-pipeline-example. Python

12

rvc-python. Using RVC via console or python scripts

12

seed-vc. zero-shot voice conversion & singing voice conversion, with real-time support

11

SmartGPT. Python

11

voice_recorder. Python

10

StyleTTS-ZS. StyleTTS-ZS: Efficient High-Quality Zero-Shot Text-to-Speech Synthesis with Distilled Time-Varying Style Diffusion

10

rvc. Installable package for rvc voice inferencing

9

youtube-transcriber. Python

7

openai_tts_example. Python

5

whisper_real_time. Real time transcription with OpenAI Whisper.

5

tortoise-tts. Python

4

VibeVoice-finetuning. Unofficial WIP LoRa Finetuning repository for VibeVoice

4

VoiceStar. VoiceStar: Robust, Duration-controllable TTS that can Extrapolate

3

kits-api-test. Python

3

tortoise_tts_api. Python

3

Midjourney-Prompts.

2

bitsandbytes-windows. Python

2

audiocraft. Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor / tokenizer, along with MusicGen, a simple and controllable music generation LM with textual and melodic conditioning.

2

beatrice_trainer. Python

2

higgs-audio. Text-audio foundation model from Boson AI

2

flash-attention. Fast and memory-efficient exact attention

2

ai_workbench_example. Python

2

transcriber_app. Real time speech to text transcription app.

2

styletts-api. Python

1

gpt4all_voice_wrapper. Python

1

youtube_miner. Python

1

styletts2_phonemizer. Python

1

fish-speech. SOTA Open Source TTS

1

DL-Art-School. Python

1

eSpeak-Files.

1

MECA-470-UBR.

1

uv_pyproject_windows_template. Helper for making sure uv works with pytorch

1

Meca470_UBR_Project. Moving joints of a UBR robot via a ROS Virtual machine

1

cuda-samples. Samples for CUDA Developers which demonstrates features in CUDA Toolkit

1

Meca-470. Moving joints of a UBR robot via a ROS Virtual machine

1
55
Apply