Researcher in the area of NLP, Ph.D. student at UFG, focusing on speech synthesis and recognition using deep learning and also professor at UFMT.
free-svc. [ICASSP 2025] FreeSVC: Towards Zero-shot Multilingual Singing Voice Conversion
95data_augmentation_for_asr. A set of audio augmentation techniques to perform noise insertion in datasets used for Automatic Speech Recognition.
49CML-TTS-Dataset. CML-TTS: A Multilingual Dataset for Speech Synthesis
36katube. KATube is a tool to automate the process of creating datasets for training Text-To-Speech (TTS) and Speech-To-Text (STT) models. From a list of YouTube playlists or YouTube channels, KATube will generate dataset with audios and texts.
26PTL-AI_Furnas_Dataset. PTL-AI Furnas Dataset: A Public Dataset for Fault Detection in Power Transmission Lines Using Aerial Images
24fault_detection_power_transmission_lines. Tensorflow Object Detection API for fault detection at power transmission lines.
19YourTTS. 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
16kabooks. KABooks is a tool to automate the process of creating datasets for training Text-To-Speech (TTS) and Speech-To-Text (STT) models. Using audiobooks, KABooks will generate dataset with segmented audios and aligned texts.
13BSpeech-MOS-Prediction. A model for predicting MOS that utilizes embeddings of supervised learning and self-supervised learning models, combined with embeddings of speaker verification models, to predict the MOS metric.
10useful_audio_scripts. Some useful scripts for audio
9capybara_dataset. This is a dataset composed of images of capybaras to be used for training a model for object detection
8youtube_diarization. Download audios from youtube and execute diarization
6Orpheus-TTS_unsloth_finetuning. Jupyter Notebook
5speaker_clustering. Python
5BRSpeech-Dataset. BRSpeech: A Portuguese Dataset for Speech Synthesis
5CleanSpecNet. Python
4CleanUNet2. Python
4overlapping_voices_detector. Python
4nvidia_tacotron2_multispeaker. Jupyter Notebook
4tacotron2. Tacotron 2 - PyTorch implementation with faster-than-realtime inference adapted for brazilian portuguese.
4CML-TTS-Toolkit. CML-TTS Conversion Tools
4flask_fault_detection_power_transmission_lines. Frontend and backend separated object detection tf2 demo build with Flask, TensorFlow 2.x.
4voice_gender_prediction. Python
3hifi-gan. Python
2wavegrad. A fast, high-quality neural vocoder.
2algoritmos_e_estrutura_de_dados_material_didatico. Material Didático da disciplina Algoritmos e Estrutura de Dados
2inteligencia_artificial_do_zero_ao_infinito. Material referente ao projeto sobre inteligência artificial com foco em visão computacional e detecão de objetos
2yolov5. YOLOv5 🚀 in PyTorch > ONNX > CoreML > TFLite
1XTTSv2-Finetuning-for-New-Languages. Python
1MeloTTS. High-quality multi-lingual text-to-speech library by MyShell.ai. Support English, Spanish, French, Chinese, Japanese and Korean.
1AI-Programming-using-Python. This repository contains implementation of different AI algorithms, based on the 4th edition of amazing AI Book, Artificial Intelligence A Modern Approach
1hifi-gan2. Python
1RVC-Demo. Python
1MOSNet-pytorch. Python
1Datasets-Portuguese-NLP. List of resources and tools developed with focus on Portuguese.
1capybara_image_segmentation. This repository presents how to train your own Image Segmentator Using TensorFlow Object Detection API.
1