This is your work, valued
Electronic engineer and speech researcher. I'm interested in speech synthesis and generative speech models.
danbooru-pretrained. Pretrained pytorch models for the Danbooru2018 dataset
177simple-speaker-embedding. A speaker embedding network in Pytorch that is very quick to set up and use for whatever purposes.
91transfusion-asr. Transcribing Speech with Multinomial Diffusion, training code and models.
80simple-asgan. Training code and trained checkpoints for ASGAN.
62simple-autovc. A simple, performant re-implementation of AutoVC
22lmao. Language Model Accessor Object - Neural autocomplete for Overleaf
13simple-sashimi. A simple-to-use standalone Pytorch implementation of the SaShiMi speech synthesis model
4simple-speech-commands. A pretrained Pytorch classifier for the Google Speech Commands dataset that is very quick to set up and use.
3simple-pi-motion. Simple Raspberry Pi with camera module motion detection with Telegram notification
1mos-finetune-ssl. MOS predictor network with torchhub integration
1HierSpeechpp. Fork of HierSpeechpp with torchhub integration
1Experiment-K. Random experiments, data, and testing 'out-there' ideas involving machine learning and stats
1simple-diffwave. DiffWave, adapted to be simple to use and train.
1DiffWave-unconditional. Pytorch Reimplementation of DiffWave unconditional generation: a high quality waveform synthesizer.
1