The answer is out there, Neo, and it's looking for you, and it will find you if you want it to.
UnsupSeg. Self-Supervised Contrastive Learning for Unsupervised Phoneme Segmentation (INTERSPEECH 2020)
147SegFeat. Phoneme Boundary Detection using Learnable Segmental Features (ICASSP 2020)
83HideAndSpeak. Python
42audiogen. HTML
23speech-resynthesis. An official reimplementation of the method described in the INTERSPEECH 2021 paper - Speech Resynthesis from Discrete Disentangled Self-Supervised Representations.
2felixkreuk.github.io. A beautiful, simple, clean, and responsive Jekyll theme for academics
2Score-Entropy-Discrete-Diffusion. Discrete Diffusion Modeling by Estimating the Ratios of the Data Distribution (https://arxiv.org/abs/2310.16834)
1_felixkreuk.github.io. HTML
1