Research Scientist working on ML/AI for speech tech. PhD from the University of Edinburgh, CSTR
Angular-Penalty-Softmax-Losses-Pytorch. Angular penalty loss functions in Pytorch (ArcFace, SphereFace, Additive Margin, CosFace)
499TDNN. Time delay neural network (TDNN) implementation in Pytorch using unfold method
204simple_diarizer. Simplified diarization pipeline using some pretrained models - audio file to diarized segments in a few lines of code
158Factorized-TDNN. PyTorch implementation of the Factorized TDNN (TDNN-F) from "Semi-Orthogonal Low-Rank Matrix Factorization for Deep Neural Networks" and Kaldi
149GE2E-Loss. Pytorch implementation of Generalized End-to-End Loss for speaker verification
88nn-similarity-diarization. Neural network based similarity scoring for diarization (pytorch implementation of "LSTM based Similarity Measurement with Spectral Clustering for Speaker Diarization")
43MTL-Speaker-Embeddings. Code for the paper: "Leveraging speaker attribute information using multi task learning for speaker verification and diarization" presented at Interspeech 2021
26dropclass_speaker. DropClass and DropAdapt - repository for the paper accepted to Speaker Odyssey 2020
22splitdim_disentangle. Disentangling speaker embeddings using multi-task learning and adversarial training
9