Speech AI | Research Scientist @ ByteDance Seed | PhD @ NTU | ex-Meta, Google DeepMind, Amazon
Full-Duplex-Bench. A Benchmark for Evaluating Turn-Taking and Overlap Handling in Full-Duplex Spoken Dialogue Models
247StyleTalk. Official release of StyleTalk dataset.
75DUAL-textless-SQA. Textless (ASR-transcript free) Spoken Question Answering. The official release of NMSQA dataset and the implementation of "DUAL: Textless Spoken Question Answering with Speech Discrete Unit Adaptive Learning" paper.
35Test-time-adaptation-ASR-SUTA. Test-time adaptation for speech recognition model by single utterance. The official implementation of "Listen, Adapt, Better WER: Source-free Single-utterance Test-time Adaptation for Automatic Speech Recognition" paper.
23E2E-ASR-Pytorch. Python
7unsupervised_ASR_challenge. Facebook AI Research Sequence-to-Sequence Toolkit written in Python.
4End-to-End-jointCTC-Attention-ASR. Jupyter Notebook
3DanielLin94144.github.io. Github Pages template for academic personal websites, forked from mmistakes/minimal-mistakes
1Protect-Your-Voice. Official implementation of Meta-StyleSpeech and StyleSpeech
1software-studio-NTHU-. this is a repository for SOFTWARE STUDIO course from NTHU
1Algorithm. NTHU EE3980
1FixMatch-pytorch. Unofficial PyTorch implementation of "FixMatch: Simplifying Semi-Supervised Learning with Consistency and Confidence"
1FDB-v2-dev. Python
1Emphasized-Talk. Official release of Emphasized-Talk
1