This is your work, valued

Tsinghua University

Kai Li (李凯)

Elite
@JusperLee

Speech-Separation-Paper-Tutorial. A must-read paper for speech separation based on neural networks

952

Conv-TasNet. Conv-TasNet: Surpassing Ideal Time-Frequency Magnitude Masking for Speech Separation Pytorch's Implement

550

Dual-Path-RNN-Pytorch. Dual-path RNN: efficient long sequence modeling for time-domain single-channel speech separation implemented by Pytorch

468

TIGER. TIGER: Time-frequency Interleaved Gain Extraction and Reconstruction for Efficient Speech Separation

437

Apollo. Music repair method to convert lossy MP3 compressed music to lossless music.

396

TDANet. An efficient speech separation method

278

SonicSim. SonicSim: A customizable simulation platform for speech processing in moving sound source scenarios

277

SPMamba. Python

227

AudioTrust. AudioTrust: Benchmarking the Multi-faceted Trustworthiness of Audio Large Language Models

215

Dolphin. Python

186

Looking-to-Listen-at-the-Cocktail-Party. Executable code based on Google articles

166

Deep-Clustering-for-Speech-Separation. Pytorch implements Deep Clustering: Discriminative Embeddings For Segmentation And Separation

133

AFRCNN-For-Speech-Separation. Speech Separation Using an Asynchronous Fully Recurrent Convolutional Neural Network

131

IIANet. This is the demo of our paper "IIANet: An Intra- and Inter-Modality Attention Network for Audio-Visual Speech Separation".

111

Calculate-SNR-SDR. Script to calculate SNR and SDR using python

93

LRS3-For-Speech-Separation. Multi-modal speech separation task data generation script on LRS3 data set.

88

CTCNet. An Audio-Visual Speech Separation Model Inspired by Cortico-Thalamo-Cortical Circuits

82

UtterancePIT-Speech-Separation. According to funcwj's uPIT, the training code supporting multi-gpu is written, and the Dataloader is reconstructed.

67

Deep-Encoder-Decoder-Conv-TasNet. A PyTorch implementation of " AN EMPIRICAL STUDY OF CONV-TASNET "

51

AV-ConvTasNet. Unofficial Time Domain Audio Visual Speech Separation Implementation

45

S4M. Official implementation of Efficient Speech Separation Framework Based on Neural State-Space Models

28

Swift-Net. Power-Guided Grouped SRU for Real-Time Causal Audio-Visual Speech Separation

26

TFACM. Python

23

DANet-For-Speech-Separation. Pytorch implement of DANet For Speech Separation

21

Look2hear. A toolkit for researchers in the multimodal sound separation.

16

awesome-speech-enhancement. speech enhancement\speech seperation\sound source localization

15

Apollo-data-preprocess. Apollo training data preprocessing scripts

14

speechbrain-docs-zh-cn. SpeechBrain中文文档

12

Gull-Codec-Training. Python

12

Arxiv-New-Paper-Server. Arxiv automatically obtains the latest article service.

11

ExamOnline. This is a complete online exam system

10

My-Script-For-Audio-Process. Some convenient scripts for your own use

9

WeChatApp. Complete code of WeChat Mini Program

8

Grass. Python

7

GrabCut. C++

7

Deep-Learning. Learn to deep learning the code of your own records.

7

jusperlee.github.io. HTML

7

JusperLee.

7

Accelerator. Openmp Accelerator

7

player. Android Homework(3)

7

ELF-SR. Python

7

Time. My Android Project

7

speech_separation. Include some core functions and model to handle speech separation

6

Dive-into-DL-PyTorch. 本项目将《动手学深度学习》(Dive into Deep Learning)原书中的MXNet代码实现改为PyTorch实现。

6

BasicSR. Basic Super-Resolution Toolbox, including SRResNet, SRGAN, ESRGAN, etc.

5

Sound-of-Pixels. Codebase for ECCV18 "The Sound of Pixels"

5

pytorch-template. PyTorch deep learning projects made easy.

5

book-mate. 一款用于图书推荐、搜索、借阅、交流的微信小程序

4

espnet. End-to-End Speech Processing Toolkit

3

AdversarialNetsPapers. The classical paper list with code about generative adversarial nets

2

avlit. Official source code of the INTERSPEECH 2023 paper: "Audio-Visual Speech Separation in Noisy Environments with a Lightweight Iterative Model" (AVLIT)

2

GenerSpeech. PyTorch Implementation of GenerSpeech (NeurIPS'22): a text-to-speech model towards zero-shot style transfer of OOD custom voice.

1

FRA-RIR. Python

1

voxceleb_trainer. In defence of metric learning for speaker recognition

1

RSTnet. Real-time Speech-Text Foundation Model Toolkit

1

taichicon. TaichiCon: Taichi (Virtual) Conference

1