This is your work, valued

China, Beijing

Heinrich Dinkel

Elite
@RicherMans

日新月异

GPV. Repository for our Interspeech2020 general-purpose voice activity detection (GPVAD) paper

141

PLDA. An LDA/PLDA estimator using KALDI in python for speaker verification tasks

102

Datadriven-GPVAD. The codebase for Data-driven general-purpose voice activity detection.

93

Dasheng. Source for the Interspeech 2024 Paper "Scaling up masked audio encoder learning for general audio classification"

86

AudioCaption. Dataset and baseline for the first Audiocaption task

79

CED. Source code for Consistent ensemble distillation for audio tagging

75

SAT. Streaming Audiotransformers for online Audio tagging

57

text_based_depression. Source code for the paper "Text-based Depression Detection: What Triggers An Alert"

50

PSL. Source code for ICASSP2022 "Pseudo Strong labels for large scale weakly supervised audio tagging"

31

UIT_Mobile. Source Code for the Paper "UNIFIED KEYWORD SPOTTING AND AUDIO TAGGING ON MOBILE DEVICES WITH TRANSFORMERS"

24

CDur. Repository for the paper "Towards duration robust weakly supervised sound event detection"

23

Speaker-Anti-Spoofing-Classifiers. Baselines and Classifiers for speaker anti-spoofing detection

18

HEAR2021_EfficientLatent. Submission to the HEAR2021 Challenge

17

Dcase2018_pooling. Repo for our pooling approach on the DCASE2018 task4

16

XiaomiVPN. A short introduction how to successfully install a VPN client on a Xiaomi router.

15

pymir. Music IR Library for Python

13

SpokenLanguageClassifiers. Pretrained spoken language classifiers from audio.

10

HEAR_CED. Hear evaluation for CED models.

9

SpecView. A spectrogram viewer - Made for TTS/TTA deployment

5

NumericalAnalysis. Homework for Numerical Analysis

3

audiodataload. Audiodataloaders for raw wave and HTK features in torch.

2

ImageNet21K. Official Pytorch Implementation of: "ImageNet-21K Pretraining for the Masses"(NeurIPS, 2021) paper

2

Sublime3-pydoc. Sublime 3 Pydoc plugin

2

hifi-gan. HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

1

torchhtk. A simple HTK (Hidden markov kit) dataloader for torch

1

Nanopi-R4S. My NanoPi R4S builds

1

richermans.github.io. My Blog / Jekyll Themes / PWA

1

Gentleplayer. Simple, easy, no extras playlist generator for Android

1

coc-pyright. Pyright extension for coc.nvim

1

MatTheory. Repo for the Latex files

1

hf_transformers_custom_model_dasheng. 🤗 Transformers custom models (Dasheng)

1

kaldi-io-for-python. Python functions for reading kaldi data formats. Useful for rapid prototyping with python.

1