This is your work, valued
dcase2023_task4b_baseline. Baseline code for DCASE 2023 task 4 B
★ 15dcase2021_task1a_baseline. Python
★ 14dcase2022_task1_baseline. Baseline system for DCASE 2022 task 1
★ 13pcutil. Rust
★ 1nlu. 1 line for thousands of State of The Art NLP models in hundreds of languages The fastest and most accurate way to solve text problems.
★ 968tools. Various utilities for processing the data.
★ 221OpenIE-standalone. PostScript
★ 587salt. SALT: STANDARDIZED AUDIO EVENT LABEL TAXONOMY
★ 16X-ACE. Python
★ 2macs-captioning-start-token. Python
★ 1ssl4birdsounds. Self-supervised representation learning for bird sounds (ICASSPW SASB 2024)
★ 10ATST-SED. This repo includes the official implementations of "Fine-tune the pretrained ATST model for sound event detection".
★ 175HTS-Audio-Transformer. The official code repo of "HTS-AT: A Hierarchical Token-Semantic Audio Transformer for Sound Classification and Detection"
★ 504DCASE2021_task6_v2. Code for CVSSP submission to DCASE 2021 Task 6
★ 36a-mask-guided-transformer-with-topic-token-for-remote-sensing-image-captioning. Python
★ 7TextToAudioGrounding. The dataset and baseline code for Text-to-Audio Grounding (TAG)
★ 49whisper-at. Code and Pretrained Models for Interspeech 2023 Paper "Whisper-AT: Noise-Robust Automatic Speech Recognizers are Also Strong Audio Event Taggers"
★ 422aac-datasets. Audio Captioning datasets for PyTorch.
★ 128transformer_workshop. Code for the Transformer workshop
★ 4audio-and-speech-tech-2022. Audio and Speech Technologies Workshop 2022, code examples
★ 4kapre. kapre: Keras Audio Preprocessors
★ 947netron. Visualizer for neural network, deep learning and machine learning models
★ 33kinterpretable_predictions. Interpretable Neural Predictions with Differentiable Binary Variables
★ 85byol-a. BYOL for Audio: Self-Supervised Learning for General-Purpose Audio Representation
★ 237wiki. This repo contains the source code for the deployment of the unofficial crowdsourced wiki for the Faculty of Information Technology and Communication Sciences at Tampere University.
★ 3sed_eval. Evaluation toolbox for Sound Event Detection
★ 161dcase_util. A collection of utilities for Detection and Classification of Acoustic Scenes and Events
★ 135sed_vis. Visualization toolbox for Sound Event Detection
★ 122fense. Fluency ENhanced Sentence-bert Evaluation (FENSE), metric for audio caption evaluation. And Benchmark dataset AudioCaps-Eval, Clotho-Eval.
★ 21pytorchforaudio. Code for the "PyTorch for Audio + Music Processing" series on The Sound of AI YouTube channel.
★ 280soundata. Python library for downloading, loading & working with sound datasets
★ 357dcase_datalist. Collection of DCASE related datasets
★ 18