This is your work, valued

Nashville, TN

Scott H. Hawley

Elite
@drscotthawley

Physics prof, musician, code tinkerer. Specialty: ML + Musical Audio

panotti. A multi-channel neural network audio classifier using Keras

270

ml-audio-start. Suggestions for those interested in developing audio applications of machine learning

215

audio-classifier-keras-cnn. Audio Classifier in Keras using Convolutional Neural Network

159

signaltrain. learning audio effects with neural networks

112

aeiou. (ML) audio engineering i/o utils

55

DLAIE. Materials for Hawley's Deep Learning & AI Ethics course

45

fad_pytorch. Frechet Audio Distance evaluation in PyTorch

36

audio-algebra. alchemy with embeddings

34

SHAART. SHAART is a Python-based audio analysis toolkit, for educational purposes

33

midi-player. Python launcher of animated MIDI player by @cifkao & @magenta

22

vibrary. Vibrary is a GUI client for a user-trainable neural network tool to help producers find audio files on their hard drives

18

vicregaddon. A lightweight and modular parallel PyTorch implementation of VICReg (intended for audio)

13

PolarPatternPlotter. iOS app for measuring sound directivity of loudspeakers and microphones

11

blog_fastpages. Scott H. Hawley's Blog

10

SoundFieldsForever. App Suite for Visualizing Sound in 3D

9

flocoder. A teaching and research repository for exploring generative latent flow matching

8

prefigure. Run configuration management utils: combines configparser, argparse, and wandb.API

6

devblog3. another dev blog attempt

6

fastproaudio. End-to-end audio deep learning with fastai

5

FaceOSC-iOS. Port to iOS of Christopher Baker's FaceOSC for Kyle McDonald's ofxFaceTracker, with output set to be compatible with Rebecca Fiebrink's Wekinator

5

NASH_time_align. Learning to do Time Alignment. Built within the fastproaudio repo, cf. https://drscotthawley.github.io/fastproaudio/time_align.html

4

oplas. Official repository for "Operational Latent Spaces"

4

espiownage. Ownage of ESPI image inference. (Pronounced like "espionage" but with a little "own" in the middle.)

4

blog. my new blog site

4

image-capture-opencv. Python utility for image capture, frame subtraction, (e.g. for ESPI)

3

botograder. An autograder for jupyter notebooks

3

talks. talks I've given

2

audio-diffusion. zach's audio-diffusion work. check various branches (not main!)

2

sample-generator. Tools to train a generative model on arbitrary audio samples

2

SPNet. Object detection for ESPI images of oscillating steelpan drums

2

music-interpolation. Mix between music tracks using machine learning

2

room-shape. 'Deep' neural network learns (boxy) room shape given mode frequencies, or vice versa

2

jukebox. Code for the paper "Jukebox: A Generative Model for Music"

1

fastai2_audio. Audio Module for fastai v2

1

flow-matching-flowers. An exercise in flow-matching modeling of the Oxford Flowers dataset

1

Mask_RCNN_SPNet. Modificaton of "Mask R-CNN for object detection and instance segmentation on Keras and TensorFlow", outputs floats instead of class labels

1

all-in-one. All-In-One Music Structure Analyzer

1

blog_quarto. Quarto version of former fastpages blog

1

descript-audio-codec. Training and inference scripts for the Descript audio codec.

1

RAVE. Bastardized version of Official implementation of the RAVE model: a Realtime Audio Variational autoEncoder

1

diffusers. 🤗 Diffusers: State-of-the-art diffusion models for image and audio generation in PyTorch

1

torchopenl3. openl3 audio embedding for PyTorch

1

LoopGAN. StyleGAN2 + MelGAN Audio Loop Generation

1

audiocraft. Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor / tokenizer, along with MusicGen, a simple and controllable music generation LM with textual and melodic conditioning.

1

M2UGen. This is the official repository for M2UGen

1

TTTT. Trying To Teach Transformers - dump/playground for code

1

audiotools. Object-oriented handling of audio data, with GPU-powered augmentations, and more.

1

mrspuff. A library for Deep Learning education. (deep learning <=> having a school at the bottom of the ocean)

1

midi-model. Midi event transformer for music generation

1

add-menu-popover-demo. Swift: Example of a popover window with a navigation controller and multiple pages for user selection

1