This is your work, valued
UnityOSC. Open Sound Control (OSC) C# classes interface for the Unity3d game engine
★ 504UnityFreesound. Editor tool for Unity3d that allows loading sounds from Freesound.org and assigning them to audio sources
★ 14ArtworkMaker. Code used for the hack made at MIDEM Hack Day in Cannes. The aim was to generate cover artwork by means of music descriptors.
★ 4SingEMAll. Python code for the application SingEMAll, the standalone version. Presented at Music Hack Day Barcelona and Boston editions.
★ 3mhdlondon2014. FFTM
★ 2ADC18. Material for the poster presented at the Audio Developer Conference 2018 (London, UK)
★ 1tinytasks. Simple and single-header tasks library in C++11
★ 1gamevelop-oct2012. Código y scripts demo presentados en la edición de Octubre 2012 de Gamevelop en Madrid. Demo code and scripts presented at Gamevelop, October 2012 in Madrid.
★ 1neutone_sdk. Join the community on Discord for more discussions around Neutone! https://discord.gg/VHSMzb8Wqp
★ 624Fragmenta. A desktop app to fine-tune, generate and perform live music with diffusion models. Runs on MacOS, Windows and Linux.
★ 50zeroclaw. Fast, small, and fully autonomous AI personal assistant infrastructure, any OS, any platform — deploy anywhere, swap anything 🦀
★ 32ksupertonic. Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.
★ 14kKimi-Audio. Kimi-Audio, an open-source audio foundation model excelling in audio understanding, generation, and conversation
★ 4.7kml-intern. 🤗 ml-intern: an open-source ML engineer that reads papers, trains models, and ships ML models
★ 11ksherpa-onnx. Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages
★ 14kpywho. One command to explain your Python environment, trace imports, and detect shadows.
★ 87SneakPeak. Precision waveform editor for REAPER — dockable item editor with spectral analysis, multi-item layering, and real-time metering
★ 17ReaBeat. Neural beat detection and tempo mapping for REAPER — powered by beat-this (ISMIR 2024)
★ 17Woosh. Public release of the Sound Effect Foundation model by Sony AI.
★ 359TonnExamples. Python and Node.js examples of how to work with RoEx's Tonn API
★ 7reaper-reapy-mcp. Reaper and MCP or AI integration A Python application for controlling REAPER Digital Audio Workstation (DAW) using the MCP(Model context protocol).
★ 78liteparse. A fast, helpful, and open-source document parser
★ 12kpypdf. A pure-python PDF library capable of splitting, merging, cropping, and transforming the pages of PDF files
★ 10klocal-deep-researcher. Fully local web research and report writing assistant
★ 9.3kproject-nomad. Project NOMAD is an offline-first knowledge and education server. Wikipedia, thousands of books, courses, maps, and optional local AI, all running on hardware you own with no internet required.
★ 35kreapy-next. Reapy project lives on
★ 16hermes-agent. The agent that grows with you
★ 223kunsloth. Unsloth is a local UI for training and running Kimi K3, Gemma 4, Qwen3.6, DeepSeek, GLM and other models.
★ 69kllmfit. Hundreds of models & providers. One command to find what runs on your hardware.
★ 31kcli. Google Workspace CLI — one command-line tool for Drive, Gmail, Calendar, Sheets, Docs, Chat, Admin, and more. Dynamically built from Google Discovery Service. Includes AI agent skills.
★ 30kvllm. A high-throughput and memory-efficient inference and serving engine for LLMs
★ 88ktadpole-studio. Easy to use AI music generation UI for local models with QoL features
★ 78KittenTTS. State-of-the-art TTS model under 25MB 😻
★ 15ktauri. Build smaller, faster, and more secure desktop and mobile applications with a web frontend.
★ 110kvoicebox. The open-source AI voice studio. Clone, dictate, create.
★ 48kLocalAI. LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.
★ 48kPyMuPDF. PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
★ 10kheartlib. HeartMuLa Official Repo: The Most Powerful Open-Source Music Generation Model of 2026
★ 3.8kHeartMuLa-Studio. Suno-like music generation studio for HeartMuLa/heartlib - AI-powered music creation with reference audio style transfer
★ 618ace-step-ui. 🎵 The Ultimate Open Source Suno Alternative - Professional UI for ACE-Step 1.5 AI Music Generation. Free, local, unlimited. Stop paying for Suno!
★ 4.6kACE-Step-1.5. The most powerful local music generation model that outperforms almost all commercial alternatives, supporting Mac, AMD, Intel, and CUDA devices.
★ 12kpyfmodex. Python bindings for fmod ex sound library.
★ 47daggr. Chain apps and models to build robust AI workflows 🤗
★ 572Kimi-K2. Kimi K2 is the large language model series developed by Moonshot AI team
★ 11kComfyUI-FL-Qwen3TTS. Qwen3-TTS text-to-speech nodes for ComfyUI with voice cloning, voice design, and fine-tuning UI
★ 135runanywhere-sdks. Production ready toolkit to run AI locally
★ 10kopenclaw. Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
★ 385kLuxTTS. A high-quality rapid TTS voice cloning model that reaches speeds of 150x realtime.
★ 4.9kQwen3-TTS. Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voice cloning.
★ 13kpocket-tts. A TTS that fits in your CPU (and pocket)
★ 8kgoose. an open source, extensible AI agent that goes beyond code suggestions - install, execute, edit, and test with any LLM
★ 52kairllm. AirLLM 70B inference with single 4GB GPU
★ 24kfreqlab. Audio plugin creation engine for macOS
★ 22DSPlayground. Terminal environment for rapid prototyping DSP code in C++ 👾
★ 33openflam. OpenFLAM: Framewise Language Audio Model
★ 110openvino-plugins-ai-audacity. A set of AI-enabled effects, generators, and analyzers for Audacity®.
★ 2kjson-render. The Generative UI framework
★ 16kmlx. MLX: An array framework for Apple silicon
★ 28kaudio-flamingo. PyTorch implementation of Audio Flamingo: Series of Advanced Audio Understanding Language Models
★ 1.2kNuitka. Nuitka is a Python compiler written in Python. It's fully compatible with Python 2.6, 2.7, 3.4-3.14. You feed it your Python app, it does a lot of clever things, and spits out an executable or extension module.
★ 15kfastai. The fastai deep learning library
★ 28kmlx-audio. A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
★ 7.7kWwise-MCP. Wwise-MCP is a Model Context Protocol (MCP) server that enables large language models (LLMs) to interact with the Wwise Authoring application. It exposes a set of tools built on a custom Python WAAPI library, allowing MCP clients such as Claude or Cursor to automate and compose complex, multi-step Wwise workflows.
★ 68astro. The web framework for content-driven websites. ⭐️ Star to support our work!
★ 61kn8n. Fair-code workflow automation platform with native AI capabilities. Combine visual building with custom code, self-host or cloud, 400+ integrations.
★ 199kBubbleLab. Open-core workflow engine powering Bubble Lab — and fully runnable, hostable, and extensible on its own.
★ 1.1ksam-audio. The repository provides code for running inference with the Meta Segment Anything Audio Model (SAM-Audio), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 3.6kchatterbox. SoTA open-source TTS
★ 26ksim. Build, deploy, and orchestrate AI agents. Sim is the central intelligence layer for your AI workforce.
★ 29kfoundry-local. C++
★ 2.5kdeepagents. The batteries-included agent harness.
★ 27kagents.md. AGENTS.md — a simple, open format for guiding coding agents
★ 23klibsodium. A modern, portable, easy to use crypto library.
★ 14kHRM. Hierarchical Reasoning Model Official Release
★ 13klangextract. A Python library for extracting structured information from unstructured text using LLMs with precise source grounding and interactive visualization.
★ 38kembedding-atlas. Embedding Atlas is a tool that provides interactive visualizations for large embeddings. It allows you to visualize, cross-filter, and search embeddings and metadata.
★ 4.9kLLM2Fx. Large Language Models for Music Post Production
★ 46anira. an architecture for neural network inference in real-time audio applications
★ 223gemini-cli. An open-source AI agent that brings the power of Gemini directly into your terminal.
★ 106kLoudness. Sony PlayStation Studios' Audio Standards Working Group - Average Loudness and Peak Levels of Audio Content on Sony Interactive Entertainment Platforms
★ 5awesome-mcp-servers. Awesome MCP Servers - A curated list of Model Context Protocol servers
★ 5.7khvcc. The heavy compiler collection for Pure Data patches. Updated to python3 and additional generators
★ 411awesome-audio-dsp. My curated list of audio DSP and plugin development resources (Github fork)
★ 1.4kpython-sdk. The official Python SDK for Model Context Protocol servers and clients
★ 24kvisualblocks. Visual Blocks for ML is a Google visual programming framework that lets you create ML pipelines in a no-code graph editor. You – and your users – can quickly prototype workflows by connecting drag-and-drop ML components, including models, user inputs, processors, and visualizations.
★ 1.4kadk-python. An open-source, code-first Python toolkit for building, evaluating, and deploying sophisticated AI agents with flexibility and control.
★ 21kramalama. RamaLama is an open-source developer tool that simplifies the local serving of AI models from any source and facilitates their use for inference in production, all through the familiar language of containers.
★ 3kollama-python. Ollama Python library
★ 10kollama. Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
★ 177kCavern. Object-based audio engine and codec pack with Dolby Atmos rendering, room correction, HRTF, one-click Unity audio takeover, and much more.
★ 549ableton-mcp. Python
★ 2.8kMMAudio. [CVPR 2025] MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis
★ 2.2kparler-tts. Inference and training library for high-quality TTS models.
★ 5.6kYuE. YuE: Open Full-song Music Generation Foundation Model, something similar to Suno.ai but open
★ 6.4kANUS. TypeScript
★ 6.5knoisecraft. Browser-based visual programming language and platform for sound synthesis.
★ 1.2ksignalflow. A sound synthesis framework for Python, designed for clear and concise expression of musical ideas
★ 254claude-code. Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflows - all through natural language commands.
★ 140kvisage. C++ UI library meets creative coding
★ 605unDAW. Editor and Runtime modules that aim to add some Digital Audio Workstation capabilities to Unreal Engine 5, requires Unreal Editor 5.4
★ 58uv. An extremely fast Python package and project manager, written in Rust.
★ 88kpywwise. PyWwise is an open-source Python wrapper around the Wwise Authoring API (WAAPI). PyWwise is meant to make Wwise scripting (e.g. automation, data validation, etc.) much easier to achieve.
★ 55awesome-juce. A curated list of JUCE modules, templates, plugins, oh my!
★ 1.3kdemucs. Code for the paper Hybrid Spectrogram and Waveform Source Separation
★ 3kdemucs. Code for the paper Hybrid Spectrogram and Waveform Source Separation
★ 10kfish-speech. SOTA Open Source TTS
★ 32kgigi. A framework for rapid prototyping and development of real-time rendering techniques.
★ 1.2kpiper. A fast, local neural text to speech system
★ 11kbenny. a live music environment
★ 212pyaudiodsptools. Numpy Audio DSP Tools
★ 215soundata. Python library for downloading, loading & working with sound datasets
★ 357TTS. 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
★ 46kImHex. 🔍 A Hex Editor for Reverse Engineers, Programmers and people who value their retinas when working at 3 AM.
★ 54ktyper. Typer, build great CLIs. Easy to code. Based on Python type hints.
★ 20krez. An integrated package configuration, build and deployment system for software
★ 1.1kmatplotlib. matplotlib: plotting with Python
★ 23kJUCE. JUCE is an open-source cross-platform C++ application framework for desktop and mobile applications, including VST, VST3, AU, AUv3, LV2 and AAX audio plug-ins.
★ 8.7kmetavoice-src. Foundational model for human-like, expressive TTS
★ 4.2kwaapi-python-tools. Wwise tools using WAAPI and Python
★ 53python-osc. Open Sound Control server and client in pure python
★ 580speechbrain. A PyTorch-based Speech Toolkit
★ 12kunreal-audio-dsp-template-UE5. A simple Unreal 5 plugin with demo content that shows how audio DSP can be implemented as Metasound, SourceEffect and SubmixEffect. The intention is to use it as a template to create more complex audio DSP effects.
★ 48steam-audio. Steam Audio
★ 2.9kprofessional-programming. A collection of learning resources for curious software engineers
★ 51kstable-audio-tools. Generative models for conditional audio generation
★ 3.8kLabSound. :microscope: :speaker: graph-based audio engine
★ 786gradio. Build and share delightful machine learning apps, all in Python. 🌟 Star to support our work!
★ 43kEmotiVoice. EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine
★ 8.5kVALL-E-X. An open source implementation of Microsoft's VALL-E X zero-shot TTS model. Demo is available in https://plachtaa.github.io/vallex/
★ 7.9ktortoise-tts. A multi-voice TTS system trained with an emphasis on quality
★ 15knih-plug. Rust VST3 and CLAP plugin framework and plugins - because everything is better when you do it yourself
★ 2.9kpyminiaudio. python interface to the miniaudio audio playback, recording, decoding and conversion library
★ 185Ryven. Flow-based visual scripting for Python
★ 4.1kPyFlow. Visual scripting framework for python
★ 3.2kComfyUI. The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
★ 123ksnoop. A powerful set of Python debugging tools, based on PySnooper
★ 1.5ksurge-python. This repo contains examples of how to use surgepy, Python bindings for the Surge synthesizer.
★ 35mustango. Mustango: Toward Controllable Text-to-Music Generation
★ 394hifi-gan. HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis
★ 2.4kspaCy. 💫 Industrial-strength Natural Language Processing (NLP) in Python
★ 34kjust_playback. A small library for playing audio files in python, with essential playback functionality.
★ 99nendo. The Nendo AI Audio Tool Suite
★ 219imgui_bundle. Interactive Python & C++ apps for desktop, mobile, and web - powered by Dear ImGui. Stop fighting GUI frameworks. Start building.
★ 1.3kimgui-node-editor. Node Editor built using Dear ImGui
★ 4.5kinsanely-fast-whisper. Jupyter Notebook
★ 13knn-inference-template. Neural network inference template for real-time cricital audio environments - presented at ADC23
★ 134sointu. Fork of 4klang that can target 386, amd64 and WebAssembly. Tools run on Windows, Mac & Linux
★ 350pydub. Manipulate audio with a simple and easy high level interface
★ 9.8kpydantic. Data validation using Python type hints
★ 28kStartup-CTO-Handbook. The Startup CTO's Handbook, a book covering leadership, management and technical topics for leaders of software engineering teams
★ 14kdawproject. Open exchange format for DAWs
★ 1kdocstring_parser. Parse Python docstrings in various flavors.
★ 270syntheon. Parameter inference of music synthesizers to simplify sound design process. Supports Vital.
★ 178pydags. Simple, lightweight, extensible DAG framework for Python with a Kubeflow-like API
★ 87paradag. A robust DAG implementation for parallel execution
★ 71SimpleDataFlow. A simple tool demonstrating the DPG node editor.
★ 24dagster. An orchestration platform for the development, production, and observation of data assets.
★ 16ktuning_playbook. A playbook for systematically maximizing the performance of deep learning models.
★ 30kstatic-analysis. ⚙️ A curated list of static analysis (SAST) tools and linters for all programming languages, config files, build tools, and more. The focus is on tools which improve code quality.
★ 15kmagenta. Magenta: Music and Art Generation with Machine Intelligence
★ 20kaudioFlux. A library for audio and music analysis, feature extraction.
★ 3.3kwhisper. Robust Speech Recognition via Large-Scale Weak Supervision
★ 106kbark. 🔊 Text-Prompted Generative Audio Model
★ 39kspleeter. Deezer source separation library including pretrained models.
★ 28knoisebandnet. Code for the "NoiseBandNet: Controllable Time-Varying Neural Synthesis of Sound Effects Using Filterbanks" paper.
★ 39awesome-pipeline. A curated list of awesome pipeline toolkits inspired by Awesome Sysadmin
★ 6.6kelementary. Elementary is a JavaScript library for digital audio signal processing.
★ 507mypy. Optional static typing for Python
★ 21kDearPyGui_Animate. DearPyGui_Animate is an add-on to bring DearPyGUI to life.
★ 71DearPyGui_Ext. Dear PyGui Extensions: A collection of useful tools, abstractions, and simplification layers built with/for Dear PyGui users.
★ 107DearPyGui-Examples. Repo for more advanced examples
★ 203DearPyGui. Dear PyGui: A fast and powerful Graphical User Interface Toolkit for Python with minimal dependencies
★ 16kpolymath. Convert any music library into a music production sample-library with ML
★ 1.6kmultimodal_emotion_recognition. Jupyter Notebook
★ 3einops. Flexible and powerful tensor operations for readable and reliable code (for pytorch, jax, TF and others)
★ 9.6kaudio-ai-timeline. A timeline of the latest AI models for audio generation, starting in 2023!
★ 1.9ktorchemotion. Emotion recognition library for PyTorch
★ 22ai-audio-startups. Community list of startups working with AI in audio and music technology
★ 1.8kaudio-diffusion-pytorch. Audio generation using diffusion models, in PyTorch.
★ 2.1ksample-generator. Tools to train a generative model on arbitrary audio samples
★ 1.1ktinygrad. You like pytorch? You like micrograd? You love tinygrad! ❤️
★ 33kmicrograd. A tiny scalar-valued autograd engine and a neural net library on top of it with PyTorch-like API
★ 17klibadm. Audio Definition Model (ITU-R BS.2076) handling library
★ 50engine-sim. Combustion engine simulator that generates realistic audio.
★ 9.5kml-audio-start. Suggestions for those interested in developing audio applications of machine learning
★ 215rpp. Read and write Reaper RPP files with Python.
★ 79reathon. Construct REAPER projects programatically in python.
★ 85Awesome-Diffusion-Models. A collection of resources and papers on Diffusion Models
★ 12kJSFX. A bundle of JSFX and scripts for reaper.
★ 625madronalib. Madronalib: a C++ framework for DSP applications.
★ 335tixl. TiXL is an open source software to create realtime motion graphics.
★ 5klibgammatone. gammatone auditory filterbanks in C++
★ 6clap. Audio Plugin API
★ 2.3kLibXtract. LibXtract is a simple, portable, lightweight library of audio feature extraction functions.
★ 231audiomentations. A Python library for audio data augmentation. Useful for making audio ML models work well in the real world, not just in the lab.
★ 2.3kHowToBeAProgrammer. A guide on how to be a Programmer - originally published by Robert L Read
★ 16kkira. Library for expressive game audio.
★ 1.1kDawDreamer. Digital Audio Workstation with Python; VST instruments/effects, parameter automation, FAUST, JAX, Warp Markers, and JUCE processors
★ 1.3kpedalboard. 🎛 🔊 A Python library for audio.
★ 6.2klmms. Cross-platform music production software
★ 10kimgui. Dear ImGui: Bloat-free Graphical User interface for C++ with minimal dependencies
★ 75kWaveFunctionCollapse. Bitmap & tilemap generation from a single example with the help of ideas from quantum mechanics
★ 25kvital. Spectral warping wavetable synth
★ 2.1kentt. Gaming meets modern C++ - a fast and reliable entity component system (ECS) and much more
★ 13kUnityAbletonLink. An Ableton Link plugin for Unity
★ 101dotween. A Unity C# animation engine. HOTween v2
★ 2.7kxNode. Unity Node Editor: Lets you view and edit node graphs inside Unity
★ 3.7kconcurrentqueue. A fast multi-producer, multi-consumer lock-free concurrent queue for C++11
★ 12k