This is your work, valued
goodhosts. Simple hosts file management in Golang (deprecated).
★ 67notes. A collection of notes using the Zettelkasten methodology.
★ 63PyWit. Python bindings for the Wit HTTP API (unofficial)
★ 24obsidian-title-as-link-text. An Obsidian plugin to set the Link Text using the document title
★ 24bfs-php. Breadth-First Search implementation in PHP
★ 15txes2. A Twisted ElasticSearch client loosely based on PyES
★ 11ansible-linode. A Linode module for Ansible (deprecated).
★ 10magicranker. An open-source stock ranking and filtering tool based on the Magic Formula.
★ 8idle. Cross-platform idle time detection in Go (golang).
★ 8consul-formula. Salt formula for installing and configuring Consul
★ 6plugin.video.joeroganexperience. Watch the Joe Rogan Experience video podcast on XBMC
★ 3plugin.video.udacity. Complete Udacity courses in XBMC
★ 2hubot-test-examples. Some example Hubot scripts with unit tests
★ 2igor. XBMC and Wit.ai voice recognition experiment
★ 2blog-archive. Static version of my blog powered by Pelican
★ 2aws-cloudwatch. AWS Lambda for logging into LogDNA
★ 1fastai-notes. Some additional notes for Fast.ai courses.
★ 1pelican-jupyter. Pelican plugin for blogging with Jupyter/IPython Notebooks
★ 1rotest. A tiny unit test framework for Roblox
★ 1bengaliai-cv19. Bengaliai CV19
★ 1simple-ranker. An extremely simple ranking module for pandas dataframes
★ 1txstripe. Stripe Twisted (Python) bindings
★ 1movie-collection-organiser. A simple script to organise my movie collection for XBMC
★ 1node-webkit-hipster-seed-decaf. node-webkit-hipster-seed without Jade and CoffeeScript
★ 1GiantDwarf. A simple Campfire bot written in Python
★ 1dotfiles. My ever changing Dot Files including Bash, Screen and Vim
★ 1reddit-stealer. Download every video in Subreddits and convert them to audio.
★ 1MrHappy. An IRC Bot (with some support for Campfire - see the campfire branch)
★ 1plugin.video.rsa. Watch videos from RSA (Royal Society for the encouragement of Arts, Manufactures and Commerce) on XBMC
★ 1aws-monitor. AWS Cloudwatch Monitor script
★ 1ansible-jenkins. Ansible playbook: Jenkins
★ 1python-youtube-download. A simple, yet versatile Python package for downloading YouTube videos.
★ 1deltafin. Run full Kimi K3 on a single device. And an OpenAI-compatible API server for local chat and coding agents.
★ 396yt-dlp. A feature-rich command-line audio/video downloader
★ 181kFableCut. Zero-dependency browser video editor that AI agents can drive — JSON timeline, MCP + REST, live-reloading UI
★ 561pxpipe. cut Claude Code token usage by rendering text context as images
★ 6.8khyperframes. Write HTML. Render video. Built for agents.
★ 39kvideo-use. Edit videos with coding agents
★ 18kOpenMontage. World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files. Turn your AI coding assistant into a full video production studio.
★ 44kmarimo. A reactive notebook for Python — run reproducible experiments, query with SQL, execute as a script, deploy as an app, and version with git. Stored as pure Python. All in a modern, AI-native editor.
★ 22kmlx-demucs. A port of Demucs to Apple's MLX framework for efficient audio source separation on Apple Silicon.
★ 4mlx-uniface. MLX-UniFace: Blazing-fast face analysis on Apple Silicon with native MLX backend. Face detection, recognition, landmarks, age/gender/emotion.
★ 1hermes-agent. The agent that grows with you
★ 223kqmd. mini cli search engine for your docs, knowledge bases, meeting notes, whatever. Tracking current sota approaches while being all local
★ 28kobsidian-charted-roots. Professional genealogical tools for Obsidian. Interactive family charts, geographic maps, PDF reports, GEDCOM/Gramps import-export, evidence tracking, and fictional calendar support. Built for genealogists, historians, and worldbuilders.
★ 147mastodon-comments. Web component to show comments from mastodon and bluesky
★ 51bluesky-comments-tag. JavaScript
★ 96mlx-drifting-models. MLX implementation of Generative Modeling via Drifting (Deng et al., 2026)
★ 2obsidian-markdown-notebook. Jupyter-style executable notebooks in plain Markdown: run code directly in Obsidian, with outputs stored in the .md file itself.
★ 4litrepl. Litrepl is a command-line tool and Vim plugin for evaluating code sections within Markdown or LaTeX documents.
★ 8obsidian-jupyter. Edit .ipynb Jupyter files directly in Obsidian.
★ 94graphify. Turn any codebase, with its docs, SQL schemas, configs, and PDFs, into a queryable knowledge graph. A /graphify skill for Claude Code, Cursor, Codex, and Gemini CLI: local deterministic AST parsing, every edge explained, no vector store.
★ 99kpocketsmith-skill. Manage PocketSmith transactions, categories, and financial data in OpenClaw/Claude.
★ 1bluesky-skill. Bluesky CLI skill for Clawdbot - bird-like interface for AT Protocol
★ 3canva-skill. 🎨 Canva integration skill for Clawdbot/Moltbot. Create, export, and manage designs via the Connect API.
★ 3awesome-openclaw-skills. The awesome collection of OpenClaw skills. 5,400+ skills filtered and categorized from the official OpenClaw Skills Registry.🦞
★ 52kLTX-2. Official Python inference and LoRA trainer package for the LTX-2 audio–video generative model.
★ 8.5kApplio. A simple, high-quality voice conversion tool focused on ease of use and performance.
★ 3.5kend2end-all-conv. Deep Learning to Improve Breast Cancer Detection on Screening Mammography
★ 391LLM-judge-reporting. A simple plug-in framework that corrects bias and computes confidence intervals in reporting LLM-as-a-judge evaluation, and an adaptive algorithm that efficiently allocates calibration samples to reduce uncertainty in estimates.
★ 81google-cloud-java. Google Cloud Client Library for Java
★ 2.1kOvi. Python
★ 1.7kmlx-image. mlx image models for Apple Silicon machines
★ 100echomimic. [AAAI 2025] EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditioning
★ 4.3kfloat. [ICCV 2025] Official Pytorch Implementation of FLOAT: Generative Motion Latent Flow Matching for Audio-driven Talking Portrait.
★ 487Sonic. Official implementation of "Sonic: Shifting Focus to Global Audio Perception in Portrait Animation"
★ 3.3kddsm_tools. C
★ 32pal-mcp-server. The power of Claude Code / GeminiCLI / CodexCLI + [Gemini / OpenAI / OpenRouter / Azure / Grok / Ollama / Custom Model / All Of The Above] working as one.
★ 12kIkemen-GO. An open-source fighting game engine that supports MUGEN resources.
★ 1.5kWan2.1. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kWan2.1-Mac. Wan2.1 for Mac.
★ 80chatterbox. SoTA open-source TTS
★ 26kdia. A TTS model capable of generating ultra-realistic dialogue in one pass.
★ 19kDI-engine. OpenDILab Decision AI Engine. The Most Comprehensive Reinforcement Learning Framework B.P.
★ 3.6kACE-Step. ACE-Step: A Step Towards Music Generation Foundation Model
★ 4.7kFramePack. Lets make video diffusion practical!
★ 17kwhisperX. WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
★ 23kDiffRhythm. Di♪♪Rhythm: Blazingly Fast and Embarrassingly Simple End-to-End Full-Length Song Generation with Latent Diffusion
★ 2.3kevals. Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.
★ 19kcsm. A Conversational Speech Generation Model
★ 15kwhere-is-satoshi. Stylometric analysis of Satoshi & comparison of Satoshi with 75,000+ authors
★ 44BrainPress. BrainPress is a simple NextJS app to self-publish Obsidian vaults. It supports the new canvas files.
★ 68re-arc. Reverse Engineering the Abstraction and Reasoning Corpus
★ 355obsidian-image-converter. ⚡️ Convert, compress, resize, annotate, markup, draw, crop, rotate, flip, align images directly in Obsidian. Drag-resize, rename with variables, batch process. WEBP, JPG, PNG, HEIC, TIF.
★ 807tiny-gpu. A minimal GPU design in Verilog to learn how GPUs work from the ground up
★ 13kcaptacity. Add Automatic Captions to YouTube Shorts with AI
★ 138neuralnoise. The AI Podcast Studio: generate podcasts scripts and their audio version with a team of AI workers in a Podcast Studio 🎙️📜
★ 227ShortGPT. 🚀🎬 ShortGPT - Experimental AI framework for youtube shorts / tiktok channel automation
★ 7.7kaider. aider is AI pair programming in your terminal
★ 48ktotalmix-volume-control. Provides control over RME TotalMix master volume via OSC.
★ 44YuE. YuE: Open Full-song Music Generation Foundation Model, something similar to Suno.ai but open
★ 6.4kLatentSync. Taming Stable Diffusion for Lip Sync!
★ 5.9kFabric. Fabric is an open-source framework for augmenting humans using AI. It provides a modular system for solving specific problems using a crowdsourced set of AI prompts that can be used anywhere.
★ 43kdocling. Get your documents ready for gen AI
★ 64kspacy-layout. 📚 Process PDFs, Word documents and more with spaCy
★ 910uv. An extremely fast Python package and project manager, written in Rust.
★ 88kopen-webui. User-friendly AI Interface (Supports Ollama, OpenAI API, ...)
★ 147klucide. Beautiful & consistent icon toolkit made by the community. Open-source project and a fork of Feather Icons.
★ 24kminchin.pelican.readers.commonmark. CommonMark reader for Pelican (via Markdown-IT-Py)
★ 1pelican-obsidian. Makes it possible to bridge work in obsidian to pelican seamlessly
★ 45markdown-it-py. Markdown parser, done right. 100% CommonMark support, extensions, syntax plugins & high speed. Now in Python!
★ 1.3kmarkdown-callouts. Markdown extension: a classier syntax for admonitions
★ 45mkdocs-publisher. Publisher for MkDocs - a set of plugins for content creators
★ 128obsidian-quiz-generator. Generate interactive flashcards from your notes using models from OpenAI (ChatGPT), Google (Gemini), Ollama (local LLMs), and more. Or manually create your own to use with the quiz UI.
★ 175obsidian-markmind. A mind map, outline for obsidian,It support mobile and desktop
★ 1kMetal-Puzzles. Solve Puzzles. Learn Metal 🤘
★ 615GPU-Puzzles. Solve puzzles. Learn CUDA.
★ 12kmarkdownlint. Markdown lint tool
★ 2.1kThe-Little-Book-of-ML-Metrics. The book every data scientist needs on their desk.
★ 998unsloth. Unsloth is a local UI for training and running Kimi K3, Gemma 4, Qwen3.6, DeepSeek, GLM and other models.
★ 69kobsidian-copilot. THE Copilot in Obsidian
★ 7.5kWeekly-Top-LLM-Papers. Curated list of weekly published LLM papers
★ 202Weekly-Top-Computer-Vision-Papers. A curated list of important published computer vision paper on a weekly basis
★ 178open-notebooklm. Convert any PDF into a podcast episode!
★ 2.6kdualdiffusion. Dual Diffusion is a generative diffusion model for music trained on video game soundtracks.
★ 93Podcast. Python
★ 260anthology-of-modern-ml. Collection of important articles to be treated as a textbook
★ 879podcastfy. An Open Source Python alternative to NotebookLM's podcast feature: Transforming Multimodal Content into Captivating Multilingual Audio Conversations with GenAI
★ 6.5kGeneFace. GeneFace: Generalized and High-Fidelity 3D Talking Face Synthesis; ICLR 2023; Official code
★ 2.7kLLMAgentPapers. Must-read Papers on LLM Agents.
★ 3.1kperpetual. Perpetual is a high-performance gradient boosting machine. It delivers optimal accuracy in a single run without complex tuning through a simple budget parameter. It features out-of-the-box support for causal ML, continual learning, native calibration, and robust drift monitoring, along with Rust core and zero-copy bindings for Python and R
★ 703llm-table-survey. Resources on Large Language Models for Table Processing
★ 111weave. Weave is a toolkit for developing AI-powered applications, built by Weights & Biases.
★ 1.1kminitorch. The full minitorch student suite.
★ 2.4kMusic-Representation-Comparison. This is the repo with the code to conduct a comparative analysis of different audio representation models.
★ 11pyfluidsynth. Python bindings for FluidSynth
★ 255agc. Audiogen Codec
★ 146MidiTok. MIDI / symbolic music tokenizers for Deep Learning models 🎶
★ 884whisper-vits-svc. Core Engine of Singing Voice Conversion & Singing Voice Clone
★ 2.9kMusicTI_AAAI2024. " Music Style Transfer with Time-Varying Inversion of Diffusion Models"
★ 59FunCodec. FunCodec is a research-oriented toolkit for audio quantization and downstream applications, such as text-to-speech synthesis, music generation et.al.
★ 445RepCodec. Models and code for RepCodec: A Speech Representation Codec for Speech Tokenization
★ 196UniAudio. The Open Source Code of UniAudio
★ 605terminalizer. 🦄 Record your terminal and generate animated gif images or share a web player
★ 16kOMG. [ECCV 2024] OMG: Occlusion-friendly Personalized Multi-concept Generation In Diffusion Models
★ 701flash-attention. Fast and memory-efficient exact attention
★ 25kbeartype. Unbearably fast near-real-time pure-Python runtime-static type-checker.
★ 3.5kseamless_communication. Foundational Models for State-of-the-Art Speech and Text Translation
★ 12kobsidian-pseudocode. An obsidian plugin that helps to render a LaTeX-style pseudocode inside a code block.
★ 135WhisperSpeech. An Open Source text-to-speech system built by inverting Whisper.
★ 4.6klibriheavy. Libriheavy: a 50,000 hours ASR corpus with punctuation casing and context
★ 220icefall. Python
★ 1.5kvocos. Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis
★ 1.1kaudio-ai-timeline. A timeline of the latest AI models for audio generation, starting in 2023!
★ 1.9kgenmusic_demo_list. a list of demo websites for automatic music generation research
★ 794CoMoSVC. CoMoSVC: One-Step Consistency Model Based Singing Voice Conversion & Singing Voice Clone
★ 148identifying-rhyming-words. Given a target word and a set of words, find the word which best rhymes with the target
★ 9RAVE. Official implementation of the RAVE model: a Realtime Audio Variational autoEncoder
★ 1.8kSVSELM. Python
★ 3VI-SVS. Singing Voice Synthesis based on VITS, different from VISinger
★ 198resemble-enhance. AI powered speech denoising and enhancement
★ 2.4kFullSubNet-plus. The official PyTorch implementation of "FullSubNet+: Channel Attention FullSubNet with Complex Spectrograms for Speech Enhancement".
★ 293voicefixer. General Speech Restoration
★ 1.4kkeytotext. Keywords to Sentences
★ 451stable-audio-tools. Generative models for conditional audio generation
★ 3.8kaugtxt. yet another text augmentation python package
★ 2typo. A python package to simulate typographical errors.
★ 40audiotools. Object-oriented handling of audio data, with GPU-powered augmentations, and more.
★ 350CALM. CALM (Contrastive Alignment of Language and Music) aligns songs with their natural language description using CLIP (Project for the 1st Sound of AI Hackathon).
★ 8memray. Memray is a memory profiler for Python
★ 15kpesto. Self-supervised learning for real-time pitch estimation
★ 297Wav2Lip. This repository contains the codes of "A Lip Sync Expert Is All You Need for Speech to Lip Generation In the Wild", published at ACM Multimedia 2020. For HD commercial model, please try out Sync Labs
★ 13kinsightface. State-of-the-art 2D and 3D Face Analysis Project
★ 29kdawproject. Open exchange format for DAWs
★ 1kRemFx. General Purpose Audio Effect Removal
★ 116all-in-one. All-In-One Music Structure Analyzer
★ 808syntheon. Parameter inference of music synthesizers to simplify sound design process. Supports Vital.
★ 178musicnn. Pronounced as "musician", musicnn is a set of pre-trained deep convolutional neural networks for music audio tagging.
★ 712python-soxr. Fast and high quality sample-rate conversion library for Python
★ 109AnimateDiff. Official implementation of AnimateDiff.
★ 12kcog-deforum-stable-diffusion. Python
★ 16MVSEP-CDX23-Cinematic-Sound-Demixing. Model for CDX23 (Cinematic Sound Demixing) contest
★ 57MVSEP-MDX23-music-separation-model. Model for MDX23 music separation contest
★ 852sdx23-aimless. Source Separation training codebase for the Sound Demixing Challenge 2023.
★ 45outlines. Structured Outputs
★ 15kFooocus. Focus on prompting and generating
★ 52kstaff. Music theory and score rendering library with midi, notes, chords, scales, and more.
★ 275multires-conv. Sequence Modeling with Multiresolution Convolutional Memory (ICML 2023)
★ 127descript-audio-codec. State-of-the-art audio codec with 90x compression factor. Supports 44.1kHz, 24kHz, and 16kHz mono/stereo audio.
★ 1.8kbinary-assets. Store all binary files such as PDFs from UoL into this repository
★ 104golf. A DDSP-based neural voice synthesiser.
★ 135autolabel. Label, clean and enrich text datasets with LLMs.
★ 2.3kcrepe. CREPE: A Convolutional REpresentation for Pitch Estimation -- pre-trained model (ICASSP 2018)
★ 1.4ktorchcrepe. Pytorch implementation of the CREPE pitch tracker
★ 523MERT. Official implementation of the paper "Acoustic Music Understanding Model with Large-Scale Self-supervised Training".
★ 482SadTalker. [CVPR 2023] SadTalker:Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face Animation
★ 14kfirst-order-model. This repository contains the source code for the paper First Order Motion Model for Image Animation
★ 15kScyclone. Real-time Neural Timbre Transfer
★ 442infinite-zoom-stable-diffusion. resources for creating Ininite zoom video using Stable Diffiusion, you can use multiple prompts and it is easy to use.
★ 88stable-diffusion-webui. Stable Diffusion web UI
★ 164ksvoice. We provide a PyTorch implementation of the paper Voice Separation with an Unknown Number of Multiple Speakers In which, we present a new method for separating a mixed audio sequence, in which multiple voices speak simultaneously. The new method employs gated neural networks that are trained to separate the voices at multiple processing steps, while maintaining the speaker in each output channel fixed. A different model is trained for every number of possible speakers, and the model with the largest number of speakers is employed to select the actual number of speakers in a given sample. Our method greatly outperforms the current state of the art, which, as we show, is not competitive for more than two speakers.
★ 1.3kshap-e. Generate 3D objects conditioned on text or images
★ 12kpyannote-audio. Neural building blocks for speaker diarization: speech activity detection, speaker change detection, overlapped speech detection, speaker embedding
★ 10kwavelets-ext. A re-implementation of the Wavelets package using Cython to improve the speed.
★ 13rvc-webui. liujing04/Retrieval-based-Voice-Conversion-WebUI reconstruction project
★ 522Voice_Separation_and_Selection. Separation voice and delete files with majority silence
★ 52potassium. An HTTP serving framework by Banana
★ 103so-vits-svc-fork. so-vits-svc fork with realtime support, improved interface and more features.
★ 9.3kpyo3. Rust bindings for the Python interpreter
★ 16koutotune. An opensource harmonizer implementation leveraging the DISTRHO Plugin Framework.
★ 81de-ess. De-essing software to reduce sibilance in speech
★ 24Constrained-Text-Generation-Studio. Code repo for "Most Language Models can be Poets too: An AI Writing Assistant and Constrained Text Generation Studio" at the (CAI2) workshop, jointly held at (COLING 2022)
★ 217textgen. Open-source desktop app for local LLMs. Text, vision, tool-calling, OpenAI/Anthropic-compatible API. 100% private.
★ 48kobsidian-front-matter-title. Plugin for Obsidian.md
★ 367tuning_playbook. A playbook for systematically maximizing the performance of deep learning models.
★ 30kautoscraper. A Smart, Automatic, Fast and Lightweight Web Scraper for Python
★ 7.8kqdrant. Qdrant - High-performance, massive-scale Vector Database and Vector Search Engine for the next generation of AI. Also available in the cloud https://cloud.qdrant.io/
★ 34ktiktoken. tiktoken is a fast BPE tokeniser for use with OpenAI's models.
★ 19kautoxgb. XGBoost + Optuna
★ 727YaLM-100B. Pretrained language model with 100B parameters
★ 3.8kvintage-online-store. An early online store I wrote in '94/95 in all it's vintage glory
★ 6strapi. 🚀 Strapi is the leading open-source headless CMS. It’s 100% JavaScript/TypeScript, fully customizable, and developer-first.
★ 73kdiff2html-cli. Pretty diff to html javascript cli (diff2html-cli)
★ 598ipyvizzu. Build animated charts in Jupyter Notebook and similar environments with a simple Python syntax.
★ 972TRACER. TRACER: Extreme Attention Guided Salient Object Tracing Network (AAAI 2022) implementation in PyTorch
★ 200Swin-Transformer-TF. Tensorflow implementation of Swin Transformer model.
★ 216Weighted-Boxes-Fusion. Set of methods to ensemble boxes from different object detection models, including implementation of "Weighted boxes fusion (WBF)" method.
★ 1.8kimgsize. Python library to get the type and size of an image quickly
★ 7tensorflow-image-models. TensorFlow port of PyTorch Image Models (timm) - image models with pretrained weights.
★ 290TokenCut. (CVPR 2022) Pytorch implementation of "Self-supervised transformers for unsupervised object discovery using normalized cut"
★ 339tf-image. TensorFlow2+ graph image augmentation library optimized for tf.data.Dataset.
★ 25Humpback-Whale-Identification. Humpback Whale Identification
★ 253long-text-token-classification. Python
★ 160manim. A community-maintained Python framework for creating mathematical animations.
★ 40kPlotNeuralNet. Latex code for making neural networks diagrams
★ 25kstylegan2-ada. StyleGAN2 with adaptive discriminator augmentation (ADA) - Official TensorFlow implementation
★ 123kubectx. Faster way to switch between clusters and namespaces in kubectl
★ 20kplaycanvas-sync. Real-time synchronization of files between PlayCanvas and your local machine
★ 96GA-SDK-JAVASCRIPT. Official repository for GameAnalytics JavaScript SDK
★ 23asteroid. The PyTorch-based audio source separation toolkit for researchers
★ 2.6kplaycanvas-rest-api-tools. A set of tools to use with the PlayCanvas REST API for common jobs such as downloading a build and archiving a project
★ 30osumapper. An automatic beatmap generator using Tensorflow / Deep Learning.
★ 461