This is your work, valued
koi. A plug-in for Krita that enables the use of AI models for img2img generation.
★ 441robo-diffusion. make cool robot concept art
★ 104dream-bench. A tool for benchmarking image generation models.
★ 32DALLE2-pytorch. Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis neural network, in Pytorch
★ 6mezo_alpaca. Train Alpaca-LLaMA with the zero-order MeZO optimizer
★ 5dis-entangle. use DIS to remove the backgrounds from / generate masks for a folder of images
★ 4k-samplers. a minimal export of the k-diffusion samplers
★ 3prior. prior models for embedding translation
★ 3maestro-class. CS 175: Team Gibbon
★ 1StableLM. StableLM: Stability AI Language Models
★ 1audio-trainer-pytorch. train audio models with ease
★ 1h14_nsfw_detector. clip-based nsfw detector
★ 1clip-retrieval. Easily compute clip embeddings and build a clip retrieval system with them
★ 1ponytail. Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.
★ 92kharness-1. 🚀 Ultra Recipe for Training Long-Horizon Search Agents - matching frontier AI's search capability with a 20B model + stateful harness
★ 892rlm. General plug-and-play inference library for Recursive Language Models (RLMs), supporting various sandboxes.
★ 5.3kHOPEJr. HOPEJr_open-source_DIY_Humanoid_Robot_with_dexterous_hands
★ 812ouroboros. Agent OS: Stop prompting. Start specifying.
★ 5.2kcamofox-browser. Stealth headless browser for AI agents — bypass Cloudflare, bot detection, and anti-scraping. Drop-in Puppeteer/Playwright replacement.
★ 8.2kendless-terminals. Python
★ 135loophole. Python
★ 410llama.cpp. LLM inference in C/C++
★ 122kle-wm. Official code base for LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
★ 4.2kautonovel. An autonomous novel writing pipeline, by Hermes Agent
★ 1.4kstop-slop. A skill file for removing AI tells from prose
★ 15ktlaplus. TLC is a model checker for specifications written in TLA+. The TLA+Toolbox is an IDE for TLA+.
★ 3kpinchtab. High-performance browser automation bridge and multi-instance orchestrator with advanced stealth injection and real-time dashboard.
★ 9.8kfirecrawl. The API to search, scrape, and interact with the web at scale. 🔥
★ 158kECC. The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
★ 236ktenacitOS. OpenClaw Mission Control Dashboard
★ 1.2kdesloppify. Agent harness to make your slop code well-engineered and beautiful.
★ 3kcraftium. A framework for creating rich, 3D, Minecraft-like single and multi-agent environments for AI research. (Accepted at ICML 2025).
★ 206claude-subconscious. Give Claude Code a subconscious
★ 2.9khermes-agent. The agent that grows with you
★ 223kpicoclaw. Tiny, Fast, and Deployable anywhere — automate the mundane, unleash your creativity
★ 30kopenclaw. Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
★ 385kopenevolve. Open-source implementation of AlphaEvolve
★ 6.8kmusic-audio-representations. Results and Models for Learning Audio Representations of Music Content
★ 107node-autoit-koffi. This Node.js module provides support for all AutoIt functions, allowing users to automate Windows GUI tasks seamlessly.
★ 16mcp-windows-desktop-automation. A Model Context Protocol (MCP) server for Windows desktop automation using AutoIt.
★ 114copyparty. Portable file server with accelerated resumable uploads, dedup, WebDAV, SFTP, FTP, TFTP, zeroconf, media indexer, thumbnails++ all in one file
★ 46kneko. A self hosted virtual browser that runs in docker and uses WebRTC.
★ 22kvaultwarden. Unofficial Bitwarden compatible server written in Rust, formerly known as bitwarden_rs
★ 65katropos. Atropos is a Language Model Reinforcement Learning Environments framework for collecting and evaluating LLM trajectories through diverse environments
★ 1.3kkontext-lab. HTML
★ 1lerobotdepot. LeRobotDepot is a community-driven repository listing open-source hardware, components, and 3D-printable projects compatible with the LeRobot library. It helps users easily discover, build, and contribute to affordable, accessible robotics solutions powered by state-of-the-art AI.
★ 229CM6_COBOT_ROBOT. Files and intro page for first version of CM6 COBOT robotic arm
★ 231setta. Streamline Python coding, configuration, UI creation, and onboarding.
★ 18sunnypilot. sunnypilot is an open source driver assistance system. sunnypilot offers the user a unique driving experience for over 350 supported car makes and models with modified behaviors of driving assist engagements. sunnypilot complies with the safety policy from comma.ai's openpilot as accurately as possible.
★ 2kopendbc. a Python API for your car
★ 3.3kcomfyui_dygen. Dy Mokomi Custom Nodes for ComfyUI
★ 1Core-R-Theta-4-Axis-Printer. 4 Axis 3D Printer
★ 902Radial_Non_Planar_Slicer. G-code
★ 327fsdp_optimizers. supporting pytorch FSDP for optimizers
★ 84SPiT. A Spitting Image: Modular Superpixel Tokenization in Vision Transformers
★ 23tt-metal. :metal: TT-NN operator library, and TT-Metalium low level kernel programming model.
★ 1.6kMuon. Muon is an optimizer for hidden layers in neural networks
★ 2.8klayered_vectorization. JavaScript
★ 196scrap. 📸 Screen capture made easy!
★ 650diamond. DIAMOND (DIffusion As a Model Of eNvironment Dreams) is a reinforcement learning agent trained in a diffusion world model. NeurIPS 2024 Spotlight.
★ 2.1kDynamicVectorQuantization. Official Pytorch Implementation of Our CVPR2023 Paper: "Towards Accurate Image Coding: Improved Autoregressive Image Generation with Dynamic Vector Quantization"
★ 194tt-buda. Tenstorrent TT-BUDA Repository
★ 314pyglove. Manipulating Python Programs
★ 717flux. Official inference repo for FLUX.1 models
★ 26ksigma-gpt. σ-GPT: A New Approach to Autoregressive Models
★ 77sparsify. Sparsify transformers with SAEs and transcoders
★ 735haliax. Named Tensors for Legible Deep Learning in JAX
★ 227optax. Optax is a gradient processing and optimization library for JAX.
★ 2.3korbax. Orbax provides common checkpointing and persistence utilities for JAX users
★ 526levanter. Legible, Scalable, Reproducible Foundation Models with Named Tensors and Jax
★ 708equinox. Elegant easy-to-use neural networks + scientific computing in JAX. https://docs.kidger.site/equinox/
★ 2.9kflax. Flax is a neural network library for JAX that is designed for flexibility.
★ 7.3kllm.c. LLM training in simple, raw C/CUDA
★ 31kMegatron-LM. Ongoing research training transformer models at scale
★ 17kHVM2. A massively parallel, optimal functional runtime in Rust
★ 11kmmdit. Implementation of a single layer of the MMDiT, proposed in Stable Diffusion 3, in Pytorch
★ 553FourierKAN. Python
★ 755pykan. Kolmogorov Arnold Networks
★ 16kdiffusion-kto. The official implementation of Diffusion-KTO: Aligning Diffusion Models by Optimizing Human Utility
★ 68torchtitan. A PyTorch native platform for training generative AI models
★ 5.6kpyo3. Rust bindings for the Python interpreter
★ 16kScore-Entropy-Discrete-Diffusion. [ICML 2024 Best Paper] Discrete Diffusion Modeling by Estimating the Ratios of the Data Distribution (https://arxiv.org/abs/2310.16834)
★ 740Betterfox. Firefox user.js for optimal privacy and security. Your favorite browser, but better.
★ 11kp5.brush. Unlock custom brushes, natural fill effects and intuitive hatching in p5.js
★ 702gptq. Code for the ICLR 2023 paper "GPTQ: Accurate Post-training Quantization of Generative Pretrained Transformers".
★ 2.3kbacktesting.py. 🔎 📈 🐍 💰 Backtest trading strategies in Python.
★ 8.7kWeylus. Use your tablet as graphic tablet/touch screen on your computer.
★ 9.4ktypst. A markup-based typesetting system that is powerful and easy to learn.
★ 55kPocketKitchen-UI. Pocket Kitchen | Node.js & Express web app and MongoDB data pipleline for presenting actionable food waste prevention insights
★ 1hccl_demo. C++
★ 26candle. Minimalist ML framework for Rust
★ 21kHotshot-XL. ✨ Hotshot-XL: State-of-the-art AI text-to-GIF model trained to work alongside Stable Diffusion XL
★ 1.1kProPainter. [ICCV 2023] ProPainter: Improving Propagation and Transformer for Video Inpainting
★ 6.8kbark. 🔊 Text-Prompted Generative Audio Model
★ 39kgenerative_agents. Generative Agents: Interactive Simulacra of Human Behavior
★ 22kminSDXL. Huggingface-compatible SDXL Unet implementation that is readily hackable
★ 439novelai-aspect-ratio-bucketing. Implementation of aspect ratio bucketing for training generative image models as described in: https://blog.novelai.net/novelai-improvements-on-stable-diffusion-e10d38db82ac
★ 397OnnxStream. Lightweight inference library for ONNX files, written in C++. It can run Stable Diffusion XL 1.0 on a RPI Zero 2 (or in 298MB of RAM) but also Mistral 7B on desktops and servers. ARM, x86, WASM, RISC-V supported. Accelerated by XNNPACK. Python, C# and JS(WASM) bindings available.
★ 2.1kblibla-comfyui-extensions. Extensions for ComfyUI
★ 178LLaVAR. Code/Data for the paper: "LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding"
★ 268MeZO. [NeurIPS 2023] MeZO: Fine-Tuning Language Models with Just Forward Passes. https://arxiv.org/abs/2305.17333
★ 1.2kuss. This is the PyTorch implementation of the Universal Source Separation with Weakly labelled Data.
★ 368prolificdreamer. ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score Distillation (NeurIPS 2023 Spotlight)
★ 1.6kaudio-diffusion. Python
★ 87StableLM. StableLM: Stability AI Language Models
★ 16kAll-In-One-Deflicker. [CVPR2023] Blind Video Deflickering by Neural Filtering with a Flawed Atlas
★ 761composer. Official implementation of "Composer: Creative and Controllable Image Synthesis with Composable Conditions"
★ 1.6kComfyUI. The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
★ 123kChatGPT. Reverse engineered ChatGPT API
★ 28kControlNet. Let us control diffusion models!
★ 34kxg2xg. by ex-googlers, for ex-googlers - a lookup table of similar tech & services
★ 16kui. A set of beautifully-designed, accessible components and a code distribution platform. Works with your favorite frameworks. Open Source. Open Code.
★ 120kJWildfire. JWildfire - powerful, flexible and user-friendly fractal flame editor
★ 252audiolm-pytorch. Implementation of AudioLM, a SOTA Language Modeling Approach to Audio Generation out of Google Research, in Pytorch
★ 2.6kaudio-dataset. Audio Dataset for training CLAP and other models
★ 748RDM-Region-Aware-Diffusion-Model. Python
★ 185dygen. Swift
★ 14sd-leap-booster. Fast finetuning using a booster model that puts the initial state to a local minimum
★ 113awesome-design-patterns. A curated list of software and architecture related design patterns.
★ 48kdiffvg. Differentiable Vector Graphics Rasterization
★ 1.3kDreambooth-Stable-Diffusion. Implementation of Dreambooth (https://arxiv.org/abs/2208.12242) by way of Textual Inversion (https://arxiv.org/abs/2208.01618) for Stable Diffusion (https://arxiv.org/abs/2112.10752). Tweaks focused on training faces, objects, and styles.
★ 3.2kstable-dreamfusion. Text-to-3D & Image-to-3D & Mesh Exportation with NeRF + Diffusion.
★ 8.8ksubmitit. Python 3.8+ toolbox for submitting jobs to Slurm
★ 1.6kAITemplate. AITemplate is a Python framework which renders neural network into high performance CUDA/HIP C++ code. Specialized for FP16 TensorCore (NVIDIA GPU) and MatrixCore (AMD GPU) inference.
★ 4.7kDreambooth-Stable-Diffusion. Implementation of Dreambooth (https://arxiv.org/abs/2208.12242) with Stable Diffusion
★ 7.7ktextual_inversion. Jupyter Notebook
★ 3.1kstable-diffusion. Jupyter Notebook
★ 1.5kDiSS. Adaptively-Realistic Image Generation from Stroke and Sketch with Diffusion Model (WACV 2023)
★ 120stability-sdk. SDK for interacting with stability.ai APIs (e.g. stable diffusion inference)
★ 2.4kkrita. Krita is a free and open source cross-platform application that offers an end-to-end solution for creating digital art files from scratch built on the KDE and Qt frameworks.
★ 10kstable-diffusion. A latent text-to-image diffusion model
★ 73ktoma. Helps you write algorithms in PyTorch that adapt to the available (CUDA) memory
★ 436GODEL. Large-scale pretrained models for goal-directed dialog
★ 880dalle2-laion. Pretrained Dalle2 from laion
★ 505brain-tokyo-workshop. 🧠🗼
★ 1.3kema-pytorch. A simple way to keep track of an Exponential Moving Average (EMA) version of your Pytorch model
★ 658diffusers. 🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
★ 34kpytorch. Tensors and Dynamic neural networks in Python with strong GPU acceleration
★ 102kimagen-pytorch. Implementation of Imagen, Google's Text-to-Image Neural Network, in Pytorch
★ 8.4kinstant-ngp. Instant neural graphics primitives: lightning fast NeRF and more
★ 18kaccelerate. 🚀 A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (including fp8), and easy-to-configure FSDP and DeepSpeed support
★ 9.8kcheckout. Action for checking out a repo
★ 8.6ksetup-python. Set up your GitHub Actions workflow with a specific version of Python
★ 2.2kgh-action-pypi-publish. The blessed :octocat: GitHub Action, for publishing your :package: distribution files to PyPI, the tokenless way: https://github.com/marketplace/actions/pypi-publish
★ 1.2kaesthetic-predictor. A linear estimator on top of clip to predict the aesthetic quality of pictures
★ 727DALLE-pytorch. Implementation / replication of DALL-E, OpenAI's Text to Image Transformer, in Pytorch
★ 5.6klightweight-gan. Implementation of 'lightweight' GAN, proposed in ICLR 2021, in Pytorch. High resolution image generations that can be trained within a day or two
★ 1.7kThin-Plate-Spline-Motion-Model. [CVPR 2022] Thin-Plate Spline Motion Model for Image Animation.
★ 3.6kCLIP. CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
★ 34kdalle-mini. DALL·E Mini - Generate images from a text prompt
★ 15ktortoise-tts. A multi-voice TTS system trained with an emphasis on quality
★ 15kembedding-dataset-reordering. Reorders the embeddings generated by CLIP to be in line with the webdatasets generated by img2dataset.
★ 5python-template. Simple python template
★ 44Awesome-Diffusion-Models. A collection of resources and papers on Diffusion Models
★ 12kimg2dataset. Easily turn large sets of image urls to an image dataset. Can download, resize and package 100M urls in 20h on one machine.
★ 4.4kclip_benchmark. clip retrieval benchmark
★ 17STEGO. Unsupervised Semantic Segmentation by Distilling Feature Correspondences
★ 791flamingo-pytorch. Implementation of 🦩 Flamingo, state-of-the-art few-shot visual question answering attention net out of Deepmind, in Pytorch
★ 1.3kmedical. This repository will be a summary and outlook on all our open, medical, AI advancements.
★ 30clip-image-search. Fine-tuning OpenAI CLIP Model for Image Search on medical images
★ 76glid-3-xl. 1.4B latent diffusion model fine tuning
★ 265NeuralNeighborStyleTransfer. Optimization based style transfer
★ 259django-tailwind. Django + Tailwind CSS = 💚
★ 1.8kblack. The uncompromising Python code formatter
★ 42kwebdataset. A high-performance Python-based I/O system for large (and small) deep learning problems, with strong support for PyTorch.
★ 3.1kembedding-reader. Efficiently read embedding in streaming from any filesystem
★ 106contrastive-unpaired-translation. Contrastive unpaired image-to-image translation, faster and lighter training than cyclegan (ECCV 2020, in PyTorch)
★ 2.5kDALLE2-pytorch. Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis neural network, in Pytorch
★ 11kDetic. Code release for "Detecting Twenty-thousand Classes using Image-level Supervision".
★ 2kpipdeptree. A command line utility to display dependency tree of the installed Python packages
★ 3klatent-diffusion. High-Resolution Image Synthesis with Latent Diffusion Models
★ 14kaiaiart. Course content and resources for the AIAIART course.
★ 567Tensor-Puzzles. Solve puzzles. Improve your pytorch.
★ 4.3keg3d. Python
★ 3.3kmagma. MAGMA - a GPT-style multimodal model that can understand any combination of images and language. NOTE: The freely available model from this repo is only a demo. For the latest multimodal and multilingual models from Aleph Alpha check out our website https://app.aleph-alpha.com
★ 489GAMA-GAT. Guided Adversarial Attack for Evaluating and Enhancing Adversarial Defenses, NeurIPS Spotlight 2020
★ 27ru-dolph. RUDOLPH: One Hyper-Tasking Transformer can be creative as DALL-E and GPT-3 and smart as CLIP
★ 254dl_tutor. Jupyter Notebook
★ 73