This is your work, valued
Awesome-Diffusion-Language-Models.
★ 35GeoPrivacy. Jupyter Notebook
★ 6Urban-CI-LitReview. Python
★ 3GeoDifferentialPrivacy. Jupyter Notebook
★ 3NGSIM-IDM. This repo illustrates how driver models can be tuned using NGSIM data.
★ 2quao627.github.io. HTML
★ 2gpt-wrapper. JavaScript
★ 1RMP. Python
★ 1AI4UrbanScience. JavaScript
★ 1ActiveLearningWithCompositeKernel. Jupyter Notebook
★ 1LLM-learning. GPT, LLaMa, Mixtral MoE, RLHF and other LLM learning notes
★ 1RLBus. Python
★ 1AGRI9999-Seminar-in-Python. Jupyter Notebook
★ 1IDM-CitySim. Jupyter Notebook
★ 1osmint-fix. Python
★ 1openevolve-fixed. Python
★ 1Dr-CiK. Dr-CiK: A Testbed for Foresight-driven Agents
★ 6FrontierOR. A benchmark of 180 literature-grounded OR problems evaluating LLM-generated efficient algorithms
★ 25AwesomeOPD. Awesome List for On-Policy Distillation
★ 773EoM. Python
★ 49continual-learning-bench. Continual Learning Bench
★ 190marin. Open-source framework for the research and development of foundation models.
★ 1.2kWorldCupArena. ⚽️🤖 Benchmarking LLMs and deep-research agents on real-world football prediction — from the tactical "who scores in minute 67" to the strategic "who wins the World Cup."
★ 20awesome-open-ended. Awesome Open-ended AI
★ 459GoNavi. Native multi-data-source DB client — ~30MB, AI & MCP ready. Not Electron. 丨原生轻量数据库客户端(Go + Wails + React,~30MB):SQL/缓存/向量/消息/国产库,内置 AI 与 MCP,告别 Electron。
★ 1.8kgingiris-opensource. ⭐ Open Source Marketing Playbook — How to get 10k+ GitHub stars. Complete OSS launch strategy from AFFiNE (60k stars in 24 months). GitHub trending, HackerNews launch, Reddit r/programming, developer community, awesome-lists, contributor acquisition.
★ 240Ozone. Next-generation Benchmark by Digital Twin
★ 24Frontier-CS. A benchmark for evaluating LLMs on open-ended CS problems. Exploring the Next Frontier of Computer Science.
★ 288AlphaOPT. Formulating Optimization Programs with Self-Improving LLM Experience Library
★ 27ResearchClawBench. 🦞 ResearchClawBench: Evaluating AI Agents for Automated Research from Re-Discovery to New-Discovery
★ 227CORAL. 🔥🔥COLM 2026🔥🔥 CORAL is a robust, lightweight infrastructure for multi-agent autonomous self-evolution, built for autoresearch. Works with Claude Code, Codex, Cursor, OpenCode, Kiro, and more.
★ 860InfinityStar. [NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation
★ 774reaction-video. Python
★ 1Stable-Video-Infinity. [ICLR 26 Oral] Stable Video Infinity: Infinite-Length Video Generation with Error Recycling
★ 2.5kmirl. verl: Volcano Engine Reinforcement Learning for LLMs
★ 13Ovi. Python
★ 1.7kDeepSeek-OCR. Contexts Optical Compression
★ 24kquao627.github.io. HTML
★ 3Wan2.2. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kDiffSynth-Studio. Enjoy the magic of Diffusion models!
★ 13kDolphin. Python
★ 186Qwen3-Omni. Qwen3-omni is a natively end-to-end, omni-modal LLM developed by the Qwen team at Alibaba Cloud, capable of understanding text, audio, images, and video, as well as generating speech in real time.
★ 3.9ktinyworlds. A minimal implementation of DeepMind's Genie world model
★ 1.3kITINERA. [EMNLP 2024 Industry Track & KDD UrbComp 2024 Best Paper Award] ITINERA: Integrating Spatial Optimization with Large Language Models for Open-domain Urban Itinerary Planning
★ 67DiT. Official PyTorch Implementation of "Scalable Diffusion Models with Transformers"
★ 8.7kPosterGen. Official Repository for PosterGen - CVPR Findings 2026
★ 247crisp-pipeline. crisp-pipeline
★ 27HuMo. HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning
★ 1.3km3-agent. Python
★ 1.4kSmellNet. Jupyter Notebook
★ 69seamless_interaction. Foundation Models and Data for Human-Human and Human-AI interactions.
★ 405FonTS-WordCon. [ICCV 2025 / TCSVT 2026] FonTS: Text Rendering with Typography and Style Controls / WordCon: Word-level Typography Control in Visual Text Rendering
★ 44GraphGen. GraphGen: Enhancing Supervised Fine-Tuning for LLMs with Knowledge-Driven Synthetic Data Generation
★ 1.1kMimeQA. Python
★ 5hallo3. [CVPR 2025] Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer
★ 1.4kMatrix-Game. Matrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory
★ 2.3kDreamVLA. [NeurIPS 2025] DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge
★ 364Awesome-Out-Of-Distribution-Detection. [ACM CSUR 2025] Out-of-Distribution Detection: A Task-Oriented Survey of Recent Advances
★ 171MEM1. Python
★ 326LLaDOU. Implementation of "Reinforcing the Diffusion Chain of Lateral Thought with Diffusion Language Models" [NeurIPS 2025]
★ 82InPO. [CVPR 2025 Highlight] InPO: Inversion Preference Optimization with Reparametrized DDIM for Efficient Diffusion Model Alignment
★ 44ATOM. ATOM: A Framework of Detecting Query-Based Model Extraction Attacks for Graph Neural Networks
★ 18MonSter. 【CVPR 2025 Highlight】MonSter: Marry Monodepth to Stereo Unleashes Power
★ 699ml-diffucoder. DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation
★ 832awesome-latex-drawing. Drawing Bayesian networks, graphical models, tensors, technical frameworks, and illustrations in LaTeX.
★ 2kconrft. This is the official implementation of the paper "ConRFT: A Reinforced Fine-tuning Method for VLA Models via Consistency Policy".
★ 361CACFM. ECCV 2026 accepted CACFM code: RL-guided curvature-adaptive consistency flow matching for few-step FLUX/SDXL generation
★ 55Paper-List-for-Medical-Anomaly-Detection. :octocat: A paper list for medical anomaly detection. ℱℯℯ𝓁 𝒻𝓇ℯℯ to contribute!
★ 58Awesome-Diffusion-Language-Models.
★ 35Awesome-Diffusion-Models. A collection of resources and papers on Diffusion Models
★ 12kScore-Entropy-Discrete-Diffusion. [ICML 2024 Best Paper] Discrete Diffusion Modeling by Estimating the Ratios of the Data Distribution (https://arxiv.org/abs/2310.16834)
★ 740Academic-project-page-template. A project page template for academic papers. Demo at https://eliahuhorwitz.github.io/Academic-project-page-template/
★ 5.1kPuzzleWorld. Python
★ 14CompAct. [EMNLP 2024] CompAct: Compressing Retrieved Documents Actively for Question Answering
★ 37272-dim-Motion-Representation. The modified 272-dimensional motion representation processing script.
★ 232MotionGPT. [NeurIPS 2023] MotionGPT: Human Motion as a Foreign Language, a unified motion-language generation model using LLMs
★ 1.9kcsm. A Conversational Speech Generation Model
★ 15kopeninterpreter. A coding agent for open models like Kimi K3
★ 67kosmint-fix. Python
★ 1self-llm. 《开源大模型食用指南》针对中国宝宝量身打造的基于Linux环境快速微调(全参数/Lora)、部署国内外开源大模型(LLM)/多模态大模型(MLLM)教程
★ 31kLLM-RLHF-Tuning-with-PPO-and-DPO. Comprehensive toolkit for Reinforcement Learning from Human Feedback (RLHF) training, featuring instruction fine-tuning, reward model training, and support for PPO and DPO algorithms with various configurations for the Alpaca, LLaMA, and LLaMA2 models.
★ 190LLM101n. LLM101n: Let's build a Storyteller
★ 38kOnline-RLHF. A recipe for online RLHF and online iterative DPO.
★ 544vocode-core. 🤖 Build voice-based LLM agents. Modular + open source.
★ 3.8kGrounded-SAM-2. Grounded SAM 2: Ground and Track Anything in Videos with Grounding DINO, Florence-2 and SAM 2
★ 3.7ksegment-geospatial. A Python package for segmenting geospatial data with the Segment Anything Model (SAM)
★ 4.1ksemantic-segmentation-pytorch. Pytorch implementation for Semantic Segmentation/Scene Parsing on MIT ADE20K dataset
★ 5.1kGrounded-Segment-Anything. Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything
★ 18kSemantic-SAM. [ECCV 2024] Official implementation of the paper "Semantic-SAM: Segment and Recognize Anything at Any Granularity"
★ 2.9kGroundingDINO. [ECCV 2024] Official implementation of the paper "Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection"
★ 10kaudio-ai-hub. The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
★ 949llama-cookbook. Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama model family and using them on various provider services
★ 19kmodular_pluralism. Modular Pluralism @ EMNLP 2024
★ 26gpt-wrapper. JavaScript
★ 1RealtimeTTS. Converts text to speech in realtime
★ 4kBayLing-Speech. LLaMA-Omni is a low-latency and high-quality end-to-end speech interaction model built upon Llama-3.1-8B-Instruct, aiming to achieve speech capabilities at the GPT-4o level.
★ 3.1kQwen2-Audio. The official repo of Qwen2-Audio chat & pretrained large audio language model proposed by Alibaba Cloud.
★ 2.1kms-swift. Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
★ 15kVLMEvalKit. Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks
★ 4.3kindie-hacker-tools. 收录独立开发者出海技术栈和工具
★ 7.9kdeveloper-roadmap. Interactive roadmaps, guides and other educational content to help developers grow in their careers.
★ 363kchrome-extensions-samples. Chrome Extensions Samples
★ 18kflutter_wechat. 🔥🔥🔥 利用 Flutter 来高仿微信(WeChat) 7.0.0+ App,代码规范惊为天人、注释详解令人发指、细节处理精益求精、核心功能配备文档、接近98%还原度的原生App视觉体验。代码不多,注释多。(持续更新,敬请期待,欢迎Star和Fork…)
★ 721distributed-web-crawler. The Architecture of a Web Crawler: Building a Google-Inspired Distributed Web Crawler
★ 125ollama. Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
★ 177kllm-course. Course to get into Large Language Models (LLMs) with roadmaps and Colab notebooks.
★ 81kReal3DPortrait. Real3D-Portrait: One-shot Realistic 3D Talking Portrait Synthesis; ICLR 2024 Spotlight; Official code
★ 1.1kSadTalker. [CVPR 2023] SadTalker:Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face Animation
★ 14kHeyGenClone. A simple and open-source analogue of the HeyGen system
★ 1kTTS. 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
★ 46kdeepface. A Lightweight Face Recognition and Facial Attribute Analysis (Age, Gender, Emotion and Race) Library for Python
★ 23kvideo-retalking. [SIGGRAPH Asia 2022] VideoReTalking: Audio-based Lip Synchronization for Talking Head Video Editing In the Wild
★ 7.3kawesome-bilibili. b 站的宝藏 up 主名单,他们教技术。关于 Web 开发、计算机科学、机器学习、游戏开发、网络安全等 😎
★ 1.1kIP-Adapter. The image prompt adapter is designed to enable a pretrained text-to-image diffusion model to generate images with image prompt.
★ 6.6kInstantID. InstantID: Zero-shot Identity-Preserving Generation in Seconds 🔥
★ 12kTikTok-Api. The Unofficial TikTok API Wrapper In Python
★ 6.5kfacechain. FaceChain is a deep-learning toolchain for generating your Digital-Twin.
★ 9.5kLLMsNineStoryDemonTower. 【LLMs九层妖塔】分享 LLMs在自然语言处理(ChatGLM、Chinese-LLaMA-Alpaca、小羊驼 Vicuna、LLaMA、GPT4ALL等)、信息检索(langchain)、语言合成、语言识别、多模态等领域(Stable Diffusion、MiniGPT-4、VisualGLM-6B、Ziya-Visual等)等 实战与经验。
★ 2.2kLLM-Tuning. Tuning LLMs with no tears💦; Sample Design Engineering (SDE) for more efficient downstream-tuning.
★ 1kRLBus. Python
★ 1multi-agent-emergence-environments. Environment generation code for the paper "Emergent Tool Use From Multi-Agent Autocurricula"
★ 1.8kawesome-diffusion-model-in-rl. A curated list of Diffusion Model in RL resources (continually updated)
★ 1.6kdreamer-pytorch. Dream to Control: Learning Behaviors by Latent Imagination, implemented in PyTorch.
★ 326undetected-chromedriver. Custom Selenium Chromedriver | Zero-Config | Passes ALL bot mitigation systems (like Distil / Imperva/ Datadadome / CloudFlare IUAM)
★ 13kMP-DQN. Source code for the dissertation: "Multi-Pass Deep Q-Networks for Reinforcement Learning with Parameterised Action Spaces"
★ 239DI-engine. OpenDILab Decision AI Engine. The Most Comprehensive Reinforcement Learning Framework B.P.
★ 3.6kvista. Data-driven simulation for training and evaluating full-scale autonomous vehicles.
★ 409DI-sheep. 羊了个羊 + 深度强化学习(Deep Reinforcement Learning + 3 Tiles Game)
★ 518VehicleFinder-CTIM. Python
★ 5FindVehicle. FindVehicle: A NER dataset in transportation to extract keywords describing vehicles on the road
★ 40PyTrack. a Map-Matching-based Python Toolbox for Vehicle Trajectory Reconstruction
★ 79mit602. Solutions to MIT 6.02 (Introduction to EECS II: Digital Communication Systems) assignments
★ 3spinningup. An educational resource to help anyone learn deep reinforcement learning.
★ 12kimitation-learning. Imitation learning algorithms
★ 571tensor_tools. 张量分解算法整理
★ 64awesome-autonomous-vehicles. Curated List of Self-Driving Cars and Autonomous Vehicles Resources
★ 2.4kawesome-self-supervised-gnn. Papers about pretraining and self-supervised learning on Graph Neural Networks (GNN).
★ 1.7kdifferential-privacy-library. Diffprivlib: The IBM Differential Privacy Library
★ 918tutorial. Jupyter Notebook
★ 791stanford-cs-229-machine-learning. VIP cheatsheets for Stanford's CS 229 Machine Learning
★ 20kGeoPrivacy. Jupyter Notebook
★ 6numpy-100. 100 numpy exercises (with solutions)
★ 14kGeoDifferentialPrivacy. Jupyter Notebook
★ 3deeprl_signal_control. multi-agent deep reinforcement learning for large-scale traffic signal control.
★ 425deeprl_network. multi-agent deep reinforcement learning for networked system control.
★ 448transferlearning. Transfer learning / domain adaptation / domain generalization / multi-task learning etc. Papers, codes, datasets, applications, tutorials.-迁移学习
★ 14kmacad-gym. Multi-Agent Connected Autonomous Driving (MACAD) Gym environments for Deep RL. Code for the paper presented in the Machine Learning for Autonomous Driving Workshop at NeurIPS 2019:
★ 374Flappy-bird-deep-Q-learning-pytorch. Deep Q-learning for playing flappy bird game
★ 559Rebuild-My-Professor.
★ 1AGRI9999-Seminar-in-Python. Jupyter Notebook
★ 1Rebuild-My-Professor. Jupyter Notebook
★ 2huizhi-photo. Using A Jekyll theme for photo for Huizhi, Fu. 📸
★ 1RMP. Python
★ 1