This is your work, valued
Think deep, work hard.
EasyPR. (CGCSTCD'2017) An easy, flexible, and accurate plate recognition project for Chinese licenses in unconstrained situations. CGCSTCD = China Graduate Contest on Smart-city Technology and Creative Design
★ 6.4kmini-AlphaStar. (JAIR'2022) A mini-scale reproduction code of the AlphaStar program. Note: the original AlphaStar is the AI proposed by DeepMind to play StarCraft II. JAIR = Journal of Artificial Intelligence Research.
★ 371HierNet-SC2. (AAAI'2019) The codes, models, logs, and data for an extended paper of the original paper "On Reinforcement Learning for Full-length Game of StarCraft". AAAI = AAAI Conference on Artificial Intelligence.
★ 34EasyPR_Win. An windows exe using EasyPR, get and try it!
★ 23EasyPR_Dll. An dll to do the same function of EasyPR, and the sample !
★ 16Thought-SC2. (TG'2021) Code for paper "Efficient Reinforcement Learning for StarCraft by Abstract Forward Models and Transfer Learning". TG = Transactions on Games.
★ 11Raw-vs-Human-in-AlphaStar. (TG'2023) Official code for the paper "Revisiting of AlphaStar" (previously called "Rethinking of AlphaStar"). It compares the raw interface with the human interface in the architecture of AlphsStar. TG = Transactions on Games
★ 10EasyPR_Dll_src. EasyPR_Dll all src
★ 10BetaStar. (SSCAIT'2019) BetaStar is a StarCraft AI, written by a team at Nanjing University. It is written for StarCraft I. SSCAIT=Student StarCraft AI Tournament.
★ 5Tea. Tea is a simple and easy Deep Learning framework aims for teaching, reading and learning. More details will come soon in a few months.
★ 3Useful-Big-Resources. Provide data, gif images, and other big resources for other projects of mine.
★ 2muzero-general. MuZero
★ 2liuruoze. My presonal repository
★ 1StarCraft. Implementations of QMIX, VDN, COMA, QTRAN, MAVEN, CommNet, DyMA-CL, and G2ANet on SMAC, the decentralised micromanagement scenario of StarCraft II
★ 1jieba. 结巴中文分词
★ 1EasyPR_GTDS. The general test data set reside here.
★ 1spark. Mirror of Apache Spark
★ 1FramePack. Lets make video diffusion practical!
★ 17kcleanrl. High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)
★ 10kppo-implementation-details. The source code for the blog post The 37 Implementation Details of Proximal Policy Optimization
★ 946invalid-action-masking. Source Code for A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
★ 168grok-1. Grok open release
★ 52kNSFC-application-template-latex. 国家自然科学基金申请书正文(面上项目)LaTeX 模板(非官方)
★ 1.1kAnimateAnyone. Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation
★ 15kmagic-animate. [CVPR 2024] Official repository for "MagicAnimate: Temporally Consistent Human Image Animation using Diffusion Model"
★ 11kgenerative_agents. Generative Agents: Interactive Simulacra of Human Behavior
★ 22kaudiocraft. Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor / tokenizer, along with MusicGen, a simple and controllable music generation LM with textual and melodic conditioning.
★ 24kmamba. This code accompanies the paper "Scalable Multi-Agent Model-Based Reinforcement Learning".
★ 67jax. Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and more
★ 36kxformers. Hackable and optimized Transformers building blocks, supporting a composable construction.
★ 11kllama. Inference code for Llama models
★ 60kdreamerv3. Mastering Diverse Domains through World Models
★ 3.6kso-vits-svc. SoftVC VITS Singing Voice Conversion
★ 28kstreet-fighter-ai. This is an AI agent for Street Fighter II Champion Edition.
★ 6.5kDeepSpeed. DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
★ 43kgpt_academic. 为GPT/GLM等LLM大语言模型提供实用化交互接口,特别优化论文阅读/润色/写作体验,模块化设计,支持自定义快捷按钮&函数插件,支持Python和C++等项目剖析&自译解功能,PDF/LaTex论文翻译&总结功能,支持并行问询多种LLM模型,支持chatglm3等本地模型。接入通义千问, deepseekcoder, 讯飞星火, 文心一言, llama2, rwkv, claude2, moss等。
★ 71knju-thesis. 南京大学学位论文XeLaTeX模板
★ 441NLP-Tutorials. Simple implementations of NLP models. Tutorials are written in Chinese on my website https://mofanpy.com
★ 950continual-learning-papers. Continual Learning papers list, curated by ContinualAI
★ 726LawRefBook. 中华人民共和国法律手册
★ 2.4kmujoco. Multi-Joint dynamics with Contact. A general purpose physics simulator.
★ 14kobjgraph. Visually explore Python object graphs
★ 842alphafold. Open source code for AlphaFold 2.
★ 15krlpyt. Reinforcement Learning in PyTorch
★ 2.3kRLzoo. A Comprehensive Reinforcement Learning Zoo for Simple Usage 🚀
★ 640Deep-Reinforcement-Learning-Algorithms-with-PyTorch. PyTorch implementations of deep reinforcement learning algorithms and environments
★ 5.9kQuantumExplorer. A quantum reinforcement learning framework based on PyTorch and PennyLane.
★ 31awesome-offline-rl. An index of algorithms for offline reinforcement learning (offline-rl)
★ 1.1kGitHubDaily. 坚持分享 GitHub 上高质量、有趣实用的开源技术教程、开发者工具、编程网站、技术资讯。A list cool, interesting projects of GitHub.
★ 47kglide-text2im. GLIDE: a diffusion-based text-conditional image synthesis model
★ 3.7kUseful-Big-Resources. Provide data, gif images, and other big resources for other projects of mine.
★ 2liuruoze. My presonal repository
★ 1HierNet-SC2. (AAAI'2019) The codes, models, logs, and data for an extended paper of the original paper "On Reinforcement Learning for Full-length Game of StarCraft". AAAI = AAAI Conference on Artificial Intelligence.
★ 35PySnooper. Never use print for debugging again
★ 17kPrefixSpan-py. Data mining algorithm PrefixSpan based on Python/数据挖掘算法PrefixSpan的简单Python实现
★ 22PrefixSpan-py. The shortest yet efficient Python implementation of the sequential pattern mining algorithm PrefixSpan, closed sequential pattern mining algorithm BIDE, and generator sequential pattern mining algorithm FEAT.
★ 427pytorch-a2c-ppo-acktr-gail. PyTorch implementation of Advantage Actor Critic (A2C), Proximal Policy Optimization (PPO), Scalable trust-region method for deep reinforcement learning using Kronecker-factored approximation (ACKTR) and Generative Adversarial Imitation Learning (GAIL).
★ 3.9kpytorch-distributed. A quickstart and benchmark for pytorch distributed training.
★ 1.7kenvpool. C++-based high-performance parallel environment execution engine (vectorized env) for general RL environments.
★ 1.5kTiKick. Learning-based agent for Google Research Football (足球游戏智能体)
★ 120pytorch-multigpu. Multi GPU Training Code for Deep Learning with PyTorch
★ 211K-Net. [NeurIPS2021] Code Release of K-Net: Towards Unified Image Segmentation
★ 484Soft-Actor-Critic-and-Extensions. PyTorch implementation of Soft-Actor-Critic and Prioritized Experience Replay (PER) + Emphasizing Recent Experience (ERE) + Munchausen RL + D2RL and parallel Environments.
★ 297DI-star. An artificial intelligence platform for the StarCraft II with large-scale distributed training and grand-master agents.
★ 1.4kAlphaStar_Implementation. This project is implementation code of AlphaStar
★ 208yzhao062.
★ 12pypeln. Concurrent data pipelines in Python >>>
★ 1.6kgin-config. Gin provides a lightweight configuration framework for Python
★ 2.2kRaw-vs-Human-in-AlphaStar. (TG'2023) Official code for the paper "Revisiting of AlphaStar" (previously called "Rethinking of AlphaStar"). It compares the raw interface with the human interface in the architecture of AlphsStar. TG = Transactions on Games
★ 10BetaStar. (SSCAIT'2019) BetaStar is a StarCraft AI, written by a team at Nanjing University. It is written for StarCraft I. SSCAIT=Student StarCraft AI Tournament.
★ 5sc2_imitation_learning. StarCraft 2 Imitation Learning
★ 29seed_rl. SEED RL: Scalable and Efficient Deep-RL with Accelerated Central Inference. Implements IMPALA and R2D2 algorithms in TF2 with SEED's architecture.
★ 836paper.
★ 43AirSim. Open source simulator for autonomous vehicles built on Unreal Engine / Unity, from Microsoft AI & Research
★ 18kDefogGAN. Python
★ 30Transfer-Learning-Library. Transfer Learning Library for Domain Adaptation, Task Adaptation, and Domain Generalization
★ 3.9kSMARTS. Scalable Multi-Agent RL Training School for Autonomous Driving
★ 1.1kModelRepo. reproduce some RL or Multi-Agent models
★ 35deepmind_MAS_enviroment. some Multiagent enviroment in 《Multi-agent Reinforcement Learning in Sequential Social Dilemmas》 and 《Value-Decomposition Networks For Cooperative Multi-Agent Learning》
★ 131LastOrder-Dota2. Dota2 AI bot
★ 419rainbow-is-all-you-need. Rainbow is all you need! A step-by-step tutorial from DQN to Rainbow
★ 2kyolort. yolort is a runtime stack for yolov5 on specialized accelerators such as tensorrt, libtorch, onnxruntime, tvm and ncnn.
★ 729BackgroundMattingV2. Real-Time High-Resolution Background Matting
★ 7.2kResnetGPT. 用Resnet101+GPT搭建一个玩王者荣耀的AI
★ 3krecommenders. Best Practices on Recommendation Systems
★ 22kminirts. We release dataset collected for our research, code that implement neural network models described in the paper, and scripts to reproduce all of our results, and visualization tool for visualize dataset.
★ 161deep-rts. A Real-Time-Strategy game for Deep Learning research
★ 249rl-starter-files. RL starter files in order to immediately train, visualize and evaluate an agent without writing any line of code
★ 726gym-carla. An OpenAI gym wrapper for CARLA simulator
★ 618slimevolleygym. A simple OpenAI Gym environment for single and multi-agent reinforcement learning
★ 786HighwayEnv. A collection of environments for autonomous driving and tactical decision-making tasks
★ 3.3kgym-anytrading. The most simple, flexible, and comprehensive OpenAI Gym trading environment (Approved by OpenAI Gym)
★ 2.4kMiniworld. Simple and easily configurable 3D FPS-game-like environments for reinforcement learning
★ 773Minigrid. Simple and easily configurable grid world environments for reinforcement learning
★ 2.5kgym-maze. A basic 2D maze environment where an agent start from the top left corner and try to find its way to the bottom left corner.
★ 375retro. Retro Games in Gym
★ 3.6kprocgen. Procgen Benchmark: Procedurally-Generated Game-Like Gym-Environments
★ 1.2kFishSchoolSim. JavaScript
★ 49summarize-from-feedback. Code for "Learning to summarize from human feedback"
★ 1.1kmuzero-general. MuZero
★ 2StarCraft. Implementations of QMIX, VDN, COMA, QTRAN, MAVEN, CommNet, DyMA-CL, and G2ANet on SMAC, the decentralised micromanagement scenario of StarCraft II
★ 1mini-AlphaStar. (JAIR'2022) A mini-scale reproduction code of the AlphaStar program. Note: the original AlphaStar is the AI proposed by DeepMind to play StarCraft II. JAIR = Journal of Artificial Intelligence Research.
★ 371TRGPPO. Python
★ 34MARL-Algorithms. Implementations of IQL, QMIX, VDN, COMA, QTRAN, MAVEN, CommNet, DyMA-CL, and G2ANet on SMAC, the decentralised micromanagement scenario of StarCraft II
★ 1.8kQTRAN. There will be updates later
★ 87overcooked_ai. A benchmark environment for fully cooperative human-AI performance.
★ 989opponent. Implementation for ICML 16 paper "Deep reinforcement learning with opponent modeling"
★ 71lstms.pth. PyTorch implementations of LSTM Variants (Dropout + Layer Norm)
★ 137Balanced-DataParallel. 这里是改进了pytorch的DataParallel, 用来平衡第一个GPU的显存使用量
★ 229transformer-xl. Python
★ 3.7kInteraction-networks_tensorflow. Tensorflow Implementation of Interaction Networks for Learning about Objects, Relations and Physics
★ 158stable-baselines3. PyTorch version of Stable Baselines, reliable implementations of reinforcement learning algorithms.
★ 14ktrfl. TensorFlow Reinforcement Learning
★ 3.1kscalable_agent. A TensorFlow implementation of Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures.
★ 1kt5_in_bert4keras. 整理一下在keras中使用T5模型的要点
★ 173multilingual-t5. Python
★ 1.3kpytorch-tabular. Some examples of using PyTorch for tabular data
★ 67AdaNorm. Code for "Understanding and Improving Layer Normalization"
★ 46TCN. Sequence modeling benchmarks and temporal convolutional networks
★ 4.5kIBN-Net. Instance-Batch Normalization Networks (ECCV2018)
★ 805tutorials. PyTorch tutorials.
★ 9.3kCollaQ. A code implementation for our arXiv paper "Multi-agent Adhoc Team Play using Decompositional Q function"
★ 132Thought-SC2. (TG'2021) Code for paper "Efficient Reinforcement Learning for StarCraft by Abstract Forward Models and Transfer Learning". TG = Transactions on Games.
★ 11Reinforcement-Implementation. Implementation of benchmark RL algorithms
★ 471ResNeXt-Tensorflow. Simple Tensorflow implementation of ResNeXt using Cifar10
★ 158ResNet-Tensorflow. Simple Tensorflow implementation of pre-activation ResNet18, 34, 50, 101, 152
★ 181UGATIT. Official Tensorflow implementation of U-GAT-IT: Unsupervised Generative Attentional Networks with Adaptive Layer-Instance Normalization for Image-to-Image Translation (ICLR 2020)
★ 6.1ktensorflow2-deep-reinforcement-learning. Code accompanying the blog post "Deep Reinforcement Learning with TensorFlow 2.1"
★ 204reaver. Reaver: Modular Deep Reinforcement Learning Framework. Focused on StarCraft II. Supports Gym, Atari, and MuJoCo.
★ 561atari-reset. Code for the blog post "Learning Montezuma’s Revenge from a Single Demonstration"
★ 207GymGo. An environment of the board game Go using OpenAI's Gym API
★ 176evolution-strategies-starter. Code for the paper "Evolution Strategies as a Scalable Alternative to Reinforcement Learning"
★ 1.6kRL_TSP_4static. Deep Reinforcement Learning for Multiobjective Optimization. Code for this paper
★ 185MORL. Multi-Objective Reinforcement Learning
★ 308WorldModelsExperiments. World Models Experiments
★ 725alphastar. A* algorithm implementation in Python
★ 1SC2AI. 星际2 AI中文教程 StarCraft2 AI with python-sc2/pysc2 API
★ 238muzero-general. MuZero
★ 2.8kMarioZero. Testing MuZero from deepmind with Super Mario Bros
★ 1muzero-pytorch. Pytorch Implementation of MuZero
★ 356muzero. A simple implementation of MuZero algorithm for connect4 game
★ 96MuZero. A structured implementation of MuZero
★ 205TrulyPPO. Python
★ 29pymarl. Python Multi-Agent Reinforcement Learning framework
★ 2.2kMadBot. Starcraft 2 Bot written in Python using the python-sc2 libary
★ 9PurpleWave. StarCraft: Brood War AI in Scala
★ 106SAIDA_RL. C++
★ 66nndl. 邱锡鹏《神经网络与深度学习》(蒲公英书)理论书 v2 与通识版
★ 19kMSC. MSC: A Dataset for Macro-Management in StarCraft II
★ 144smac. SMAC: The StarCraft Multi-Agent Challenge
★ 1.4kMARL-Papers. Paper list of multi-agent reinforcement learning (MARL)
★ 4.9knsfw_data_scraper. Collection of scripts to aggregate image data for the purposes of training an NSFW Image Classifier
★ 13kpaper-gestalt. Deep Paper Gestalt
★ 451softlearning. Softlearning is a reinforcement learning framework for training maximum entropy policies in continuous domains. Includes the official implementation of the Soft Actor-Critic algorithm.
★ 1.4kmcts. An implementation of Monte Carlo Tree Search in python
★ 162homework. Assignments for CS294-112.
★ 1.7kbaselines-rudder. RUDDER for ATARI games with delayed rewards in OpenAI Baselines package
★ 268Locutus. AI for StarCraft: Brood War
★ 84bluebluesky. C++
★ 8SAIDA.
★ 78TorchCraft. Connecting Torch to StarCraft
★ 1.4klarge-scale-curiosity. Code for the paper "Large-Scale Study of Curiosity-Driven Learning"
★ 828random-network-distillation. Code for the paper "Exploration by Random Network Distillation"
★ 930bert. TensorFlow code and pre-trained models for BERT
★ 40koneDNN. oneAPI Deep Neural Network Library (oneDNN)
★ 4kpytorch-scripts. A few Windows specific scripts for PyTorch
★ 411ELF. ELF: a platform for game research with AlphaGoZero/AlphaZero reimplementation
★ 3.4kddp-gym. Differential Dynamic Programming controller operating in OpenAI Gym environment.
★ 87iLQR-REINFORCE-DAGGER. Python
★ 3IRCNN. Learning Deep CNN Denoiser Prior for Image Restoration (CVPR, 2017) (Matlab)
★ 610alpr-unconstrained. License Plate Detection and Recognition in Unconstrained Scenarios
★ 1.8kDetectron.pytorch. A pytorch implementation of Detectron. Both training from scratch and inferring directly from pretrained Detectron weights are available.
★ 2.8kDistributed-TensorFlow-Guide. Distributed TensorFlow basics and examples of training algorithms
★ 641stanford-tensorflow-tutorials. This repository contains code examples for the Stanford's course: TensorFlow for Deep Learning Research.
★ 10ktensorflow-value-iteration-networks. TensorFlow implementation of the Value Iteration Networks (NIPS '16) paper
★ 549VIN. Value Iteration Networks
★ 291crnn. Convolutional Recurrent Neural Network (CRNN) for image-based sequence recognition.
★ 2.1kLicense-Plate-Detect-Recognition-via-Deep-Neural-Networks-accuracy-up-to-99.9. works in real-time with detection and recognition accuracy up to 99.8% for Chinese license plates: 100 ms/plate
★ 1.4kclarity-analyzer. JavaFX-Application to interactively visualize the raw data of a Dota 2, CSGO, CS2 or Deadlock replay.
★ 55dota2dqn. This is a deep Q-network (reinforcement learning) AI for Dota 2
★ 58dota2Bots. God like lina
★ 39Dota2-FullOverwrite. Work in progress for a full-overwrite Dota 2 bot framework
★ 99TSCL. Teacher-Student Curriculum Learning code
★ 85gym-minecraft. Minecraft environment for Open AI Gym, based on Microsoft's Malmo.
★ 280dota2ai. Ranked Matchmaking AI: An improved Dota2 AI based on Valve's default AI. Has more than 3 million subscribers on Steam.
★ 343gps. Guided Policy Search
★ 600Dota2_ai_bot. The infrastructure and the reinforcement learning algorithm for training a Dota 2 AI bot
★ 9dota2ai. Dota2 AI Framework
★ 398pydota2_archive. Dota 2 Python AI
★ 102Dota2_DPPO_bots. Python
★ 14genetics. A python library for genetic algorithms
★ 235deap. Distributed Evolutionary Algorithms in Python
★ 6.4kwargus. Importer and scripts for Warcraft II: Tides of Darkness, the expansion Beyond the Dark Portal, and Aleonas Tales
★ 426high-speed-downloader. 已不再维护
★ 7.1kSNNs. Tutorials and implementations for "Self-normalizing networks"
★ 1.6kml-agents. The Unity Machine Learning Agents Toolkit (ML-Agents) is an open-source project that enables games and simulations to serve as environments for training intelligent agents using deep reinforcement learning and imitation learning.
★ 20kscikit-optimize. Sequential model-based optimization with a `scipy.optimize` interface
★ 2.8kgail_ppo_tf. Tensorflow implementation of Generative Adversarial Imitation Learning(GAIL) with discrete action
★ 114sc2aibot. Implementing reinforcement-learning algorithms for pysc2 -environment
★ 91Tensorflow-Tutorial. Some interesting TensorFlow tutorials for beginners.
★ 894InfoGAN. Code for reproducing key results in the paper "InfoGAN: Interpretable Representation Learning by Information Maximizing Generative Adversarial Nets"
★ 1.1kimitation. Code for the paper "Generative Adversarial Imitation Learning"
★ 729gail-tf. Tensorflow implementation of generative adversarial imitation learning
★ 198PyTorch-RL. PyTorch implementation of Deep Reinforcement Learning: Policy Gradient methods (TRPO, PPO, A2C) and Generative Adversarial Imitation Learning (GAIL). Fast Fisher vector product TRPO.
★ 1.3kADNet. Action-Decision Networks for Visual Tracking with Deep Reinforcement Learning (CVPR 2017)
★ 106dpcluster. Efficient Dirichlet process clustering
★ 26gmm. Gaussian Mixture Models in Python
★ 159irl_rocks. Cool Inverse Reinforcement Learning Papers
★ 124tsc-dl. Visual Transition State Clustering
★ 13tsc. Implements experiments to evaluate transition state clustering
★ 13hierarchical-deep-RL. Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstractions and Intrinsic Motivation
★ 89hierarchical_IL_RL. Code for hierarchical imitation learning and reinforcement learning
★ 301reinforcement_learning. Implementation of selected reinforcement learning algorithms in Tensorflow. A3C, DDPG, REINFORCE, DQN, etc.
★ 153irl-imitation. Implementation of Inverse Reinforcement Learning (IRL) algorithms in Python/Tensorflow. Deep MaxEnt, MaxEnt, LPIRL
★ 678Inverse-Reinforcement-Learning. Implementations of selected inverse reinforcement learning algorithms.
★ 1.1kXX-Net. A proxy tool to bypass GFW.
★ 33kevent-driven-rllab. Extending rllab to event-driven multiagent environments
★ 13