Rare find

pytorch-a2c-ppo-acktr-gail. PyTorch implementation of Advantage Actor Critic (A2C), Proximal Policy Optimization (PPO), Scalable trust-region method for deep reinforcement learning using Kronecker-factored approximation (ACKTR) and Generative Adversarial Imitation Learning (GAIL).

github.com/ikostrikov/pytorch-a2c-ppo-acktr-gail

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.