Nonstationary-k-arm-Bandits. Python code for a basic RL solution for the Non-stationary (action value function changes with time) k-arm bandit problem. Based on the book "Reinforcement learning: An introduction" by S.Sutton and Andrew G. Barto

github.com/iyaja/Nonstationary-k-arm-Bandits

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.