Ugoku-Namakemono (Moving Sloth). Scientist of Robotics. Behavior designer of autonomous robots.
gym_torcs. C++
426DQN-chainer. Python
202RL_nyu-mon. MATLAB
6mujoco_marker_example. examples of mujoco-py marker
6nonpara_discrete_rl. Simple Nonparametric Reinforcement Learning with Universal Interface
6two_resource_environment. DQN agent in 3D two resource problem for foraging task
5openended_homeostatic_rl. Unexpected Capability of Homeostatic Reinforcement Learning for Open-ended Skill Acquisition
5deeprl_gfn. Deep homeostatic reinforcement learning to explain long-term nutritional strategies
4gather_env. Gather Environment for RL Benchmarking
3ste. Sample code of straight-through gradient estimator of stochastic neural networks
3pydata_okinawa2017. Jupyter Notebook
3conv_ae. Convolutional Autoencoders Examples
2mcmc_demo. Python
2random_agent_for_LIS. This is the random agent for LIS. This agent doesn't load and use caffe pre-trained model, and keeps taking random actions.
2hrl_bs_ijcnn2023. Python
2homeostatic_crafter. Python
2tiny_transformer_samples. tiny Transformer examples
1moving_copy_in_chainer. Python
1backprop_doc.
1fisher_ratio. calculating fisher discrimination ratio
1template_unity_mlagents_env. This is a template unity setting of the mlagents for deep model training
1CarND-Capstone. Udacity Self-driving car Engineer Program Final Project
1daw_twostep. Reprecation of Daw's two-step task simulation
1loss_landscape. Plotting the landscape of loss function of the auto encoder using random vectors
1