Kyoto, Japan

naoto yoshida

Expert
@ugo-nama-kun

Ugoku-Namakemono (Moving Sloth). Scientist of Robotics. Behavior designer of autonomous robots.

gym_torcs. C++

426

DQN-chainer. Python

202

RL_nyu-mon. MATLAB

6

mujoco_marker_example. examples of mujoco-py marker

6

nonpara_discrete_rl. Simple Nonparametric Reinforcement Learning with Universal Interface

6

two_resource_environment. DQN agent in 3D two resource problem for foraging task

5

openended_homeostatic_rl. Unexpected Capability of Homeostatic Reinforcement Learning for Open-ended Skill Acquisition

5

deeprl_gfn. Deep homeostatic reinforcement learning to explain long-term nutritional strategies

4

gather_env. Gather Environment for RL Benchmarking

3

ste. Sample code of straight-through gradient estimator of stochastic neural networks

3

pydata_okinawa2017. Jupyter Notebook

3

conv_ae. Convolutional Autoencoders Examples

2

mcmc_demo. Python

2

random_agent_for_LIS. This is the random agent for LIS. This agent doesn't load and use caffe pre-trained model, and keeps taking random actions.

2

hrl_bs_ijcnn2023. Python

2

homeostatic_crafter. Python

2

tiny_transformer_samples. tiny Transformer examples

1

moving_copy_in_chainer. Python

1

backprop_doc.

1

fisher_ratio. calculating fisher discrimination ratio

1

template_unity_mlagents_env. This is a template unity setting of the mlagents for deep model training

1

CarND-Capstone. Udacity Self-driving car Engineer Program Final Project

1

daw_twostep. Reprecation of Daw's two-step task simulation

1

loss_landscape. Plotting the landscape of loss function of the auto encoder using random vectors

1
24
Apply