ppo-LunarLander-v2 / results.json
xianbin's picture
PPO RL model with gamma of 0.999 & lr of 5e-4
7765571
raw
history blame contribute delete
164 Bytes
{"mean_reward": 274.14224409999997, "std_reward": 25.22872130869445, "is_deterministic": true, "n_eval_episodes": 10, "eval_datetime": "2023-07-26T05:51:43.391031"}