bstad's picture
Upload PPO LunarLander-v2 trained agent with n_envs = 32
26fa895
raw
history blame contribute delete
No virus
163 Bytes
{"mean_reward": 149.0667677429552, "std_reward": 88.31244782073007, "is_deterministic": true, "n_eval_episodes": 10, "eval_datetime": "2022-05-13T22:37:13.107619"}