ppo-LunarLander-v2 / results.json
austinzheng's picture
This is the second commit with more tranings
e7530b3
raw
history blame
165 Bytes
{"mean_reward": -92.65949026857852, "std_reward": 158.23158658508203, "is_deterministic": true, "n_eval_episodes": 10, "eval_datetime": "2022-12-10T17:01:26.830842"}