deep-rl-LunarLander-v2
This model was trained as part of the Hugging Face Deep Reinforcement Learning Course.
Model Description
- Environment:
LunarLander-v2 - Library:
deep-rl-course - Algorithm:
ppo - Mean Reward:
200.00 +/- 10.00
Evaluation Results
The model was evaluated on LunarLander-v2 and achieved an average reward of 200.00 with a standard deviation of 10.00.
- Downloads last month
- 14
Evaluation results
- mean_reward on LunarLander-v2self-reported200.00 +/- 10.00