Trained PPO Agent for LunarLander-v2
This model is a trained PPO agent submitted for Unit 8 PI of the Hugging Face Deep Reinforcement Learning Course.
Model Details
- User: mohanpoduri2005
- Unit: Unit 8 PI
- Environment:
LunarLander-v2 - Library:
deep-rl-course - Algorithm: PPO
- Evaluation Score:
260.0 +/- 15.0 - Min Passing Result Required:
-500
Usage
This repository contains the trained model weights and evaluation metadata ready to be benchmarked and evaluated on the Deep RL Course Leaderboard.
- Downloads last month
- -
Evaluation results
- mean_reward on LunarLander-v2self-reported260.0 +/- 15.0