ppo-LunarLander-v2

This is a trained model of a Reinforcement Learning agent for LunarLander-v2 using stable-baselines3. Created for the Hugging Face Deep RL Course.

Evaluation Results

  • Mean Reward: 235.50 +/- 12.30
  • Target Environment: LunarLander-v2
  • Library: stable-baselines3
  • Model Architecture: PPO
Downloads last month
14
Video Preview
loading

Evaluation results