ppo-SpaceInvadersNoFrameskip-v4

This model was trained as part of the Hugging Face Deep Reinforcement Learning Course.

Model Description

  • Environment: SpaceInvadersNoFrameskip-v4
  • Library: stable-baselines3
  • Algorithm: ppo
  • Mean Reward: 300.00 +/- 20.00

Evaluation Results

The model was evaluated on SpaceInvadersNoFrameskip-v4 and achieved an average reward of 300.00 with a standard deviation of 20.00.

Downloads last month
6
Video Preview
loading

Evaluation results

  • mean_reward on SpaceInvadersNoFrameskip-v4
    self-reported
    300.00 +/- 20.00