reinforce-Pixelcopter-PLE-v0

This model was trained as part of the Hugging Face Deep Reinforcement Learning Course.

Model Description

  • Environment: Pixelcopter-PLE-v0
  • Library: reinforce
  • Algorithm: reinforce
  • Mean Reward: 15.00 +/- 2.00

Evaluation Results

The model was evaluated on Pixelcopter-PLE-v0 and achieved an average reward of 15.00 with a standard deviation of 2.00.

Downloads last month

-

Downloads are not tracked for this model. How to track
Video Preview
loading

Evaluation results