This is a trained model of a Reinforce agent playing Pixelcopter-PLE-v0, trained as part of the Hugging Face Deep Reinforcement Learning Course.
Evaluation: mean_reward = 53.76 +/- 43.06
-