reinforce-Pixelcopter-PLE-v0
This model was trained as part of the Hugging Face Deep Reinforcement Learning Course.
Model Description
- Environment:
Pixelcopter-PLE-v0 - Library:
reinforce - Algorithm:
reinforce - Mean Reward:
15.00 +/- 2.00
Evaluation Results
The model was evaluated on Pixelcopter-PLE-v0 and achieved an average reward of 15.00 with a standard deviation of 2.00.
Evaluation results
- mean_reward on Pixelcopter-PLE-v0self-reported15.00 +/- 2.00