reinforce-CartPole-v1
This model was trained as part of the Hugging Face Deep Reinforcement Learning Course.
Model Description
- Environment:
CartPole-v1 - Library:
reinforce - Algorithm:
reinforce - Mean Reward:
500.00 +/- 10.00
Evaluation Results
The model was evaluated on CartPole-v1 and achieved an average reward of 500.00 with a standard deviation of 10.00.
Evaluation results
- mean_reward on CartPole-v1self-reported500.00 +/- 10.00