ppo-ML-Agents-Pyramids

This model was trained as part of the Hugging Face Deep Reinforcement Learning Course.

Model Description

  • Environment: ML-Agents-Pyramids
  • Library: ml-agents
  • Algorithm: ppo
  • Mean Reward: 0.00 +/- 0.00

Evaluation Results

The model was evaluated on ML-Agents-Pyramids and achieved an average reward of 0.00 with a standard deviation of 0.00.

Downloads last month
9
Video Preview
loading

Evaluation results