ppo-Pyramids

This is a trained model of a PPO agent playing ML-Agents-Pyramids. This model was trained as part of the Hugging Face Deep Reinforcement Learning Course.

Evaluation Results

  • Environment: ML-Agents-Pyramids
  • Algorithm: PPO
  • Library: ml-agents
  • Mean Reward: 1.85 +/- 0.15

Usage

To evaluate this model locally or play with it, download the model files from this repository.

Downloads last month
-
Video Preview
loading

Evaluation results