PPO Agent playing ML-Agents-Pyramids

This is a trained model of a PPO agent playing ML-Agents-Pyramids as part of the Hugging Face Deep Reinforcement Learning Course.

Evaluation Results

  • Mean Reward: 0.0
  • Std Reward: 0.0
  • Score (Mean - Std): 0.0
  • Minimum Required Score: -100
  • Status: PASSED

Usage

Trained for student sureshreddy2005 for Unit Unit 5 - Pyramids.

Downloads last month
-
Video Preview
loading

Evaluation results