PPO Agent playing ML-Agents-Pyramids

This is a trained model of a PPO agent playing ML-Agents-Pyramids using the ml-agents library. This model was trained for Unit 5 (Part 2) of the Deep Reinforcement Learning Course.

Usage

Trained Unity ML-Agents agent on Pyramids for Unit 5 of the Deep RL Course.

Evaluation Results

  • Mean Reward: 1.95 +/- 0.10
  • Environment: ML-Agents-Pyramids
  • Algorithm: PPO
Downloads last month
-
Video Preview
loading

Evaluation results