ppo-Pyramids

This is a trained model of a PPO agent playing ML-Agents-Pyramids.

  • Unit: Unit 5 of the Hugging Face Deep RL Course.
  • Library: ml-agents
  • Algorithm: PPO
  • Environment: ML-Agents-Pyramids
  • Mean Reward: 1.85 +/- 0.05

Description

Trained Unity ML-Agents PPO agent on Pyramids for Hugging Face Deep RL Course Unit 5.

Usage

This agent was uploaded as part of the Hugging Face Deep Reinforcement Learning Course assignments by Surya198382.

Downloads last month
-
Video Preview
loading

Evaluation results