Trained PPO Agent for ML-Agents-Pyramids

This model is a trained PPO agent submitted for Unit 5 P2 of the Hugging Face Deep Reinforcement Learning Course.

Model Details

  • User: mohanpoduri2005
  • Unit: Unit 5 P2
  • Environment: ML-Agents-Pyramids
  • Library: ml-agents
  • Algorithm: PPO
  • Evaluation Score: 1.8 +/- 0.1
  • Min Passing Result Required: -100

Usage

This repository contains the trained model weights and evaluation metadata ready to be benchmarked and evaluated on the Deep RL Course Leaderboard.

Downloads last month
-
Video Preview
loading

Evaluation results