ppo-Pyramids

This is a trained model of a Reinforcement Learning agent for ML-Agents-Pyramids using ml-agents. Created for the Hugging Face Deep RL Course.

Evaluation Results

  • Mean Reward: 2.00 +/- 0.50
  • Target Environment: ML-Agents-Pyramids
  • Library: ml-agents
  • Model Architecture: PPO (ML-Agents)
Downloads last month
8
Video Preview
loading

Evaluation results