PPO Agent playing ML-Agents-Pyramids

This is a trained model of a PPO agent playing ML-Agents-Pyramids using Unity ML-Agents for Unit 5 of the Hugging Face Deep RL Course.

Evaluation Results

  • Mean Reward: 2.0 +/- 0.1
  • Score (Mean - Std): 1.9 (Requirement: >= -100.0)

Usage

This ONNX model can be used inside Unity or with ML-Agents inference.

Downloads last month
4
Video Preview
loading

Evaluation results