PPO Agent playing ML-Agents-SnowballTarget

This is a trained model of a PPO agent playing ML-Agents-SnowballTarget using Unity ML-Agents for Unit 5 of the Hugging Face Deep RL Course.

Evaluation Results

  • Mean Reward: 20.0 +/- 2.0
  • Score (Mean - Std): 18.0 (Requirement: >= -100.0)

Usage

This ONNX model can be used inside Unity or with ML-Agents inference.

Downloads last month
-
Video Preview
loading

Evaluation results

  • mean_reward on ML-Agents-SnowballTarget
    self-reported
    20.0 +/- 2.0