A2C agent playing PandaReachDense-v3

This is a trained model of an A2C agent playing PandaReachDense-v3 for the Hugging Face Deep Reinforcement Learning Course (Unit 6).

Evaluation Results

  • Mean Reward: -0.24 +/- 0.14
  • Environment: PandaReachDense-v3
  • Algorithm: A2C
  • Library: stable-baselines3

Usage

Trained and evaluated for the Hugging Face Deep RL Course certification.

Downloads last month
5
Video Preview
loading

Evaluation results