Reinforce Agent playing {CartPole-v1}

This is a trained model of a Reinforce agent playing {env_id} . To learn to use this model and train yours check Unit 4 of the Deep Reinforcement Learning Course: https://huggingface.co/deep-rl-course/unit4/introduction

Usage (with Stable-baselines3)

def load_policy(ckpt_path, s_size, a_size, h_size, device="cpu"):
    policy = Policy(s_size, a_size, h_size).to(device)
    policy.load_state_dict(torch.load(ckpt_path, map_location=device))
    policy.eval()
    return policy


from stable_baselines3 import ...
from huggingface_sb3 import load_from_hub
from inference import load_policy()

...
Downloads last month
-
Video Preview
loading

Evaluation results