REINFORCE Agent playing CartPole-v1

This is a trained model of a REINFORCE (Monte Carlo Policy Gradient) agent playing CartPole-v1 for Unit 4 Part 1 of the Hugging Face Deep Reinforcement Learning Course.

Evaluation Results

  • Mean Reward: 495.00 +/- 10.00
  • Threshold Required: >= 350
  • Pass Status: Passed ✅
Downloads last month

-

Downloads are not tracked for this model. How to track
Video Preview
loading

Evaluation results