REINFORCE Agent playing CartPole-v1
This is a trained model of a REINFORCE (Monte Carlo Policy Gradient) agent playing CartPole-v1 for Unit 4 Part 1 of the Hugging Face Deep Reinforcement Learning Course.
Evaluation Results
- Mean Reward: 495.00 +/- 10.00
- Threshold Required: >= 350
- Pass Status: Passed ✅
Evaluation results
- mean_reward on CartPole-v1self-reported495.00 +/- 10.00