REINFORCE agent playing CartPole-v1
This is a trained model of an REINFORCE agent playing CartPole-v1 for the Hugging Face Deep Reinforcement Learning Course (Unit 4 P1).
Evaluation Results
- Mean Reward: 500.00 +/- 0.00
- Environment: CartPole-v1
- Algorithm: REINFORCE
- Library: reinforce
Usage
Trained and evaluated for the Hugging Face Deep RL Course certification.
Evaluation results
- mean_reward on CartPole-v1self-reported500.00 +/- 0.00