REINFORCE Agent playing Pixelcopter-PLE-v0
This is a trained model of a REINFORCE agent playing Pixelcopter-PLE-v0 for Unit 4 Part 2 of the Hugging Face Deep Reinforcement Learning Course.
Evaluation Results
- Mean Reward: 14.50 +/- 3.20
- Threshold Required: >= 5
- Pass Status: Passed ✅