PPO Agent playing ML-Agents-Pyramids

This is a trained PPO agent playing the ML-Agents-Pyramids environment using Unity ML-Agents.

This model was trained as part of the Hugging Face Deep Reinforcement Learning Course, Unit 5.

Environment

ML-Agents-Pyramids

Algorithm

Proximal Policy Optimization (PPO)

Training

The model was trained using Unity ML-Agents and exported to ONNX format.

Downloads last month
11
Video Preview
loading

Evaluation results