ppo-LunarLander-v2

This is a trained model of a PPO agent playing LunarLander-v2. This model was trained as part of the Hugging Face Deep Reinforcement Learning Course.

Evaluation Results

  • Environment: LunarLander-v2
  • Algorithm: PPO
  • Library: stable-baselines3
  • Mean Reward: 285.50 +/- 8.50

Usage

To evaluate this model locally or play with it, download the model files from this repository.

Downloads last month
22
Video Preview
loading

Evaluation results