a2c-PandaReachDense-v3

This is a trained model of a A2C agent playing PandaReachDense-v3. This model was trained as part of the Hugging Face Deep Reinforcement Learning Course.

Evaluation Results

  • Environment: PandaReachDense-v3
  • Algorithm: A2C
  • Library: stable-baselines3
  • Mean Reward: -0.45 +/- 0.05

Usage

To evaluate this model locally or play with it, download the model files from this repository.

Downloads last month
4
Video Preview
loading

Evaluation results