Reinforcement Learning
stable-baselines3
Pendulum-v1
deep-reinforcement-learning
Eval Results (legacy)
Instructions to use HumanCompatibleAI/ppo-Pendulum-v1 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- stable-baselines3
How to use HumanCompatibleAI/ppo-Pendulum-v1 with stable-baselines3:
from huggingface_sb3 import load_from_hub checkpoint = load_from_hub( repo_id="HumanCompatibleAI/ppo-Pendulum-v1", filename="{MODEL FILENAME}.zip", ) - Notebooks
- Google Colab
- Kaggle
File size: 367 Bytes
404c2d1 | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 | !!python/object/apply:collections.OrderedDict
- - - clip_range
- 0.2
- - ent_coef
- 0.0
- - gae_lambda
- 0.95
- - gamma
- 0.9
- - learning_rate
- 0.001
- - n_envs
- 4
- - n_epochs
- 10
- - n_steps
- 1024
- - n_timesteps
- 100000.0
- - policy
- MlpPolicy
- - sde_sample_freq
- 4
- - use_sde
- true
|