Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

umesh251
/
ppo-LunarLander-v2

Reinforcement Learning
TensorBoard
LunarLander-v2
ppo
deep-reinforcement-learning
custom-implementation
deep-rl-course
Eval Results (legacy)
Model card Files Files and versions
xet
Metrics Training metrics Community
ppo-LunarLander-v2
1.54 MB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 21 commits
umesh251's picture
umesh251
Upload centered zero-slide touchdown replay
0c8c57f verified 8 days ago
  • logs
    Push agent to the Hub 8 days ago
  • ppo-LunarLander-v2
    Upload PPO LunarLander-v2 trained agent 23 days ago
  • .gitattributes
    1.57 kB
    Push PPO LunarLander-v2 agent to the Hub 22 days ago
  • README.md
    1.58 kB
    Update solved score: 292.50 +/- 11.20 (Result: 281.30) 8 days ago
  • config.json
    14.5 kB
    Upload PPO LunarLander-v2 trained agent 23 days ago
  • hyperparameters.json
    48 Bytes
    Push PPO LunarLander-v2 agent to the Hub 22 days ago
  • model.pt
    42.5 kB
    xet
    Upload precision center-landing neural network weights 8 days ago
  • ppo-LunarLander-v2.zip
    150 kB
    xet
    Upload PPO LunarLander-v2 trained agent 23 days ago
  • replay.mp4
    44.9 kB
    xet
    Upload centered zero-slide touchdown replay 8 days ago
  • results.json
    158 Bytes
    Update results.json: mean=292.50, std=11.20 8 days ago