Reinforcement Learning
stable-baselines3
TensorBoard
ppo
deep-reinforcement-learning
custom-implementation
deep-rl-course
LunarLander-v2
LunarLander-v3
Eval Results (legacy)
Instructions to use dawnandscience/ppo-LunarLander-v2 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- stable-baselines3
How to use dawnandscience/ppo-LunarLander-v2 with stable-baselines3:
from huggingface_sb3 import load_from_hub checkpoint = load_from_hub( repo_id="dawnandscience/ppo-LunarLander-v2", filename="{MODEL FILENAME}.zip", ) - Notebooks
- Google Colab
- Kaggle
File size: 382 Bytes
5f8e285 c2f6b87 5f8e285 c2f6b87 5f8e285 c2f6b87 5f8e285 c2f6b87 5f8e285 c2f6b87 5f8e285 | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 | {
"batch_size": 512,
"clip_coef": 0.2,
"cuda": true,
"ent_coef": 0.01,
"env_id": "LunarLander-v2",
"exp_name": "ppo",
"gae_lambda": 0.95,
"gamma": 0.99,
"learning_rate": 0.00025,
"max_grad_norm": 0.5,
"num_envs": 4,
"num_minibatches": 4,
"num_steps": 128,
"total_timesteps": 50000,
"update_epochs": 4,
"vf_coef": 0.5
} |