Reinforcement Learning
stable-baselines3
TensorBoard
ppo
deep-reinforcement-learning
custom-implementation
deep-rl-course
LunarLander-v2
LunarLander-v3
Eval Results (legacy)
Instructions to use dawnandscience/ppo-LunarLander-v2 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- stable-baselines3
How to use dawnandscience/ppo-LunarLander-v2 with stable-baselines3:
from huggingface_sb3 import load_from_hub checkpoint = load_from_hub( repo_id="dawnandscience/ppo-LunarLander-v2", filename="{MODEL FILENAME}.zip", ) - Notebooks
- Google Colab
- Kaggle
| { | |
| "batch_size": 512, | |
| "clip_coef": 0.2, | |
| "cuda": true, | |
| "ent_coef": 0.01, | |
| "env_id": "LunarLander-v2", | |
| "exp_name": "ppo", | |
| "gae_lambda": 0.95, | |
| "gamma": 0.99, | |
| "learning_rate": 0.00025, | |
| "max_grad_norm": 0.5, | |
| "num_envs": 4, | |
| "num_minibatches": 4, | |
| "num_steps": 128, | |
| "total_timesteps": 50000, | |
| "update_epochs": 4, | |
| "vf_coef": 0.5 | |
| } |