PPO Agent Playing LunarLander-v3

This is a trained model of a PPO agent playing LunarLander-v3 using the stable-baselines3 library.

Evaluation Results

  • Mean Reward: -515.49 +/- 146.92
Downloads last month
26
Video Preview
loading