APLab1545/ppo-LunarLander-v2
084
PPO Agent Playing LunarLander-v3
This is a trained model of a PPO agent playing LunarLander-v3 using the stable-baselines3 library.
Evaluation Results
- Mean Reward: -232.00 +/- 73.03
This is a trained model of a PPO agent playing LunarLander-v3 using the stable-baselines3 library.