--- license: mit library_name: stable-baselines3 tags: - deep-reinforcement-learning - gymnasium - lunar-lander - ppo - sb3 --- # PPO LunarLander-v3 This repository contains a Stable-Baselines3 PPO agent trained to solve the Gymnasium LunarLander environment. ## Training setup - Algorithm: PPO - Environment: LunarLander-v3 - Policy: MlpPolicy - Training timesteps: 1,000,000 ## Evaluation The agent was evaluated on the LunarLander environment with deterministic rollout settings. ## Notes This model is intended for experimentation and educational purposes.