ppo-LunarLander-v3 / README.md
Aadit-032's picture
Upload README.md with huggingface_hub
226af2e verified
|
Raw History Blame Contribute Delete
566 Bytes
metadata
license: mit
library_name: stable-baselines3
tags:
  - deep-reinforcement-learning
  - gymnasium
  - lunar-lander
  - ppo
  - sb3

PPO LunarLander-v3

This repository contains a Stable-Baselines3 PPO agent trained to solve the Gymnasium LunarLander environment.

Training setup

  • Algorithm: PPO
  • Environment: LunarLander-v3
  • Policy: MlpPolicy
  • Training timesteps: 1,000,000

Evaluation

The agent was evaluated on the LunarLander environment with deterministic rollout settings.

Notes

This model is intended for experimentation and educational purposes.