PPO Agent playing Pendulum v1 This is a trained model of a PPO agent playing Pendulum v1 using the stable baselines3 library and the RL Zoo. The RL Zoo is a training framework for Stable Baselines3 reinforcement learning agents, with hyperparameter optimization and pre trained agents included. Usage (with SB3 RL Zoo) RL Zoo: https://github.com/DLR RM/rl baselines3 zoo SB3: https://github.com/DLR RM/stable baselines3 SB3 Contrib: https://github.com/Stable Baselines Team/stable baselines3 contrib Install the RL Zoo (with SB3 and SB3 Contrib): If you installed the RL Zoo3 via pip ( pip install rl zoo3 ), from anywhere you can do: Training (with the RL Zoo) Hyperparameters Environment Arguments
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy