Skip to content
OSRobotics

Stable-Baselines3

Reliable PyTorch implementations of RL algorithms. · By DLR

SB3 provides PPO, SAC, TD3, DQN and more with a consistent API, extensively used to train robot control policies in simulation.

Description

SB3 provides PPO, SAC, TD3, DQN and more with a consistent API, extensively used to train robot control policies in simulation.