Stable-Baselines3
Reliable PyTorch implementations of RL algorithms. · By DLR
SB3 provides PPO, SAC, TD3, DQN and more with a consistent API, extensively used to train robot control policies in simulation.
Description
SB3 provides PPO, SAC, TD3, DQN and more with a consistent API, extensively used to train robot control policies in simulation.