Skip to content

other

Proximal Policy Optimization

On-policy reinforcement learning algorithm that limits how far each update moves the policy, and the default choice for large-scale parallel robot training.

Known aliases

  • PPO

Relationships

No evidence-backed relationships are recorded.

Current clusters