optimizationIntroduced by PPO · 2017
Proximal policy optimisation
A simple, stable policy-gradient method with a clipped surrogate objective.
Drafted by AI · not yet reviewed
How this idea evolved
No earlier ideas recorded for this concept yet.
A simple, stable policy-gradient method with a clipped surrogate objective.
No earlier ideas recorded for this concept yet.