Today for AI
Explore indexBack to Terms

Policy Gradient

An RL method that performs gradient ascent directly on policy parameters to maximize expected return.

No public content is connected to this entity yet.