Explore indexBack to Terms
Reward Alignment
Adjusting model behavior via reinforcement learning or preference data to align outputs with human values or specific goals.
No public content is connected to this entity yet.
Adjusting model behavior via reinforcement learning or preference data to align outputs with human values or specific goals.
No public content is connected to this entity yet.