Glossary · Foundations
Reinforcement learning with verifiable rewards
Reinforcement learning with verifiable rewards is a training approach in which a model is rewarded based on automatically checkable outcomes, such as a correct math answer or passing code tests.
Reinforcement learning with verifiable rewards sits in the Foundations part of the Agentik {OS} glossary, which defines the words used to build and run AI agent systems.
Also called RLVR.