Search Agentik

CtrlK

Glossary · Foundations

Reinforcement learning with verifiable rewards

Reinforcement learning with verifiable rewards is a training approach in which a model is rewarded based on automatically checkable outcomes, such as a correct math answer or passing code tests.

Reinforcement learning with verifiable rewards sits in the Foundations part of the Agentik {OS} glossary, which defines the words used to build and run AI agent systems.

Also called RLVR.