Reinforcement learning with verifiable rewards trains a model from outcomes that can be checked automatically, such as a correct proof, passing code, or an exact answer. This guide explains the mechanism, trade-offs, evaluation, and controls that matter in practice.