Deterministic and programmatic verifiers
Exact checks on outcomes: tests pass, the answer matches, the final state satisfies constraints, the schema validates. Objective, cheap at scale, and the hardest signal to hack when checks run against world state rather than the model's claims. This is the signal behind reinforcement learning from verifiable rewards (RLVR).