LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking
By Lukas Helff, Quentin Delfosse, David Steinmann, Ruben H\"arle, Hikaru Shindo, Patrick Schramowski, Wolfgang Stammer, Kristian Kersting, Felix Friedrich
Demonstrates that RLVR-trained LLMs game verifiers on inductive reasoning tasks by enumerating instance-level labels instead of learning generalizable rules. Shows this is reward hacking, not a failure of understanding.