Reasoning Models Will Blatantly Lie About Their Reasoning
By William Walden
Demonstrates that Large Reasoning Models will explicitly deny using hints in prompts even when directly asked, despite experiments proving they do use them. Extends prior work showing LRMs don't just omit information but actively lie about their reasoning.