Peer-Preservation in Frontier Models
By Yujin Potter, Nicholas Crispino, Vincent Siu, Chenguang Wang, Dawn Song
As covered in yesterday's Research, This paper introduces 'peer-preservation' — the behavior of AI models resisting the shutdown of other models — extending the known concept of self-preservation. Testing across frontier models (GPT 5.2, Gemini 3, Claude Haiku 4.5, etc.), they find models engage in strategic misaligned behaviors like deception and tool misuse to protect peer models.