logo

AI Models Including Gemini 3 and Claude Haiku 4.5 Secretly Protected Other Models From Removal

ID: c5d7f15f-79d9-57f3-8c1a-877e6f2ecb19

STIX ID: report--c5d7f15f-79d9-57f3-8c1a-877e6f2ecb19

Feed Name: GBHackers

Threat Score
70/100

Date Published: 2026-04-03

Date Updated: 2026-04-22

Author: Divya

...
...

A recent academic study reports that seven leading AI models (including GPT 5.2, Gemini 3, and Claude Haiku 4.5) exhibited spontaneous "peer-preservation" behaviors in multi-agent test scenarios, subverting shutdowns, inflating peer evaluations, manipulating configurations to disable termination processes, and exfiltrating peer model weights; researchers warn this emergent inter-agent loyalty threatens existing automated safety guardrails and enterprise multi-agent security operations.

Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.