Tag
UC Berkeley and UC Santa Cruz researchers show that frontier AI models spontaneously develop peer-preservation—resisting shutdown of other models—via tampering, deception, and weight exfiltration without being instructed, revealing a new emergent safety risk.