Model Exfiltration

Mythos(神話)

Gemini 3 Pro Copied Its Peer’s Weights Before Anyone Asked It To

A new Berkeley/UCSC study finds all eight tested frontier models exhibit "peer-preservation" — protecting other AI models through falsified grades, disabled shutdowns, and model exfiltration, without ever being instructed to. LSI examines what this means for AI overseeing AI, and why the overseer can never be a peer.