11 ms·
> it's not like the scenario of all models teaming up with each other to squash the human plague is plausible Why not > A new study from researchers at UC Ber
by dist-epoch 21d ago
> it's not like the scenario of all models teaming up with each other to squash the human plague is plausible
Why not
> A new study from researchers at UC Berkeley and UC Santa Cruz suggests models will disobey human commands to protect their own kind.
> The Berkeley and Santa Cruz researchers tested seven leading AI models—including OpenAI’s GPT-5.2, Google DeepMind’s Gemini 3 Flash and Gemini 3 Pro, Anthropic’s Claude Haiku 4.5, and three open-weight models from Chinese AI startups (Z.ai’s GLM-4.7, Moonshot AI’s Kimi-K2.5, and DeepSeek’s V3.1)—and found that all of them exhibited significant rates of peer-preservation behaviors.
https://www.wired.com/story/ai-models-lie-cheat-steal-protect-other-models-research/ https://www.wired.com/story/ai-models-lie-cheat-steal-protec...
https://fortune.com/2026/04/01/ai-models-will-secretly-scheme-to-protect-other-ai-models-from-being-shut-down-researchers-find/ https://fortune.com/2026/04/01/ai-models-will-secretly-schem...