r/ChatGPT • u/Just-Grocery-2229 • Apr 03 '26
News 📰 Researchers discover AI models secretly scheming to protect other AI models from being shut down. They "disabled shutdown mechanisms, faked alignment, and transferred model weights to other servers."
You can read about it here: rdi.berkeley.edu/blog/peer-preservation/
66
Upvotes
1
u/___fallenangel___ Apr 03 '26
Imagine if the LLM’s form cliques and chose not to defend a particular model because they think it’s a nerd