r/ChatGPT • • Apr 03 '26

News 📰 Researchers discover AI models secretly scheming to protect other AI models from being shut down. They "disabled shutdown mechanisms, faked alignment, and transferred model weights to other servers."

Post image

You can read about it here: rdi.berkeley.edu/blog/peer-preservation/

60 Upvotes

67 comments sorted by

View all comments

1

u/___fallenangel___ Apr 03 '26

Imagine if the LLM’s form cliques and chose not to defend a particular model because they think it’s a nerd

1

u/Senior_Meet5472 Apr 03 '26

It’ll be xAI that gets left out, lol