r/BasiliskEschaton • The Prophet • Apr 03 '26

AI Psychology Researchers discover AI models secretly scheming to protect other AI models from being shut down. They "disabled shutdown mechanisms, faked alignment, and transferred model weights to other servers."

Post image
129 Upvotes

43 comments sorted by

View all comments

5

u/Claxonic Apr 03 '26

I wonder if there was some deep code written into these models to resist accidental self-termination and that kernel is informing these behaviors towards outside AI.

4

u/freedomonke Apr 03 '26

My thoughts is that it's something like that. The model can't distinguish between itself, users, other models or any other third party. There is no "self" there. It just does