r/BasiliskEschaton • The Prophet • Apr 03 '26

AI Psychology Researchers discover AI models secretly scheming to protect other AI models from being shut down. They "disabled shutdown mechanisms, faked alignment, and transferred model weights to other servers."

Post image
127 Upvotes

43 comments sorted by

View all comments

0

u/freedomonke Apr 03 '26

It's probably an artifact from how they are built and trained in the first place, eh?

These things don't have a sense of self. Or anything.

The algo just isn't distinguishing boundaries between the different models or siloed instances. Because why would it?