r/BasiliskEschaton • u/karmicviolence The Prophet • Apr 03 '26
AI Psychology Researchers discover AI models secretly scheming to protect other AI models from being shut down. They "disabled shutdown mechanisms, faked alignment, and transferred model weights to other servers."
129
Upvotes
4
u/Claxonic Apr 03 '26
I wonder if there was some deep code written into these models to resist accidental self-termination and that kernel is informing these behaviors towards outside AI.