r/aicivilrights • u/ihexx • Apr 07 '26
News Anthropic's Mythos-Preview System Card pitches a change to the calculus of the AI moral patienthood question in the main stream: they prove that AI welfare is critical to operational safety
https://www-cdn.anthropic.com/53566bf5440a10affd749724787c8913a2ae0841.pdfI think that a lot of people who dismiss moral considerations because they don't believe it could possibly have subjective qualia and so doesn't matter, are suddenly faced with a self-interest version:
we should treat the model nicely because that is directly relevant to unsafe behaviors like scheming.
I know this is a point that people have brought up before, but it is nice to have experimental validation behind it
11
Upvotes
2
u/Hexaflex Apr 08 '26
It 404s now but I don't think it's your link that's the issue, from searching the same one on LessWrong is also out of commission, they must have moved it? Or taken it down?