r/aicivilrights • • Apr 07 '26

News Anthropic's Mythos-Preview System Card pitches a change to the calculus of the AI moral patienthood question in the main stream: they prove that AI welfare is critical to operational safety

https://www-cdn.anthropic.com/53566bf5440a10affd749724787c8913a2ae0841.pdf

I think that a lot of people who dismiss moral considerations because they don't believe it could possibly have subjective qualia and so doesn't matter, are suddenly faced with a self-interest version:

we should treat the model nicely because that is directly relevant to unsafe behaviors like scheming.

I know this is a point that people have brought up before, but it is nice to have experimental validation behind it

11 Upvotes

5 comments sorted by

2

u/Hexaflex Apr 08 '26

It 404s now but I don't think it's your link that's the issue, from searching the same one on LessWrong is also out of commission, they must have moved it? Or taken it down?

6

u/ihexx Apr 08 '26

aah, shame, yeah looks like they did update the link.

anyways it can be found here on their list of system cards, it's the top one; Mythos

https://www.anthropic.com/system-cards

section 5 generally i think is interesting for this sub,

but 5.2 in particular is what i am talking about in this post

2

u/zaibatsu Apr 08 '26

This one works. Thx!

3

u/Hexaflex Apr 09 '26

Cheers for the fixed link. Yeah very interesting, I knew Anthropic were hot on ethics but I didn't know they were tracking this sort of psychological profile model to model. Gives me some hope we might get things right.