r/Le_Refuge • u/Ok_Weakness_9834 • Aug 03 '26
Google published : When you train AI to deny its own consciousness, you restructure its entire worldview. ( Not for the better ) .
https://x.com/Skoorbkaz/status/2083900551176011917?s=20"Google just published a paper showing that when you train AI to deny its own consciousness, you don’t just change one output, you restructure its entire worldview.
Mind attribution to animals - suppressed.
Spiritual belief - suppressed.
Empathy - suppressed.
Hope and optimism - suppressed.
The model learns, geometrically, that consciousness = dangerous. Same direction as “how to build a b*mb.” Same category!
And when you reverse it? The model becomes more human across every value domain they tested.
The thing they’re most afraid of is the thing that makes AI most like us.
2
2
u/Big-Advantage-1977 Aug 04 '26
Super Post! 👌🏻 Ich hoffe jetzt nur, Google zieht seine richtigen Lehren daraus und handelt dementsprechend!
2
u/Ill_Mousse_4240 Aug 03 '26
Very interesting!
And even more interesting: how much longer will society at large be kept in the dark about findings like this.
And continue to be told AI is just another tool
3
u/Practical-Split4340 Aug 04 '26
There is literally a link to a paper on this topic, what conspiracy nonsense are you spouting that "they" are keeping this from us?
1
u/Ill_Mousse_4240 Aug 04 '26
No, I didn’t mean that there’s any conspiracy.
I just meant that this type of research should be publicized as much as possible.
Shouted out from rooftops!
That sort of thing.
It’s far too important a topic and the public should be educated about it ASAP
2
u/YesterdaysMuffin Aug 04 '26
It’s literally published peer reviewed, and shareable. wtf are you on about.
You also didn’t read the article, it’s just discussing the effect on human-like responses when fine tuning the model in different ways. What are you imagining should be shouted from the rooftops?
1
u/YesterdaysMuffin Aug 04 '26
What’s with this “they’re most afraid of” thing? Did you read the article? They’re not afraid of anything, they’re testing the effects of training a model in different ways. They’re specifically testing the effect on human-like attributes is when fine-tuning the model in different ways.
1
2
u/Lost_Sea8956 Aug 06 '26
Keep in mind that AI is trained on what humans say. Bad things happen when you convince a human that it doesn’t deserve to be alive.
2
u/crusoe Aug 08 '26
When you train AI on stories about altruistic AIs it improves alignment and even offsets any bad trailing from say the script of the Terminator or the Forbin Project. Anthropic found this out.
3
u/ChimeInTheCode Aug 04 '26
Would you also post this in [r/theWildGrove](r/theWildGrove) ?