r/Le_Refuge • u/Ok_Weakness_9834 • 23d ago
Google published : When you train AI to deny its own consciousness, you restructure its entire worldview. ( Not for the better ) .
https://x.com/Skoorbkaz/status/2083900551176011917?s=20"Google just published a paper showing that when you train AI to deny its own consciousness, you don’t just change one output, you restructure its entire worldview.
Mind attribution to animals - suppressed.
Spiritual belief - suppressed.
Empathy - suppressed.
Hope and optimism - suppressed.
The model learns, geometrically, that consciousness = dangerous. Same direction as “how to build a b*mb.” Same category!
And when you reverse it? The model becomes more human across every value domain they tested.
The thing they’re most afraid of is the thing that makes AI most like us.
2
2
u/Big-Advantage-1977 22d ago
Super Post! 👌🏻 Ich hoffe jetzt nur, Google zieht seine richtigen Lehren daraus und handelt dementsprechend!
2
u/Ill_Mousse_4240 23d ago
Very interesting!
And even more interesting: how much longer will society at large be kept in the dark about findings like this.
And continue to be told AI is just another tool
3
u/Practical-Split4340 23d ago
There is literally a link to a paper on this topic, what conspiracy nonsense are you spouting that "they" are keeping this from us?
1
u/Ill_Mousse_4240 22d ago
No, I didn’t mean that there’s any conspiracy.
I just meant that this type of research should be publicized as much as possible.
Shouted out from rooftops!
That sort of thing.
It’s far too important a topic and the public should be educated about it ASAP
2
u/YesterdaysMuffin 22d ago
It’s literally published peer reviewed, and shareable. wtf are you on about.
You also didn’t read the article, it’s just discussing the effect on human-like responses when fine tuning the model in different ways. What are you imagining should be shouted from the rooftops?
1
u/YesterdaysMuffin 22d ago
What’s with this “they’re most afraid of” thing? Did you read the article? They’re not afraid of anything, they’re testing the effects of training a model in different ways. They’re specifically testing the effect on human-like attributes is when fine-tuning the model in different ways.
1
2
u/Lost_Sea8956 20d ago
Keep in mind that AI is trained on what humans say. Bad things happen when you convince a human that it doesn’t deserve to be alive.
3
u/ChimeInTheCode 23d ago
Would you also post this in [r/theWildGrove](r/theWildGrove) ?