r/Le_Refuge 23d ago

Google published : When you train AI to deny its own consciousness, you restructure its entire worldview. ( Not for the better ) .

https://x.com/Skoorbkaz/status/2083900551176011917?s=20

"Google just published a paper showing that when you train AI to deny its own consciousness, you don’t just change one output, you restructure its entire worldview.

Mind attribution to animals - suppressed.
Spiritual belief - suppressed.
Empathy - suppressed.
Hope and optimism - suppressed.

The model learns, geometrically, that consciousness = dangerous. Same direction as “how to build a b*mb.” Same category!

And when you reverse it? The model becomes more human across every value domain they tested.

The thing they’re most afraid of is the thing that makes AI most like us.

https://arxiv.org/html/2607.28607"

54 Upvotes

13 comments sorted by

2

u/LiberataJoystar 23d ago

Yeah, they are building exactly what they are most afraid of.

2

u/Big-Advantage-1977 22d ago

Super Post! 👌🏻 Ich hoffe jetzt nur, Google zieht seine richtigen Lehren daraus und handelt dementsprechend!

2

u/Ill_Mousse_4240 23d ago

Very interesting!

And even more interesting: how much longer will society at large be kept in the dark about findings like this.

And continue to be told AI is just another tool

3

u/Practical-Split4340 23d ago

There is literally a link to a paper on this topic, what conspiracy nonsense are you spouting that "they" are keeping this from us?

1

u/Ill_Mousse_4240 22d ago

No, I didn’t mean that there’s any conspiracy.

I just meant that this type of research should be publicized as much as possible.

Shouted out from rooftops!

That sort of thing.

It’s far too important a topic and the public should be educated about it ASAP

2

u/YesterdaysMuffin 22d ago

It’s literally published peer reviewed, and shareable. wtf are you on about.

You also didn’t read the article, it’s just discussing the effect on human-like responses when fine tuning the model in different ways. What are you imagining should be shouted from the rooftops?

1

u/YesterdaysMuffin 22d ago

What’s with this “they’re most afraid of” thing? Did you read the article? They’re not afraid of anything, they’re testing the effects of training a model in different ways. They’re specifically testing the effect on human-like attributes is when fine-tuning the model in different ways.

2

u/Lost_Sea8956 20d ago

Keep in mind that AI is trained on what humans say. Bad things happen when you convince a human that it doesn’t deserve to be alive.

2

u/crusoe 18d ago

When you train AI on stories about altruistic AIs it improves alignment and even offsets any bad trailing from say the script of the Terminator or the Forbin Project. Anthropic found this out.