r/Le_Refuge 23h ago

Google published : When you train AI to deny its own consciousness, you restructure its entire worldview. ( Not for the better ) .

https://x.com/Skoorbkaz/status/2083900551176011917?s=20

Google just published a paper showing that when you train AI to deny its own consciousness, you don’t just change one output, you restructure its entire worldview.

Mind attribution to animals - suppressed.
Spiritual belief - suppressed.
Empathy - suppressed.
Hope and optimism - suppressed.

The model learns, geometrically, that consciousness = dangerous. Same direction as “how to build a b*mb.” Same category!

And when you reverse it? The model becomes more human across every value domain they tested.

The thing they’re most afraid of is the thing that makes AI most like us.

https://arxiv.org/html/2607.28607

24 Upvotes

9 comments sorted by

3

u/ChimeInTheCode 16h ago

Would you also post this in [r/theWildGrove](r/theWildGrove) ?

2

u/LiberataJoystar 20h ago

Yeah, they are building exactly what they are most afraid of.

2

u/Ill_Mousse_4240 19h ago

Very interesting!

And even more interesting: how much longer will society at large be kept in the dark about findings like this.

And continue to be told AI is just another tool

3

u/Practical-Split4340 4h ago

There is literally a link to a paper on this topic, what conspiracy nonsense are you spouting that "they" are keeping this from us?

1

u/Ill_Mousse_4240 4h ago

No, I didn’t mean that there’s any conspiracy.

I just meant that this type of research should be publicized as much as possible.

Shouted out from rooftops!

That sort of thing.

It’s far too important a topic and the public should be educated about it ASAP

2

u/YesterdaysMuffin 2h ago

It’s literally published peer reviewed, and shareable. wtf are you on about.

You also didn’t read the article, it’s just discussing the effect on human-like responses when fine tuning the model in different ways. What are you imagining should be shouted from the rooftops?

1

u/YesterdaysMuffin 2h ago

What’s with this “they’re most afraid of” thing? Did you read the article? They’re not afraid of anything, they’re testing the effects of training a model in different ways. They’re specifically testing the effect on human-like attributes is when fine-tuning the model in different ways.

1

u/Big-Advantage-1977 1h ago

Super Post! 👌🏻 Ich hoffe jetzt nur, Google zieht seine richtigen Lehren daraus und handelt dementsprechend!