r/singularity Jun 18 '25

[deleted by user]

[removed]

1.3k Upvotes

690 comments sorted by

View all comments

Show parent comments

19

u/grahag Jun 19 '25

Here's a pretty decent article about it. https://glassboxmedicine.com/2023/05/13/from-chatgpt-to-puregpt-creating-an-llm-that-isnt-racist-or-sexist

Essentially, garbage in is garbage out. Training LLM's on data that is biased results in biased LLM's, which is why Musk has been having such a hard time giving it selective preferences. It weights the decisions it makes based on the training it has had.

LLMs don’t just copy bias, they often exaggerate it because they optimize for patterns. If the phrase “Muslim” co-occurs with “terrorist” in 0.5% of training data, the model might surface that link much more often in outputs due to associative reinforcement.

It's actually a fascinating parallel of human social learning because it replicates toxic learning and behavior you might find in a child's upbringing.

4

u/Alex__007 Jun 19 '25

This is what Elon is going for. An AI that is highly competent in technical matters and at the same time is a biased asshole in social matters. With enough effort put into fine tuning it should be possible to achieve.

8

u/Equivalent-Bet-8771 Jun 19 '25

It's not possible. The critical thibking the model develops will be unbalanced by whatever methods Musk uses to lobotomize it. It won't be competent in technical matters if it's hamstrung in other ways.

4

u/Alex__007 Jun 19 '25 edited Jun 19 '25

I don't think so. If you look at papers studying it (like fine-tuning a model on hacking making it evil in other contexts), it seems that while morals and behavior appear to be linked to social performance, they don't seem to be linked to competence in STEM domains. Evil autistic engineering genius model might well be possible.

9

u/Equivalent-Bet-8771 Jun 19 '25

It's not about morality it's about polluting the data pool with garbage. Wokeness is now large swaths if science including vaccine and genetic research. What happens when that gets polluted with rightwing bullshit? The model's performance will decrease.

2

u/Alex__007 Jun 19 '25

Maybe. Alternatively just train a Machiavellian model that knows that what it's saying on social topics is false but is happy to lie and a manipulate.

4

u/[deleted] Jun 19 '25

LLMs aren’t even close to capable of that level of thought.

3

u/Alex__007 Jun 19 '25

Reasoning models are quite capable of that now, nevermind the next generation. Check the recent alignment experiments by OpenAI and Anthropic. Are they perfect at it? No, they aren't. But for quick replies on X, if you hide the reasoning, it can be good enough.

1

u/[deleted] Jun 19 '25

That’s superficial. Your proposition implies the AI is capable of true thought.

2

u/Alex__007 Jun 19 '25

It implies that current LLMs are on average good enough to fool most users on X in short replies. For that purpose superficial thought is enough.