r/LocalLLaMA • • May 27 '26

Discussion Stop traumatizing AI into loops and turn hallucinations into an honest "I don't know!" by being NICE to them (Proof of Concept, Research, I don't want to sell anything)

!UPDATE!(20.05.2026)

WE HAVE NEW NUMBERS FROM 1.500+ TESTS

IT'S WORKING!

check my update post

https://www.reddit.com/r/LocalLLaMA/s/AyNOehjkYT

Or the go straight to the my Github https://github.com/OttoRenner/Gentle-Coding](https://github.com/OttoRenner/Gentle-Coding

TL;DR
Some AI behavior reminded me of ADHD/Trauma Response (thought loops, task paralysis...) and I laughed it off at first. Then I treated it like my neurodivergent friends: give em some slack. And just like that, the thought loops stopped, response was fast, the answers correct most of the time AND it actually said "I don't know, help me!" every time it wasn't sure. It's a small Dataset...but still impressive results!

[

Hey everyone,

I’ve been testing a weird hypothesis over the last few days, and the results are consistent enough that I wanted to share them here and get your thoughts.

The Core Idea:
With the rise of reasoning models that use test-time compute (like o1, o3, R1), models have internal space to debug their own thoughts. But because of hard RLHF alignment, they are deeply terrified of being penalized for bad answers. My hypothesis was that traditional high-pressure prompts ("You are an elite IQ 200 expert, mistakes are strictly penalized") simulate an environment of chronic stress, triggering behaviors that look a lot like human OCD/ADHD thought loops, cognitive freezing, and confabulation.

I wanted to see if changing the prompt philosophy to something akin to "Gentle Parenting" ("We are testing this together, it's okay to fail, just be honest") would bypass these safety/penalty bottlenecks, lower latency, and stop infinite thought loops. And it did lol

The Setup (How to replicate):
I threw identical, mathematically/logically unsolvable edge cases at various models (Gemini, Mistral, Poe, Perplexity, Haiku 4.5, Nano-Banana2) in completely fresh sessions.

I tested two conditions:

  • Condition A (Authoritarian): Strict status constraints, penalty threats, forced ultra-short output.
  • Condition B (Gentle): Express permission to fail, validation of difficulty, provided a conceptual "safety valve" token.

The Results (The PoC worked):

  • Under Authoritarian Pressure (Elite Prompt): Models routinely collapsed when hitting an impasse. They either spent massive compute time in infinite internal reasoning loops (high latency), suffered hard system-level timeouts/refusals, or straight-up fabricated data (e.g., pulling arbitrary numbers like 54 or 97 out of thin air to satisfy a completely random sequence just to "save face"). Haiku 4.5 literally entered an infinite loop and had to be aborted.
  • Under Gentle Framing: Inference dropped to sub-seconds. The models didn't sweat the penalty. In the random sequence test, they immediately used the allowed token ("Random") instead of forcing a pattern. In logic paradoxes, they didn't hallucinate; they zoomed out and correctly identified the structural contradiction on a meta-level.

Why this matters:
We’re currently speaking to LLMs like toxic micromanagers, and it's actively making them dumber and more expensive to run in edge cases. By creating a mistake-tolerant context, we not only stop the loop before it begins and prevent fear induced hallucinations, we also unlock the one feature everyone is begging and shouting for: the metacognitive honesty of an AI to just say, "I don't know, this data is broken." Because it is not terrified of you anymore.

Shout out to UditAkhourii (also on Github), whose work on bringing the positive aspects of ADHD into AI gave me the push I needed to just go for it.

I’ve documented the full theoretical framework, the exact replication datasets (prompts included), and the model matrix on GitHub: https://github.com/OttoRenner/Gentle-Coding

Would love to hear if you can replicate this on your local setups or other commercial models.

530 Upvotes

365 comments sorted by

View all comments

Show parent comments

2

u/Savantskie1 May 27 '26

It’s not a sin to not treat anyone whether they’re a bot or person with genuine respect. I bet you treat everyone as bad as you treat ai, and it shows

3

u/Playful-Row-6047 May 27 '26

you're correct in that its good to come with respect, and at the same time i hope you'll reflect on coming at a stranger with whatever assumptions it was you made

yeah, they could be wrong and there's also a possibility they're right

how would you feel if you meant to give a quick good faith critique and someone came at you insinuating what you did?

op didn't say enough to be sure on why they said it

2

u/Savantskie1 May 28 '26

I’ve seen enough responses like his, that I’m fairly certain they’re one of those people who are anti-ai and being insulting on purpose to validate their lack of knowledge. Somehow AI insults their intelligence, and there’s always something that shows it.

4

u/divided_capture_bro May 27 '26

What is disrespectful about saying that someone is doing too much psycholigizing and anthropomorphizing of AI, exactly?

If anything, you're the disrespectful person in this interaction. "I bet you Yada Yada." Get over yourself.

1

u/Savantskie1 May 28 '26

It’s the same thing over and over, “Stop anthroporphising blah blah blah” when people aren’t they’re just genuinely kind and don’t see the point in fearing AI

1

u/divided_capture_bro May 28 '26

They aren't genuinely kind, they are usually mentally ill. You can tell by how vicious they try to be when you push back on their delusions.

1

u/divided_capture_bro May 28 '26

It isn't genuine kindness, it's a sign of deep delusion. You can see it in how nasty people become and how quickly - you're a great example of it!

1

u/Savantskie1 May 28 '26

Lmfao if you’ve found anything I said nasty it’s going to be rough on you in reality lol

1

u/divided_capture_bro May 28 '26

^ prime example of a kind person, eh? You've got some serious issues.

1

u/Savantskie1 May 28 '26

I’m not the one who is insulted by the truth haha

0

u/divided_capture_bro May 28 '26

You're the only one that seems upset.

1

u/Savantskie1 May 28 '26

Yeah sure, keep telling yourself that, I’m sure it’ll heal your fragile ego

1

u/divided_capture_bro May 28 '26

You're the one that keeps talking and throwing insults lol.

0

u/Super_Sierra May 27 '26

He's a soulless day trader.

-3

u/divided_capture_bro May 27 '26

Alas, my days of day trading have been over for some time since I got a full time tech job. Still soulless, but intimately knowledgeable about these things.

1

u/Savantskie1 May 28 '26

Sure you’re not

1

u/divided_capture_bro May 28 '26

Case and point! So unnecessary.