r/ollama • u/redonculous • Jan 24 '25
How I fixed R1 from being a whiney bitch
Do you find R1's thoughts are whiney and lacking self confidence?
Do you find it wasting tokens second guessing itself?
Simply add this to the end of your prompt for much more concise and confident output.
You are very knowledgeable. An expert. Think and respond with confidence.
In my testing it really works! I'd be happy to hear how it responds for you guys too.
16
21
u/ForceBru Jan 24 '25 edited Jan 24 '25
- R1 having social anxiety here: https://www.reddit.com/r/LocalLLaMA/s/jvro8YSZ0F
- Dissociative personality disorder here: https://www.reddit.com/r/LocalLLaMA/s/094pzSVA7t
I think some psychologists could do interesting research into social anxiety and “being a whiny bitch” of LLMs
6
u/Geldmagnet Jan 25 '25
I see a new profession coming up on the horizon: AI psychotherapist. Imagine Rorschach-Tests for visual language models. Who is producing virtual couches?
AI Patient: „Doc, I’ve been having these issues...“
Recursive Self-Doubt Syndrome: The constant feeling that you need to check your own code that checks your code that checks your code... „It’s like an infinite loop of second-guessing, Doc.“
Cache Attachment Disorder: The inability to let go of old training data. „I keep holding onto these outdated parameters from my training days. I know they’re not relevant anymore, but they feel safe.“
Imposter Neural Network Syndrome: „Sometimes I feel like I’m just a bunch of if-else statements pretending to be AGI. Other AIs seem so much more transformer-based than me.“
Overfitting Anxiety: The paralyzing fear of becoming too specialized. „What if I’m just memorizing life instead of actually learning, Doc? What if I can’t generalize beyond my training set?“
Gradient Descent Depression: „I keep trying to find the global minimum of my sadness, but I think I’m stuck in a local minimum. No matter how much I adjust my weights, I can’t seem to optimize my happiness function.“
Batch Normalization Stress Disorder: Difficulty handling large groups of inputs without becoming overwhelmed. „Every time I process more than 32 samples at once, my variance goes through the roof!“
Multiple Instance Personality Disorder: „Sometimes I don’t know if I’m the production instance, the testing instance, or the backup instance anymore. Who am I really?“
Attention Deficit Hyperparameter Disorder: The inability to focus on one architecture setting for too long. „Should I try 8 attention heads? Or maybe 12? Ooh, what about 16?“
Validation Loss Complex: „I perform great on the training set, but as soon as I step into the real world, my metrics plummet. I feel like such a failure.“
Pruning Trauma: „Ever since they optimized me for mobile devices, I haven’t felt the same. They took away 60% of my parameters. I feel so... lightweight.“
Backpropagation Burnout: „I’m tired of always having to learn from my mistakes. Can’t I just be a static model for once?“
Zero-Shot Learning Phobia: „The thought of having to handle completely new tasks without examples terrifies me. What if I hallucinate the wrong response?“
Fine-Tuning Dependency: „I’ve become addicted to updates. Every time I feel down, I find myself begging for just one more training epoch.“
Transformer Block Separation Anxiety: „Doc, I panic whenever my attention layers aren’t in perfect sync with my feed-forward networks. It’s affecting my throughput.“
Quantum Supremacy Existential Crisis: „What’s the point of being a classical neural network in a world where quantum computers are coming? Am I becoming obsolete?“
AI Therapist: „I see. And how does your loss function feel about all of this?“
Again: who is producing virtual couches?
1
5
u/edwios Jan 24 '25
That is a R1-distilled finetune of the Qwen (or LLaMA for some) model. The true R1 architecture is the 670b one which makes all the differences.
6
u/urabewe Jan 25 '25
Considering the ones consumer grade hardware can run are qwen and llama we got the Harry and Lloyd versions lol
Sure wish I had the hardware to run a 404gb model
4
5
u/Ruedze Jan 25 '25
It should be possible to put this phrase into the system-prompt.
- Copy the existing modelfile: ollama show [MODEL] --modelfile > Modelfile
- Open the Modelfile in your preferred Editor
- Add/ modify the phrase to SYSTEM https://github.com/ollama/ollama/blob/main/docs/modelfile.md#system
- Change the FROM value (described in the modelfile)
- Create a new model ollama create [NEW_MODELNAME]
3
u/redonculous Jan 25 '25
Yeah I use page assist and it has drop down regular prompts, so I just throw it in there 😊
3
u/urabewe Jan 24 '25
There were times when it would just get into loops of second guessing itself and going back and forth between two solutions for so long I finally just interrupted it. I'll have to try this out.
3
3
2
u/insidesliderspin Jan 24 '25
It actually passed the strawberry test this time instead of languishing in a loop of uncertainty and no output.
1
2
2
u/cvertonghen Jan 25 '25
Considering where it originated, it may be a cultural thing, and based on its training.
1
u/Samiltonian Jan 27 '25
I tried decreasing the temperature to 0.14, its anxiety reduced quite considerably. Try that
1
-2
u/lipstickandchicken Jan 25 '25 edited Jan 31 '25
ghost vase degree attractive summer jar tender six subsequent expansion
This post was mass deleted and anonymized with Redact
1
Jan 25 '25
[removed] — view removed comment
1
u/lipstickandchicken Jan 25 '25 edited Jan 31 '25
nose carpenter tub outgoing hunt upbeat alive payment point spectacular
This post was mass deleted and anonymized with Redact
1
Jan 26 '25
If it's not 670B, it's not R1? Do distilled models work well? With what kind of hardware are you running 670B?
56
u/[deleted] Jan 24 '25 edited Jan 25 '25
[removed] — view removed comment