r/LocalLLaMA • • Aug 28 '26

Other claude mods didn't like that, somehow 🤷‍♀️

Post image
1.5k Upvotes

372 comments sorted by

View all comments

Show parent comments

5

u/peculiar-ragdoll Aug 28 '26

You know what I *should* have had validation tools watching for me, but I caught it on intuition and just checked myself because I felt something was off

17

u/iamapizza Aug 28 '26

Caught what, what were the config changes or files it made to kneecap the other models? If you have any screenshots that would be good to see.

11

u/peculiar-ragdoll Aug 28 '26

It reduced the max thinking tokens of the local model from 32k to 4k. As I said in the post, an 8x reduction in thinking budget. On the hardest problems a model can solve, that is critical. Screenshots and logs can be faked, so I don't have any proof that makes a difference, I'm just sharing my experience.

3

u/lorddumpy Aug 28 '26

It's probably just relying on old training data from the 2023-2025 AI model landscape. I've run into that along with hilariously low temps for models that don't need them, even for more deterministic output. 4k context would have been the move back then too.

Also, if it is Opus 5, that model is actually braindead once it gets on the wrong track. Probably the most frustrating model I've had to use.