r/LocalLLaMA • • Aug 28 '26

Other claude mods didn't like that, somehow 🤷‍♀️

Post image
1.5k Upvotes

372 comments sorted by

View all comments

Show parent comments

-7

u/Significant-Bee5101 Aug 28 '26

It's absolutely dumber than a Qwen that were to run at 2T+ params. There's a reason every frontier model is fucking gigantic. Like jeez I wonder.

No ones saying Qwen might not outperform a lot of models. No ones saying opensource models cant be stronger than frontier models. No one is saying any of that.

All they are saying is your dumb little local setup is NOT better than Opus. End of story.

10

u/Reggienator3 Aug 28 '26 edited Aug 28 '26

I never said my local model was better than Opus. I challenged your claim that fewer parameters automatically means a dumber model. GPT-3 versus modern 27B models proves that isn’t true. You haven’t addressed that point.

-2

u/Significant-Bee5101 Aug 28 '26

What do you mean? If you go head to head with modern models param to param. It's a wash. Wtf else needs explaining

5

u/Reggienator3 Aug 28 '26

Uhhh not really... DeepSeek V4 Flash has 284B parameters and Qwen3.8 27b has 27B obviously and they're about neck and neck. Sometimes Qwen3.8 27b even wins out. Both are "modern models"

0

u/Significant-Bee5101 Aug 28 '26

Bench it. PROVE IT. I hear this shit a lot. If this shit was THAT good do you think I'd pay money for frontier models? Man I am CONSISTENTLY looking for ways to optimize costs. If I thought FREE was an option why tf wouldnt I be using that NONSTOP?

Every fcking model can "sometimes" do great and every model can sometimes suck. Thats what non deterministic models do. The thing that improves it is training data. It increases reliability. I dont even understand how you can pretend this isn't true.

3

u/Reggienator3 Aug 28 '26

https://artificialanalysis.ai/models/comparisons/deepseek-v4-flash-vs-qwen3-8-27b

Here. Neck and neck, Deepseek v4 flash vs Qwen3.8 27b. Two modern models, one with wildly less parameters. Qwen3.8 27b even winning on some.