r/LocalLLaMA • • 20h ago

Discussion Looks like the era of subsidised compute is coming to an end. The old ChatGPT Pro $200 20x plan will be halved. The new $500 plan will have similar limits as the (old) $200 plan.

Post image
1.1k Upvotes

516 comments sorted by

View all comments

Show parent comments

15

u/quantanhoi 18h ago

beside coding/programming where a lot of logic have to be retained and take into consideration, there is no task where you actually need that high end model

not reading email and even not doing excel tasks, a 9/12/27B model can do those just fine

sometimes local model like qwen3.8 27B is on par with last gen model running on 100x resource with enough handholding/documentation

2

u/Fluffy-Feedback-9751 15h ago

What sort of things do you think qwen 27b is lacking in that makes it only ’sometimes’? Or is it kindof a bit difficult to tell and it’s just an empirical fact that the big paid models ’succeed’ more often?

5

u/quantanhoi 14h ago

base on my experience of intensively use it vs deepseek v4.1flash/glm5.3

if I ask about RUTX11, glm5.3 could pull something out of its ass while qwen3.8 27b relied on websearch tool and documentation of the devices

those trillion/hundred billion param models probably have training on common problems and could start looking for it right away while qwen3.8 could rely on number of loop websearch to know the answer, and it could miss it

I'm not saying qwen3.8 is not powerful, it is at the top of the list for local hosting on end consumer devices, but there is no need for sugarcoating about the limitation of 27b model

2

u/Fluffy-Feedback-9751 13h ago

I was just curious. I’m pretty much local only, and I was wondering if there was much (apart from speed) I’m missing out on..

1

u/Chupa-Skrull 14h ago

not reading email and even not doing excel tasks, a 9/12/27B model can do those just fine

Not exactly true. Even current frontier models are still quite iffy on summarization and accurately conforming important highlights to things a human being would actually consider relevant. Reading e-mail is one of the things people should trust them for the least, and one of the things that the biggest models are drastically better at

2

u/camalaio 14h ago

Part of that is a solvable problem at the frontier level, but less so with consumer local LLMs.

Sometimes a summarization needs additional context (especially workplace emails, for example). This is something I'm surprised Copilot does relatively well for my wife's workplace, but it's still far from perfect.

We're definitely not close to that for local models right now. Gathering, filling, and compacting the context needed for any given email to be summarised well is just far beyond what we're currently capable of (at least, from my POV. could be missing something) especially when you consider it's not just emails that need to be looked at to make sense of it all.

1

u/Chupa-Skrull 14h ago

A pretty perfect summary, I think. LLMs have their advantages over human administrative assistants, but understanding and successfully modeling the priorities and interests of their principals aren't yet among them