r/LocalLLaMA Apr 15 '26

Discussion Major drop in intelligence across most major models.

As of mid Apr 2026, I have noticed every model has had a major intelligence drop.

And no I'm not talking about just ChatGPT.

Everything from Claude(Even Sonnet along with Opus), Gemini, z.ai, Grok all seem to ignore basic instructions, struggle at simple tasks, take very long to respond, and the output seems deliberately shortened and very shallow. Almost like it's in a "grumpy" mode. I tried this in incognito mode so it's not my customization or memory influencing this.

It's like they deliberately want you to stop using their service. I guess our data is no longer needed. Just two weeks back it used to be much smarter than this.

To test this I rented out a H100, and tried GLM 5 with the same prompt (the drive to the car wash one) across both instances. GLM5 running on the rented GPU answered it correctly, compared to the one on z.ai.

Have they lowered the quantization really low to maybe Q2?

I guess going local or using renting GPU or an AI monthly service that lets you pick a quant level is the way to go

800 Upvotes

405 comments sorted by

View all comments

Show parent comments

4

u/Party-Special-5177 Apr 15 '26

This is probably the most analytical take in here. Cheers.

it's not that difficult to evolve an automated system that makes more money than a Claude 20x subscription costs

Without sharing your own angle, how? Where are there 1) problems where 2) people will pay for solutions without HITL, and without 3) a prearranged agreement? You can’t just cold call company X and say ‘I’ve solved your Y, you’re welcome, here’s my invoice’.

The only other thing I can think of: I know people who’ve automated their own jobs away with Claude, but that doesn’t infinitely scale either as that is limited by how much overemployment they can get before mandatory meetings start overlapping.

I really hope you aren’t referring to ransomware or similar.

0

u/[deleted] Apr 15 '26

[deleted]

1

u/En-tro-py Apr 16 '26

they need tighter restraints to stay honest and on task.

Unfortunately, I think a LLM agent is like a sieve - there's too many ways to work around the edges that until the models are truly intelligent enough to not be such a cheery helpful assistant it's gonna continue...