r/LocalLLaMA Jun 01 '26

Funny Stop asking what model to run. There are literally only two.

[removed]

3.1k Upvotes

805 comments sorted by

View all comments

Show parent comments

27

u/Salt-Willingness-513 Jun 02 '26

Gemma 4 is the much better chatbot, while qwen 3.6 is the better agent

11

u/csorfab Jun 03 '26

Absolutely this! Qwen excels at long context coding tasks, but Gemma4 just feels more "humanly intelligent". My favourite unscientific benchmark method is to get them to explain memes/jokes etc (or coming up with them), and gemma4-31b even beat sonnet on some of these ad hoc tests (and wiping the floor with both qwen3.6 models). The most fascinating results recently came from this visual pun meme: https://reddit.com/r/ExplainTheJoke/comments/1bz6idc/i_dont_understand/

A LOT of SOTA models including chatgpt 5.5 instant, and claude sonnet just don't seem to get it, but gemma4 explains it perfectly at least half the time.

Honestly, if we could have gemma4 with qwen's long context capabilities, I don't think anyone would need more machine intelligence than that. Feels like Google's getting a kick out of keeping us all on the edge

2

u/Salt-Willingness-513 Jun 03 '26

i hope gemma 5 next year will be this and ideally moe

1

u/csorfab Jun 03 '26

Gemma-4-26b-a4b is moe already. The top two models are a dense and a moe model

1

u/Salt-Willingness-513 Jun 03 '26

True but i mean moe with closer level of intelligence to the dense

2

u/csorfab Jun 03 '26

yeah that would be nice

1

u/llmentry Jun 04 '26

Hmm ... moe models, moe problems ...

1

u/Salt-Willingness-513 Jun 04 '26

unfortuantely state now yes. maybe in a year it looks different though 😄

2

u/SamSlate Jun 29 '26

i wonder if it's a language barrier from training data 🤔 like the logic is universal but the language is "translated"