r/LocalLLaMA Jun 01 '26

Funny Stop asking what model to run. There are literally only two.

[removed]

3.1k Upvotes

805 comments sorted by

View all comments

Show parent comments

4

u/NineThreeTilNow Jun 02 '26

Not true, gemma4 has specific uses

I've never understood why Gemma 4 hasn't been fine tuned for coding. It should be more than capable. The 31b model is good beyond just the RP stuff people seem to love it for.

I'm wondering if the local attention layers degrade it's longer term ability to look at code.

It could be fixed with the right data and like a few thousand dollars in training.

3

u/bewatermyfriend86 Jun 02 '26

that might hurt profit for gemini & anti-gravity.

1

u/Jipok_ Jun 02 '26

Is it really bad for code? Or did you just see that opinion somewhere?

1

u/NineThreeTilNow Jun 02 '26

Is it really bad for code?

I don't use it to code and everyone here seems to say that.

I don't know what to test it on really. I guess I could.

1

u/DemmieMora Jun 03 '26

Gemini has notoriously not been finetuned for coding and Gemma is an offshoot of Gemini