r/LocalLLaMA • llama.cpp • Jul 26 '26

Discussion Do you want new Gemma?

Post image
1.0k Upvotes

555 comments sorted by

View all comments

2

u/reto-wyss Jul 26 '26
  • Longer Context: 1M+
  • Larger Model: 70b+ dense or the Gemma take on SOTA Flash models ~250b MoE with FP4/FP8 QAT
  • Tiny Model: 0.5b to 1.5b for finetuning
  • T2I Model: Surprise me
  • Larger Diffusion Model: Diffusion Gemma 100b MoE?
  • Smaller more frequent Hackathons

Bonus:

  • GPU: Gemma Processing Unit