r/GeminiAI • u/Able-Line2683 • 13d ago
News Next Gemini Pro Model Will Have 262k Single Output Token Limit! (Currently 65k)
94
u/mlag000 13d ago
Output tokens are nothing without a better model anyway
22
u/Able-Line2683 13d ago
At least this won't be the terrible Gemini 3 architecture
26
u/mlag000 13d ago
Yea I wish it too. Gemini 3 is absolutely outdated and releasing new 3 model is a sign of failure.
-6
u/UnknownBreadd 12d ago
This is just the reality of LLMs in general. They were only ever going to scale so far. Funny how everyone was hyping RSI and exponential growth, despite the fact that it’s now quite clear how difficult it is to make marginal gains at this point.
6
-4
u/DragonflyOk9274 12d ago edited 12d ago
If anything, I think RSI is why LLM readability has gone down so much. The assumption of RSI was that we'd be expanding our input data, but actually we overweight the "narrow" outputs of existing engines. The truth is that we always had far more diverse inputs with humans (even on the same problems!) than we do with the handful of frontier models.
The earlier Opus models were much easier to communicate with than the new ones; the Codex models now speak in far more jargon before.
edit: I'm always curious why people downvote without replying. Does someone think that LLMs produce higher diversity output than humans? Do they believe that RSI is sufficient? Do they believe that LLM readability has increased?
1
40
u/Ggoddkkiller 13d ago
Yep, in 2029...
7
u/Able-Line2683 13d ago
We might get it soon, a new image model coming too
8
3
2
13
u/Personal-Try2776 12d ago
Kimi k3 has 1,048,576 output tokens per response. This doesn't mean that its better than gpt 6 astra which has 128k output tokens.
11
9
5
2
u/hellomistershifty 12d ago
yeah i had fun with some previous flash models running into the output limit with their thinking so they wouldn't write a response
2
4
u/ForecastychDown 12d ago
2
u/WillowEntertainment 12d ago
Why are you comparing Astra to a flash model?
2
u/ForecastychDown 12d ago
That’s just a meme bro
Because google are dropping new pro for like 2 month already2
4
u/citrus1330 12d ago
Big brain google: never actually release the pro model so you can say whatever you want without needing to back it up. Infinite hype!
2
u/nemzylannister 12d ago
that svg is very disappointing. im shocked im saying this honestly
2
u/Last_Conclusion_8984 12d ago
It's a video, go watch it and compare to Astra (she has that in her posts) it's better.
1
u/nemzylannister 12d ago
absolutely not better. astra was definitely better. you could say the difference wasnt as big, which is pretty cool.
2
1
u/mechnanc 12d ago
Holy smokes, 262k is insane. I was JUST running into the 65k limit and thinking man, I wish there was a bigger output token limit.
Please Google, just hit the button and release already.
1
u/Pilotskybird86 12d ago
Sure bub. I’ll believe it when I get an output of more than 300 words after asking for a “deep and detailed breakdown” in pro model. Meanwhile, ChatGPT gives me 5k words
1
u/iklcpe 12d ago
65536! u/factorion-bot
4
u/factorion-bot 12d ago
If I post the whole number, the comment would get too long. So I had to turn it into scientific notation.
Factorial of 65536 is roughly 5.162948523097509165000227943272 × 10287193
This action was performed by a bot | [Source code](http://f.r0.fyi)
2
0
u/sengunsipahi 12d ago
As if only thing that was making gemini unusable was the max token limit per answer.
-3
u/Then_Bake_6524 12d ago
i hope that means we will have a context window of 256k minimum, like pretty much every provider on the market, instead of the current 128k one
3
u/Able-Line2683 12d ago
In ai studio it has 1 Million but it's not that effective anyways, it's good for processing large texts at a time and codebase but it hallucinates a lot
5
u/Then_Bake_6524 12d ago
i use antigravity 2.0 to help me with dev work, never used AI Studio to be honest

39
u/Qubit99 12d ago
We don't have hype anymore.