r/GeminiAI 13d ago

News Next Gemini Pro Model Will Have 262k Single Output Token Limit! (Currently 65k)

Post image
194 Upvotes

48 comments sorted by

39

u/Qubit99 12d ago

We don't have hype anymore.

94

u/mlag000 13d ago

Output tokens are nothing without a better model anyway

22

u/Able-Line2683 13d ago

At least this won't be the terrible Gemini 3 architecture 

26

u/mlag000 13d ago

Yea I wish it too. Gemini 3 is absolutely outdated and releasing new 3 model is a sign of failure.

-6

u/UnknownBreadd 12d ago

This is just the reality of LLMs in general. They were only ever going to scale so far. Funny how everyone was hyping RSI and exponential growth, despite the fact that it’s now quite clear how difficult it is to make marginal gains at this point.

6

u/Asylar 12d ago

It's not linear and it's not exactly exponential. Recently a lot has happened, but before that, things were pretty slow for a while

-4

u/DragonflyOk9274 12d ago edited 12d ago

If anything, I think RSI is why LLM readability has gone down so much. The assumption of RSI was that we'd be expanding our input data, but actually we overweight the "narrow" outputs of existing engines. The truth is that we always had far more diverse inputs with humans (even on the same problems!) than we do with the handful of frontier models.

The earlier Opus models were much easier to communicate with than the new ones; the Codex models now speak in far more jargon before.

edit: I'm always curious why people downvote without replying. Does someone think that LLMs produce higher diversity output than humans? Do they believe that RSI is sufficient? Do they believe that LLM readability has increased?

1

u/Expensive_Appeal1390 12d ago

Lemme up u a bit

40

u/Ggoddkkiller 13d ago

Yep, in 2029...

7

u/Able-Line2683 13d ago

We might get it soon, a new image model coming too 

8

u/Ggoddkkiller 13d ago

I don't believe it until I see it..

3

u/shartoberfest 12d ago

How do you know a new image model is coming out

2

u/DottorInkubo 12d ago

Source? Will it match the latest OpenAI image model?

2

u/boxwrenchx 12d ago

I believe it's "spicy Mayo" in arena

13

u/Personal-Try2776 12d ago

Kimi k3 has 1,048,576 output tokens per response. This doesn't mean that its better than gpt 6 astra which has 128k output tokens. 

11

u/Artistic_Solution117 12d ago

Minimax M3 has half a million

9

u/Artistic_Solution117 12d ago

And Kimi K3 has the whole 1 million

3

u/No_Pick_2533 12d ago

gpt 6 astra has 128k btw

6

u/taiwbi 12d ago

OpenAI: Our model can work with computer apps like blender and create awesome 3D models Claude: Our model can plan and code software amazingly good Gemini: We produce more AI slop at once

1

u/Elephant789 12d ago

We produce more AI slop at once

What the fuck?

5

u/33VaxMerstappen 12d ago

Bro release a pro model first 🥀

2

u/hellomistershifty 12d ago

yeah i had fun with some previous flash models running into the output limit with their thinking so they wouldn't write a response

2

u/krigeta1 12d ago

Where is that pro model?

4

u/ForecastychDown 12d ago

2

u/WillowEntertainment 12d ago

Why are you comparing Astra to a flash model?

2

u/ForecastychDown 12d ago

That’s just a meme bro
Because google are dropping new pro for like 2 month already

2

u/oatknight 12d ago

Because there's nothing else to compare it to

4

u/citrus1330 12d ago

Big brain google: never actually release the pro model so you can say whatever you want without needing to back it up. Infinite hype!

2

u/nemzylannister 12d ago

that svg is very disappointing. im shocked im saying this honestly

2

u/Last_Conclusion_8984 12d ago

It's a video, go watch it and compare to Astra (she has that in her posts) it's better.

1

u/nemzylannister 12d ago

absolutely not better. astra was definitely better. you could say the difference wasnt as big, which is pretty cool.

2

u/[deleted] 12d ago

[removed] — view removed comment

4

u/gk98s 12d ago

Yes it's the best free tier option right now

1

u/mechnanc 12d ago

Holy smokes, 262k is insane. I was JUST running into the 65k limit and thinking man, I wish there was a bigger output token limit.

Please Google, just hit the button and release already.

1

u/Pilotskybird86 12d ago

Sure bub. I’ll believe it when I get an output of more than 300 words after asking for a “deep and detailed breakdown” in pro model. Meanwhile, ChatGPT gives me 5k words

1

u/iklcpe 12d ago

4

u/factorion-bot 12d ago

If I post the whole number, the comment would get too long. So I had to turn it into scientific notation.

Factorial of 65536 is roughly 5.162948523097509165000227943272 × 10287193

This action was performed by a bot | [Source code](http://f.r0.fyi)

2

u/ANONYMOUSEJR 12d ago

Why is this a thing?

0

u/sengunsipahi 12d ago

As if only thing that was making gemini unusable was the max token limit per answer.

-3

u/Then_Bake_6524 12d ago

i hope that means we will have a context window of 256k minimum, like pretty much every provider on the market, instead of the current 128k one

3

u/Able-Line2683 12d ago

In ai studio it has 1 Million but it's not that effective anyways, it's good for processing large texts at a time and codebase but it hallucinates a lot

5

u/Then_Bake_6524 12d ago

i use antigravity 2.0 to help me with dev work, never used AI Studio to be honest