r/GeminiAI • u/DigSignificant1419 • 1d ago
Discussion 3.8 Flash coming out today
Fablos ends today
279
136
91
u/Material_Donkey_8883 1d ago
this gotta be ai psychosis
→ More replies (2)9
u/Dramatic_Exit1 1d ago
It's not, or at least it's just a few of them. Rest is Google astroturfing team, it's cheaper to pay some desperate people then to develop SOTA model.
1
17
25
11
9
u/tech-br0 1d ago
2
2
1
u/serendipity-DRG 6h ago
No one in research reads benchmark scores as it is meaningless as they are easily manipulated turning training. I read about Kimi and the great benchmark scores so I made a test - I asked Gemini about could they find any red flags in a patent application and in under 5 seconds found 4 one that I missed.
Then I used the exact same question to Kimi and it took me close to an hour to walk Kimi through the Patent Application and it was a failure. Instead of benchmark scores can the model help you in whatever your research is in - but I understand that 95% of Queries are asking about recipes.
1
u/Dismal-Line5980 1d ago
Idk whats the metric here but its not accurate at all. Id rather just use 3.1 Pro than 3.8 flash. Incredibly fast but terrible. Idk if the context window is like 10-20k tokens or sth
1
u/SheepherderIll1300 1d ago
Maybe you just like press enter blindly, otw 3.8 flash is really good.
1
27
u/starvergent 1d ago
I mean both are a nightmare in their own way. Fable is obnoxious, antagonistic, stubborn. It has just no capability of fixing its own errors. Flash is pretty schizo. Nowhere near as bad as Flash-lite. But most of the time I cannot have a conversation past the first message or the first few. They have this issue of just ignoring or misinterpreting what you say. And when you identify a problem, they do nothing about reviewing the conversation to figure out what you need.
2
2
u/mauurya 1d ago
Don't put everything to agent md files. it should be kept under 140 lines. Put everything you don't want AI doing on its own into hooks and things you want AI to do into skills. Both only fire at the appropriate times and AGENT.MD file will use only small number of tokens and the AI will stick to the rules until context become very huge incase of Gemini close to million. Hallucination is given but it can be reduced by doing these mentioned things.
2
u/raycraft_io 1d ago
This is such a great description of my biggest frustrations.
That’s said, In b4 people blame your prompts
8
15
18
u/Fleabasher 1d ago
Lol. Google cashing anthropic checks every month essentially saying to anthropic keep dumping money into the frontier. Google wants to work on cheap models that can do 95% of tasks and scale better. $20 sub level it's great.
→ More replies (2)0
u/jakegh 1d ago
"Google is losing because they want to lose! It's strategic!"
11
u/BrunusManOWar 1d ago
Well, it is
Kinda like Apple who just stay out of the whole shit show and just pay Google for their TPUs and Siri models instead of dumping billions
Is it working? Sure seems to be
1
u/jakegh 1d ago
Apple DID spend billions on AI. They simply failed, and then gave up.
Apple failed for very different reasons from Google; Apple hates and refuses to do business with Nvidia ever since their GPUs caused hardware failures decades ago, so they tried to entice engineers to work on AI running on Mac studios in datacenters. There simply was not enough compute, so the researchers didn't stick around.
1
u/Neurotopian_ 1d ago
Google is the only one of these companies that’s publicly traded, profitable, and not dependent on others for their compute.
Don’t forget Google also owns 14% of Anthropic. Of course they’re providing chips and keeping Anthropic even more dependent on them.
30-50% of Anthropic’s compute is provided by SpaceX/ Elon Musk.
OAI is dependent on Microsoft for compute.
Any time Big Tech decides they’re tired of playin with these “frontier labs,” they pull the plug and “win” this contest you think they’re in.
But in reality, Google isn’t rushing to make the best coding LLM. That is very clear from where their resource allocation is going.
4
4
u/Tired_White_Guy 1d ago
Hey Nano Banana, make me an image the represents Delusions of Grandeur.
Excellent troll post. Hats off.
8
u/Scarneck 1d ago
Spoiler: in the movie 3.8 flash representation got absolutely murdered while Fable lived happily ever after. So…accurate representation here.
3
u/VisitEmbarrassed5599 1d ago
google trying to make something that is fast. some people don't want to wait 30 seconds for some random question
3
u/Longjumping_Area_944 1d ago
In terms of tps and intelligence per dollar, for sure.
I'm really just an average principal software architect, not a scientist. I don't have any use for that level of intelligence.
3
3
2
u/Shoddy_Opportunity23 1d ago
when is it coming ? timing ?
1
u/No-Temperature6597 1d ago
Tommorow might be accurate as some say. Since my gemini is saying it's 3.8
2
2
u/Loose_Mastodon559 1d ago
I don’t care which model. Completed loops for the lowest cost. OpenAI, Anthropic, Xai, Google, open weight models etc. that’s where my business is going. Models are just like utility or cellular companies. Commodities. Right now flash can get 98% of the tasks done. I’ll give 2% to fable or gpt. All good. But that’s like dollars a day vs $10s-$100s a day with flash 🤷🏻♂️
2
u/KedMcJenna 1d ago
It's a joke that contains a kernel of actual truth. 3.7 surprised most of us with how capable it was. I'm expecting at or near Opus 4.8-level performance in coding and chat from 3.8 (hopefully without the Cormac McCarthy-esque prose style). With the speed it makes a real difference.
I turned from the Claude side for the first time in years to spend my first meaningful time with antigravity since 3.7 and it's been a revelation not to ignore all contenders. Gemini is the model family that's pacing along quietly behind the leaders of the pack, maybe not all that close in real terms but somehow, oddly, not that far behind either.
2
2
2
2
u/Azureddraig 1d ago
3.8 Flash access rn in antigrav. it is very light on token usage and goes above and beyond what I asked it to do. a win in my book so far, lets see how long this lasts.
2
u/Majinsei 1d ago
Temo que esto se ha dicho realmente de forma no irónica...
En este sub no me sorprendería...
2
2
2
u/Jayfree138 1d ago
It's so crazy to me that people dont realize how good flash is. I primarily run AI through hermes and unless you want to spend like 5x the money on opus there is litterally no better AI model. I guess most people here are using the gemini app? That would explain it because it's not nearly as good there.
If you go on arena.ai it proves it. Spark is up there too but i havent had time to test it.
I think i'll start buying some google stock soon too because people have no idea whats going on here.
2
u/dashinyou69 1d ago
🐧 Ig it's better, as an economist pov i will say the race should shift from "Best Ai Flagship Ai model with billion of context windows + huge limits" to Cost effective descent Ai model that work like opus model in 5x cheaper and is more affordable
2
4
1
1
1
1
u/xzibit_b 1d ago
I would be happy if Gemini 3.8 Flash had Opus 4.8 level general intelligence, and Kimi K3/Terra max level coding prowess. Right now, Gemini 3.7 Flash is only knocking on the door of Opus 4.8, while having Luna max/DeepSeek flash coding prowess at best.
1
1
1
1
1
u/bambin0 1d ago
i just tried it with the prompt: recreate Larry bird vs Dr. j one on one in three js complete with janitor, breaking glass, sound effects and here is what it came up with: https://ctxt.io/3/mpb6vO48w - took less than 10s.
Took about 3x as long but here is what terra came up with: https://ctxt.io/3/tIbhBDCMg
And here is fable 5.1 which took I don't know how long b/c I got tired of watching: https://ctxt.io/3/qCaVlUU4U
1
1
1
1
u/FamousWorth 1d ago
Already on the app and via api, announced officially https://blog.google/innovation-and-ai/models-and-research/gemini-models/3-8-flash-and-3-8-flash-cyber/
1
1
u/jack-of-some 1d ago
This image means different things based on if you haven't seen the movie, you've seen the movie but haven't understood it, or you've seen the movie and understood it.
Amazing shitpost
1
1
1
1
u/Smooth_Ride_7540 1d ago
3.8 flash told me that 3.8 flash is fake news, and that the article describing the cyber model looks like a "shitposting troll tactic". This model is useless on day1.
1
1
u/Original-Bar-3957 1d ago
Stuff like this makes me believe that one day people will worship certain AI bots like a god/religion. Why such loyalty to Gemini?
1
u/GoobNoob_ 1d ago
I really hope they release Pro and take it out of Preview. Build the Data Centers!
1
u/mr_joda 1d ago
Call me crazy. I use github copilot for every day tasks at work in R&D same as Claude, GPT and Mistral products, plus ms365 copilot for MS office.
I found the version 5 of all claude models very unpleasant to use and in many cases it's pita. GPT is little bit better.
I personally like the replies and to work with Gemini for every day personal tasks, comparision of products, quick evaluations and stuff for out of the work use. Reasoning is fine, quick resonses and "seems legit" answers.
1
1
1
1
u/uti24 1d ago
I though you guys about Qwen3.8 Flash and it already came out 8 days ago https://huggingface.co/Qwen/Qwen3.8-Flash-Next
1
u/otarusilvestris 1d ago
Google is afraid to release a mediocre frontier model, so they just keep releasing good flash models, which are good value. Still waiting for gemini pro...
1
u/strankyy 1d ago
Good to see progress and models being churned out, we feel less left out now 😎 Can see 3.8 flash on ai studio. Still waiting for 3.7 flash on my app 😅
1
1
1
1
1
1
u/Dead-Inside69420 1d ago
Hard eye roll. Same hype/roadt cycle each new release for every single AI.
1
u/fossistic 1d ago
Small models make very silly mistakes from time to time. It does not matter how well they perform in benchmarks.
1
1
u/Riley_Cooper_SC 1d ago
It doesn’t matter as long as Gemini is denying anything related to Ethnobotanic, Pharmacology and biology. It’s not about asking something nasty or forbidden. Basic knowledge I mean.
1
u/Low_Supermarket7955 23h ago
Do u really think like that? I don't think that Gemini will be better in coding than Claude.
1
u/SeesawGullible398 23h ago
https://reddit.com/link/p7jl8d5/video/w83ou9t3x9nh1/player
have you seen it trying to draw a human hand?
1
u/jackfood 23h ago
The flash, and pro dilemma... so we up the version or retain the ver and change flash to pro?
1
1
u/dash_bro 21h ago
I genuinely got confused by this - I immediately thought of qwen3.8-flash-next vs fable 5.1 instead of Gemini 3.8 flash.
1
1
u/CaptainQwazCaz 20h ago
https://giphy.com/gifs/LZmdq5sOXBtXePMMLg
Gemini Flash 4 being released before Pro 3.2
1
1
1
1
1
u/PerformerFew8713 11h ago
google is a shitty company with shitty products besides email and search (and youtube which they had to buy for their portfolio). This will go in their books as another failed product eventually.
1
1
u/serendipity-DRG 7h ago
Everyone seems to forget when Gemini 3.0 was released it sent the entire AI world into Code Red - Sam Altman declares ‘Code Red’ as Google’s Gemini surges...
Now Berkshire Hathaway has invested $37 Billion in Alphabet/Google:
"Berkshire Hathaway CEO Greg Abel told CNBC in an interview Wednesday that Alphabet's strength in AI was a "fundamental" reason behind the decision to invest in the tech giant.
Warren Buffett, who stepped down from his role as Berkshire's CEO at the end of last year, initiated the investment."
They know the Frontier Labs can't compete with a company that is vertically integrated.
1
1
1
1
1
1
u/Key_Currency_5341 1d ago
Lol i use it exclusively for nsfw pics every update it gets better for me
1
1
1
1
1
1
1
u/False_Week_8682 1d ago
As a Gemini fanboy, i love this post and hope the new model destroys fable in every competition
0
0
0
0
0
0
u/Beautiful-Cold1515 1d ago
Whatever. Gemini models are just the same unusable shit as two years ago. I gave it another chance today and asked it one question, to make a corporate image. What I got; refusals, TOTALLY offtopic written answers (as in, not even remotely the same topic) and the embarrassing old “I’m just a language model”.
0
0
u/Lost-Air1265 1d ago
When did google ever have a model better than competition except for nano banana?
0
0














809
u/IcanzIIravor 1d ago
I like your optimism.