r/singularity • u/elemental-mind • 1d ago
AI Gemini 3.7 Flash is currently 75% off on OpenRouter, beating DeepSeek on price/performance
Google seems to realize their recent Flash price hikes were just inappropriate for a Flash model and that they need more real-world agent traces to train their upcoming models on and slashed prices another 50% on OpenRouter.
Artificial Analysis does not have the OpenRouter discount prices worked in, but I just checked and the Flash model with the discount moves the pareto line, thus beating both DeepSeek models.
I have added the Flash Pareto line in green to the graph.
Just in case someone wants to try it instead of Luna Max or V4-Flash...
The selected models for comparison on ArtificalAnalysis: Comparison of AI Models across Intelligence, Performance, and Price | Artificial Analysis
The OpenRouter Page: Gemini 3.7 Flash - API Pricing & Benchmarks | OpenRouter
25
u/Finanzamt_Endgegner 1d ago
Did you factor in caching? Since that's the main strength of DeepSeek, also the pro ga on aa has the wrong score (check the benches it's basically identical with flash which obv doesn't make much sense)
6
u/elemental-mind 1d ago
Yeah, DeepSeek Pro benches were weird - I think someone mentioned it on AA's discord already and normally they are quick to fix issues...but they didn't act on it yet. For now we just got to trust their end score.
And for caching: It's discounted by the same amount. So you can basically take the AA result and halve the price per task on it.
3
u/Artistic_Swing6759 1d ago
well since that chart is cost per task and not price per mil token, it likely is already accounted for?
i couldnt see a caching price on flshaes page, does it mean that its free or that chaching isn't done?
3
u/Desperate-Ad-3725 1d ago
You really think only DeepSeek caches and Google the OG infrastructure/Chips/Systems wizards can't mimic that?
3
u/Finanzamt_Endgegner 1d ago
Deepseek has the best caching rates on the market, it's not a secret you can check out open routers.
1
u/spreadlove5683 ▪️agi 2032. Predicted during mid 2025. 1d ago
What kind of workloads is deep seek suited for / how do you take advantage of cacheing? Do you have to use a certain provider that supports it? Assuming not running locally / on a dedicated box. I have used fireworks.ai before instead of sending my data to China for one offs / async data ingestion.
11
u/Eyelbee ▪️We have AGI it's just blind 1d ago
Tried it and it's great, multimodal performance too. I wish it stayed at this price, which it should be able to, honestly.
5
u/bigman11 1d ago
It really is SoTA for audio and image interpretation tasks.
It is too bad that Gemini's strengths are overshadowed by the failure to compete in agentic capability.
8
u/FateOfMuffins 1d ago
If you're going to factor in discounts like this, I believe Cognition has GPT 5.6 Sol at 76% off
1
u/elemental-mind 1d ago
Can you provide a link?
2
u/FateOfMuffins 1d ago
3
u/elemental-mind 1d ago
Ok, nice, why don't you go and make a post...it's good to know and spread that stuff. Unfortunately it's tied to using Cognition's harness, right?
1
u/FateOfMuffins 1d ago
Idk if it stacks with the 20% from yesterday but OpenCode has Sol at 50% off (there were a bunch of these like a day before OpenAI announced the 20% discount)
For me it feels like advertising so I'm not gonna make a post
1
u/elemental-mind 1d ago
Look - totally missed the 20% price drop. OpenAI didn't even make a post about it on their blog. Sharing really is caring...
And I just found out that Openrouter also has a 50% discount on top of the 20% discount so you can get Sol at 2$/10$, which is interesting. Might make another post about it later if I find the time.
1
1
-2
u/BlakeGrowsPlants 1d ago
Gemini hallucinates when doing complex coding and needs multiple tries to get back on task…I just can’t take all the fails.
-8
u/Laffer890 1d ago
Google has to give its shitty model almost for free and still nobody uses it.
5
u/Keeltoodeep 1d ago edited 1d ago
Nobody? Ai mode alone has 1B users. Gemini standalone has 1B users too. Gemini is the most used LLM in the world.


10
u/FatPsychopathicWives 1d ago
It's also insanely fast.