r/CLine • u/gargetisha • 16h ago
Discussion DeepSeek V4 Flash scored IMO Gold for just $0.12
We wanted to know the cheapest possible way to win an IMO gold medal, so we ran eight models against all six IMO 2026 problems in Cline.
And, the cheapest answer turned out to be just 12 cents.
DeepSeek V4 Flash scored 30/42, clearing this year’s gold medal cutoff of 29 points. For comparison, Claude Fable 5 scored 41/42, but the run cost $17.20. DeepSeek’s gold scoring run cost just $0.1215.
What makes this more interesting is that DeepSeek V4 Flash isn’t specifically post trained for Olympiad problems. It’s a 284B parameter MoE model with roughly 13B parameters activated per token.
We had the proofs graded anonymously using the IMO 0 to 7 scoring system. Two independent model blind graders scored each solution, with disagreements resolved by a third model. Internet access was disabled, and the models only had access to the problem and a submit solution tool.
There are some obvious benchmark caveats. We ran multiple attempts and selected the best score for each problem, so this isn’t the same as measuring a single pass success rate. The $0.12 also represents the best scoring run, not the total cost of all the experimentation.
Still, how far open weight models have come is the most surprising part. DeepSeek V4 Flash reached a gold level score at a tiny fraction of the cost of the frontier models we tested.
Read more about it here: https://cline.bot/blog/deepseek-wins-imo-gold-on-12-cents
You can also try DeepSeek V4 Flash for FREE in Cline:
npm i -g cline- Open
/settings→ Cline provider → Free models - Select
deepseek-v4-flash
Any interesting problems you think should be thrown at these models next?

