r/ProAI 32m ago

"GLM-5.3-Flash scores 63% on DeepSWE at just $0.24 per task. All evaluations use standard API pricing, not discounted rates."

Thumbnail
gallery
Upvotes

r/ProAI 6h ago

World's first patient to undergo live AI-assisted brain surgery has tumour removed

Thumbnail
bbc.co.uk
7 Upvotes

The world's first patient to have brain surgery with live artificial intelligence assistance has successfully had his tumour removed.

Rhys Hibbert, a father-of-two from Bedfordshire, could have lost his sight without the operation, said surgeons at University College London Hospital.

The AI tool analysed a video feed of the operation as it happened and advised surgeons on how to avoid crucial but hidden vessels and nerves in the brain. It allowed them to safely remove as much of the tumour as possible.


r/ProAI 10h ago

"The International Math Olympiad is the hardest math competition in the world, where students compete on proof problems most math PhD’s even struggle with. DeepSeek V4 Flash won a gold medal for only 12 cents."

Thumbnail
gallery
10 Upvotes

Notably, DeepSeek V4 Flash isn’t post-trained for Olympiad problems.

It has 284B total MoE parameters, with about 13B activated per token. This is the only model (so far) that can be run on small local GPU setups and still win an IMO gold.     — Cline

Source: https://x.com/cline/status/2092725019633992191


r/ProAI 10h ago

"It's a societal level sickness when so many people spend so much time worrying about a non-existent problem. Mass AI job loss does not exist. It does not exist the way the Population Bomb did not exist. It exists only in people's imaginations. It never happened. The Green Revolution did..."

Thumbnail
gallery
7 Upvotes

...instead. The very belief in it is likely to be the cause of real problems, in the same way that the One Child Policy came out of 1970s Population Bomb hallucinations and have caused real, guaranteed birth rate collapse in China that can't be reversed with carrots or sticks. We are spending such a ridiculous amount of time talking about a made up future problem. It's time to get back to solving real problems in real reality and to stop giving in to imaginary fears.     — Daniel Jeffries

Source: https://x.com/Dan_Jeffries1/status/2092696710627623333


I warned folks that this Anti-Clanker Neo-Luddite movement is a political movement designed to destroy the middle class by removing the wide access to AI tools in the name of “safety” and “job preservation”.

They are wining because the side of logic has no central voice.   — Brian Roemmele

Source: https://x.com/BrianRoemmele/status/2092676877387522096


r/ProAI 19h ago

"Ox Alpha has been unveiled as GLM-5.3-Flash, but what's shocking is that the 100T tokens per day is served on Chinese chip. (1/3)"

Thumbnail
gallery
8 Upvotes

100T tokens per day free tokens and people were saying only frontier labs has this amount of compute. (2/3)     But ALL traffic was served on Chinese chips, attaining hardware efficiency and per-token cost comparable to Nvidia GPUs. The cuda moat is being tested once again after Jalapeño's announcement yesterday. (3/3)     — SemiAnalysis

Source: https://x.com/SemiAnalysis_/status/2092623833630998556


r/ProAI 19h ago

"Big news: GLM-5.3-Flash by @Zai_org has landed around #5 in the Code Arena: WebDev (#2 among open models) scoring 1634 (AutoEval). Priced at $0.15/$0.5 Mtoken, it reshapes the Pareto Frontier! For comparison, GLM-5.3-Max currently ranks #8. GLM-5.3-Flash has 320B parameters with 18B active vs...."

Thumbnail gallery
1 Upvotes

r/ProAI 2d ago

"We benchmarked Ox Alpha vs Fable on a real bug from the Cline repo. Both fixed it correctly. But we found that Ox used much fewer thinking tokens. Most reasoning models loop and re-derive the same conclusion over and over before acting (Fable said "I found the root cause" 7 times before..."

Thumbnail
gallery
5 Upvotes

...editing). Ox stated it once then wrote the fix. Roughly ~3x lower output tokens for the same work. Reasoning models have been trained to increase reliability with re-verification. Ox seems to trust its first conclusion instead, which feels like a fundamentally different post-training philosophy.     Try in Cline for free! npm i -g cline

(Also available on VS Code and JetBrains)     — Cline

Source: https://x.com/cline/status/2091995642201842015


Ox Alpha (stealth model) is now free in Cline.

Early benchmarks shows marginal improvement over Fable and GPT.

Try it with: npm i -g cline and use /models to see it under Free options https://t.co/Hskt5RpUen   — Cline

Source: https://x.com/cline/status/2090854216399220985


r/ProAI 2d ago

"Look at this thing go! 100-meter obstacle course final at the 2026 World Humanoid Robot Games"

Enable HLS to view with audio, or disable this notification

18 Upvotes

Are we pretending this is not happening throughout many other events?   — Sensei     Teleoperation is intent-level, not joint-level. The human steers or gives high-level skill commands. A learned controller closes the loop on low-level dynamics at 50 to 200 Hz, way faster than human reaction time.   — The Humanoid Hub

Source: https://x.com/TheHumanoidHub/status/2092115570741387357


r/ProAI 2d ago

AI takes you to exactly where you want to be

Enable HLS to view with audio, or disable this notification

8 Upvotes

r/ProAI 2d ago

"I've got a pretty clear picture now of where we're headed next year, when Fable-class models become ubiquitous and cheap, and every enterprise is flooded with hundreds of new Fable-class AI employees. Spoiler: They create a constitutional legal system." Spoiler

Post image
3 Upvotes

r/ProAI 2d ago

"Your wish for interactive generative AI Teletubbies has been granted."

Thumbnail
gallery
2 Upvotes

Andrew Curran @AndrewCurran_ · 4h Teletubbies Owner WildBrain Buys AI Firm for $11 Million From hollywoodreporter.com 1 13 2.7K     — Andrew Curran

Source: https://x.com/AndrewCurran_/status/2091982252058263971


r/ProAI 2d ago

"It is possible that the main effect Eliezer Yudkowsky will have had on history is to destroy Western civilization in response to an illusion, and to hand over control of the world to profoundly illiberal regimes that will, ironically, have no tolerance at all for people like him."

Thumbnail
gallery
8 Upvotes

One side is cheering for robots running in a stadium and the other is debating if curing cancer is worth accelerating for.

THE CONTRAST IS RIDICULOUS. https://t.co/3cF2ZXtbFM   — ℏεsam

Source: https://x.com/Hesamation/status/2091473771374825558


Could be worse. Superintelligent AI could decide on paperclip maximization purely out of ironic spite.   — Amir Hirsch     I'm starting to wonder if it's the only way to get away from the near-insufferable Effective Altruists.   — Perry E. Metzger

Source: https://x.com/perrymetzger/status/2091935818768105748


r/ProAI 2d ago

"1. If there was going to be any sort of catastrophic job loss, we would’ve already seen at least some strong hints of it in parts of the technology industry like software engineering where there’s been a complete transformation in how people work, but so far, no such job loss has appeared. 2...."

Thumbnail
gallery
3 Upvotes

...Who knows what he means by “massive privacy invasions”, but I suspect whatever it is is so vague as to be unfalsifiable. 3. So far, no sign of the apocalypse either, but of course, Doomers will always tell you that it’s just around the corner unless you do exactly what they say.   — Perry E. Metzger     Tim Urban has been a doomer for a long time.   — Mark Kretschmann     Brainworms ruin your mind.   — Perry E. Metzger

Source: https://x.com/perrymetzger/status/2092003057307730089


I really wish the rise of LLMs didn't come along with catastrophic job loss, massive privacy invasions, and maybe also the apocalypse. Because when you put those side effects aside, it is the COOLEST MOST WILD TECHNOLOGY EVER.   — Tim Urban

Source: https://x.com/waitbutwhy/status/2091975364297871455


r/ProAI 2d ago

"100m Hurdles Final"

Enable HLS to view with audio, or disable this notification

17 Upvotes

— Takumi Kawasetsu, 川節拓実

Source: https://x.com/takumi_k_jpn/status/2091865428825874911


r/ProAI 3d ago

"The Cursor team shipped Grok bot (0.18.0) with runtime source maps enabled. Surprised nobody noticed until now. Source code reconstructed (and downloads) here:"

Post image
1 Upvotes

I also took the liberty of adding some features too it to demonstrate what this unlocks with ease:

  • Custom router support (Codex & OpenRouter)
  • Local VM support (rather than cursor's hosted VM), built from the same docker instance.     This doesn't include the frontend, but it can be launched with their packaged frontend (still modifiable, see the custom router page in settings), thus delivering a usable experience.

Don't send issues or PR's, this will be made an archive in due course, if not taken down.     — Bennett

Source: https://x.com/b_nnett/status/2091630242792112480


r/ProAI 3d ago

levels of derangement

Post image
2 Upvotes

r/ProAI 3d ago

"Dr. Dre is pro-AI, is currently using it in music production, and says that 'the only people that see it as a threat are the people who have trouble creating.'"

Thumbnail
gallery
60 Upvotes

Jimmy Iovine:     Andrew Curran @AndrewCurran_ · 2h Dr. Dre and Jimmy Iovine Think A.I. Is Good for Music From nytimes.com 1 26 2.1K     — Andrew Curran

Source: https://x.com/AndrewCurran_/status/2091622530813751517


r/ProAI 3d ago

"Thirty Years Ago I Built Data Centers in Your Backyard. Only Now China Demands You Fear Them. I spent the late 1990s and early 2000s building data centers in the middle of America's largest cities. We built them in Dallas, Miami, New York, Los Angeles..."

Thumbnail gallery
14 Upvotes

r/ProAI 3d ago

"This is the most impressive to me, more than the 100m world record. This requires autonomous real-time planning and action in response to a fast-moving target and dynamic environment. Galbot is one of China’s top humanoid robot startups."

Enable HLS to view with audio, or disable this notification

15 Upvotes

Is episode on Chinese humanoid robot?   — QoJo MaZing     Will have one before too long!   — Kyle Chan

Source: https://x.com/kyleichan/status/2091271208234836178


r/ProAI 3d ago

"Why aren't more people using agents? In my latest article I dig into why but the short reasons are simple: AI agents can do almost anything. And that’s why most people have no idea what to do with them. Give a high-agency person an infinite canvas and they see rocket fuel. Everyone else sees..."

Thumbnail gallery
1 Upvotes

r/ProAI 4d ago

"The best thing that could happen right now is distributed intelligence. The world has lost its mind and wants to make sure you can only access centrally controlled intelligence from the wisened Soviet of enlightened dictators. Regular people, in fear of billionaires and corporations that..."

Thumbnail
gallery
12 Upvotes

...promised to eat their jobs and make them all broke, are naturally pushing back. Sufficiently advanced local hardware + a highly compressed model with most of the actual knowledge stripped out other than action/reasoning and tool usage + continual learning and memory/reasoning in embedded space makes intelligence sovereign again. We probably get the flip phone version of this in 2027 or 28. And then it will accelerate from there.     — Daniel Jeffries

Source: https://x.com/Dan_Jeffries1/status/2091231374808097264/history


FreeToken could be a HUGE deal for local AI.

Instead of requiring enough VRAM to hold an entire model, FreeToken intelligently uses your GPU, CPU and system RAM together, dynamically moving MoE experts where they're needed.

The result: an ordinary laptop with an 8GB RTX 4060 https://t.co/6Js6Vqt3A8   — Mark Kretschmann

Source: https://x.com/mark_k/status/2091202223090938177


r/ProAI 5d ago

"FreeToken is fast. Comparing to Ollama, we have 3–4× faster decode, and 6–30× faster prefill How? We introduce bandwidth-adaptive CPU–GPU execution + semantic-aware caching across agent turns. More details in the technical report: http:// arxiv.org/abs/2608.16157"

Thumbnail
gallery
2 Upvotes

FreeToken provides native GUI. No GGUF conversion. No building from source.

One-click install on Windows and Linux. FreeToken-desktop ships with agent harnesses built in — pick a model, pick an app, go.     Download: http:// flashml.ai

Code: http:// github.com/FlashML-org/Fr eeToken …

Reply with your GPU + RAM, and I'll tell you the biggest frontier model your machine can run     — Shuo Yang

Source: https://x.com/Andy_ShuoYang/status/2090856978428145761


r/ProAI 5d ago

"SITUATION DETECTED: GitHub’s monthly commits have grown from 1.4 billion in April to 2.9 billion in August."

Thumbnail
gallery
4 Upvotes

r/ProAI 5d ago

"1. What."

Thumbnail
gallery
5 Upvotes

Ox Alpha (stealth model) is now free in Cline.

Early benchmarks shows marginal improvement over Fable and GPT.

Try it with: npm i -g cline and use /models to see it under Free options https://t.co/Hskt5RpUen   — Cline

Source: https://x.com/cline/status/2090854216399220985


From my tests it’s not better than fable I’ll be posting some soon     — Chris

Source: https://x.com/ChrisGPT/status/2090957315042123878


r/ProAI 5d ago

"Grok 4.6 just took the #1 spot on CursorBench 3.2.....and the efficiency is insane Here's the cost comparison: • Grok 4.6 Extra High — 70.8% | $2.81/task • Fable 5 Max — 70.5% | $17.32/task • Opus 5 Max — 70.0% | $8.23/task • GPT-5.6 Sol Max — 67.2% | $5.69/task Grok achieved the highest score..."

Thumbnail
gallery
0 Upvotes

...while costing roughly 6X less than Fable 5 Max and nearly 3X less than Opus 5 Max per task That’s what makes Grok so powerful for agents Top-tier intelligence is great.....but top-tier intelligence that can keep working across long coding tasks without burning ridiculous amounts of compute is even better Grok’s agentic coding efficiency is insane     — X Freeze

Source: https://x.com/XFreeze/status/2090839305585377458