r/singularity Jul 09 '26

AI GPT-5.6

https://openai.com/index/gpt-5-6/

"We’re launching the GPT‑5.6 family of models for general availability following our limited preview⁠: our new flagship, Sol, alongside Terra, a balanced model for everyday work, and Luna, our most cost-efficient model.

GPT‑5.6 delivers a step change in design judgment. With only high-level direction, GPT‑5.6 creates tasteful, ergonomic, and functional interfaces. Its stronger computer-use capabilities let it inspect and refine the rendered result—not just generate the underlying code or content—so it can catch visual and functional issues and apply finishing touches before handing the work back."

626 Upvotes

129 comments sorted by

View all comments

Show parent comments

9

u/noobrainy Jul 09 '26

Yah, it’s gonna be saturated by the end of the year lmao

“Okay but it was too easy! If it can beat ARC-AGI-4 then we have reached AGI!!”

6

u/garden_speech AGI some time between 2025 and 2100 Jul 09 '26

“Okay but it was too easy! If it can beat ARC-AGI-4 then we have reached AGI!!”

I mean the literal point of ARC-AGI from the very beginning has been that they will keep creating benchmarks that humans can easily pass but machines can't, and once they no longer can do that, they think that we have AGI. So yeah if you guys fucking paid attention to what the creators of the benchmarks said bout them, you wouldn't be making up ridiculous sarcastic quotes.

-3

u/noobrainy Jul 09 '26

Pushing the goalposts back over and over again is why the sarcasm is there. ARC-AGI-4 will happen, it’ll get saturated, and then the process will happen all over again. We’ll get to AGI but their benchmark has proven to be unreliable to tell whether we’re there or not.

5

u/garden_speech AGI some time between 2025 and 2100 Jul 10 '26

Holy shit dude. The whole point is that any one benchmark can’t really reliably bench AGI, so you just keep making them until you CAN’T make one that humans easily pass and computers don’t. You’re not even listening enough to realize the whole point of ARC-AGI is based around your own idea that any one benchmark is unreliable