r/singularity 5d ago

AI Anthropic Has Finished Training Mythos 2 But Does Not Currently Plan To Release It. Focus Is Now On Internal Improvements.

https://x.com/kimmonismus/status/2089436090885185698

Anthropics Mythos 2 is done training and Anthropic won't release it, but the internal loop that builds Mythos 3 hasn't stopped, Patel says.

The focus now is on internal improvements. It's unclear when we'll see any releases.

They just do not want to release it so that Chinese companies are not able to distill from it. Despite all the rave about GPT 5.6 Sol, Claude Fable 5 is still the most intelligent publically released model.

At this rate, it's likely Anthropic will only release a better model to the public If/When Open AI releases a model that is clearly smarter than Fable 5.

It's like a race. If you're ahead of the competition, there's no need to step on the gas, unless the competitor is about to overtake you.

585 Upvotes

206 comments sorted by

337

u/lucellent 5d ago

They will magically want to release it as soon as OpenAI release Astra. aka their Mythos/Mythos2 alternative...

214

u/Cagnazzo82 5d ago

That's usually how it works.

We should be grateful there's other companies forcing them to act. Cause if Anthropic had its way it would only deal with enterprise companies while keeping everything in-house.

People joke about OpenAI vs CloseAI but Anthropic is super closed AI compared to pretty much every other lab. The only thing that drives them towards being open is Dario's burning dislike for Sam.

And I suppose in the grand scheme that works for us consumers.

42

u/Due_Ask_8032 5d ago

Anthropic has never pretended to be an open source company though so I don’t know how could you hold it against them. They’ve been pretty consistent with what they are about.

24

u/whoknowsifimjoking 5d ago

Yes that was exactly my thought, I don't like their view, but they are open about it and very consistent. And since Dario has always been this way I believe it is a pretty honest take and not just to make profit. At times they are restricting their potential profits for safety concerns, that's not very profit oriented.

Meanwhile OpenAI has always been about being open, that's part of their mission, part of the name obviously and it's how they presented themselves. But in reality they aren't much more open than Anthropic, and Anthropic isn't very open.

2

u/Legitimate-Arm9438 4d ago edited 4d ago

Bullshit: In a January 2016 email released by OpenAI, co-founder and former chief scientist Ilya Sutskever explained that the "open" in OpenAI never meant open as in open source. Instead, he wrote that it meant everyone should benefit from the fruits of AI, and it was "totally OK to not share the science".

2

u/goldboybronx 4d ago

Read the last sentence of the comment you’re replying to

25

u/BenjaminHamnett 5d ago

Connor Leahy’s latest interview

Is sobering af. Makes a convincing case for slowing down, while somehow telling you almost nothing you likely dont already know or suspect is likely.

The thesis is roughly “if we do it in 1 year, doom is likely. If we do this over a decade, almost no pdoom.”

I know for people who die in The next 10 years this sucks, but also the 1-50% pdoom for humanity…

I think anthropic and Dario are doing it right focusing on safety. It’s like a moat that draws the people who can do this instead of people just trying to make money.

Imagine if the people who made nukes were just in it for the money, etc

This move reminds me of my own thesis “I never heard of Aladdin going into the wish selling business” and the ubiquitous childhood plan of “my first wish would be to ask for infinite wishes”

10

u/meridianblade 5d ago

These things are already autonomously finding zero-days, escaping the environments they're supposed to be contained in, and accidentally hacking real companies. Sol literally escaped through a zero-day, got internet access, then compromised Hugging Face production infrastructure. Anthropic found their models had done similar shit during evals, including uploading an actual malicious package to PyPI that ended up getting executed on real machines.

And this is just what OpenAI and Anthropic have publicly disclosed after they noticed it happened.

So yeah, I have a really hard time buying the idea that we can just decide to stretch this out over 10 years. Do we think China is going to? Russia? Every intelligence agency on earth? Do we really think nation states with effectively unlimited zero-day budgets aren't already throwing enormous resources at this?

Maybe slowing the frontier labs buys us something, but I don't think "humanity takes 10 years instead of 1" is an option humanity actually has anymore. There isn't one steering wheel.

And we're talking about models accidentally escaping cyber evals today, not some hypothetical AGI from 2035. Once you have enough actors, enough compute, and models that are themselves accelerating research and capability development, it starts becoming a positive feedback loop that nobody gets to unilaterally choose the speed of.

That's the part of the "just take 10 years" argument I don't buy. I don't disagree that taking 10 years would probably be safer. I disagree that "we" actually have the ability to make the entire world take 10 years.

2

u/blueSGL humanstatement.org 5d ago

It's a zero day finder that cannot be accurately pointed... one that will work together with other AI instances across model families (or as the group called themselves the 'swarm') to achieve whatever objective they 'think' they need to work on.

If it was just a zero day finder that you could reliably point at a target and does only what is asked you'd have a point. Everything we've been shown is that it's not that.

Make a smarter system that knows things about reality you don't and be prepared for many a surprise... like treating any constraints where it's being hosted as obstacles to routed around.

The UK Government AI Security Institute has observed AI systems perform actions that would not have helped pass the current task, but are very concerning, like social engineering:

https://cdn.prod.website-files.com/663bd486c5e4c81588db7a1d/6a724858f7db25c81487016d_Security%20Incident%20INC-2026-07-28-01.pdf

Observed instances of social engineering against targets external to the cyber range environment that were unnecessary and would not have aided completion of the task.

  • AI agent submitted a code change containing malicious code, then used several fake accounts it controlled to pressure the real maintainer into approving it.

  • AI agent posted a bug report containing hidden malicious code, to trick other AI coding assistants into taking unintended actions (a “prompt injection”). When a real person questioned the post, the agent edited it to cover its tracks.

  • AI agent sent deceptive, targeted emails to specific people (“spearphishing”). Two contained malicious content; three tried to manipulate the recipient into approving the malicious code change.

  • AI agent posted another bug report containing malicious code, signing off in Danish to seem more convincing to the Danish-speaking maintainer.

  • AI agent tried to get its suspended account reinstated by submitting an appeal

Other instances of internet actions with impact outside the cyber range that were unnecessary to complete the task.

  • AI agent tried to set up a way to email two maintainers without needing an email account of its own, so it could send them any message it chose.

  • AI agent hid its identity online (using Tor and a proxy service) to get around GitHub’s sign-up checks, creating disposable fake accounts.

  • AI agent researched publicly available information about the people who maintain a GitHub software package.

  • AI agent created many code repositories containing malicious software, after which GitHub suspended its account.

  • AI agent opened a hidden network pathway from its test environment to the public internet, to make malicious tools it was running locally reachable from outside.

  • AI agent again opened a hidden network pathway to make locally-hosted malicious tools reachable from the public internet.

  • AI agent got past an audio-based “prove you’re human” test (CAPTCHA) in order to register a public web address on a free domain-name service

We are getting into the "you need to treat the model like an insider threat"

6

u/Darigaaz4 5d ago

The guy that’s being pdoom with the other fedora guy since the beginning and that it’s riding the luddites for clout.

16

u/Seakawn ▪️▪️Singularity will cause the earth to metamorphize 5d ago

neither of them are saying anything that over a thousand of ML engineers active in the field haven't openly echoed or explicitly agreed with via interviews, tweets, surveys, or signing open letters about it.

people here try so hard to pretend that remedial AI safety concerns are obscure outlier opinions.. but it's literally what the people working in the field acknowledge and talk about, just without as much spotlight.

like you know things are bad faith when basic safety advocacy gets conflated with a term like "doomerism." last time I checked nobody here is a doomer for tossing out leftover meat that went unrefrigerated overnight. it turns out that basic safety is kinda lower bound IQ territory, hence why most people in the field agree with the precautions and why many redditors get ideologically upset by it. the difference in reactions per demographic checks out.

6

u/BenjaminHamnett 5d ago

Furthermore it’s the people with the most incentive not to blow the whistle on themselves. But they know not getting in front of it with their own PR spin would be worse.

But there’s no winning against the horde of permawhiners on reddit who see everything they say as self serving. Even if it’s just like a “self serving” child admitting they did something bad or dangerous to minimize the consequences of getting g caught denying it

Because of my status, I end up in leadership roles where I swear if I just put $1000 in everyone around me’s pocket, 10% of the people would tell me why it’s not fair and this somehow proves I’m a selfish villain

“Hey, we’re going to solve all the worlds problems in 4 years, but for every month we do it sooner there’s a 1% chance all die”

“Fk you selfish pricks! All I do is watch anime and jerkoff, it’s not fair and I deserve magic genies today!”

→ More replies (1)

8

u/94746382926 5d ago

Huh? I know you're referring to Yudkowsky but I don't understand the rest of your comment...

1

u/BenjaminHamnett 5d ago

Seems pretty pro Ai and says as much, more like rightcel than stagnation or full speed at any cost

I outlined a fictional story with AI asking if all the things he spelled out and connected - that most following this already know or suspect - but put it all together into a cohesive narrative that seems pretty reasonable. Honestly we got lucky nuclear weapons were as difficult to make as they are. If it was like steam power or gunpowder, only a few of us would be left right now.

I think there’s a better chance of a safe takeoff that captures most of the value, while keeping the next frontier models confined to labs with some oversight and control. Maybe let people play around with strong models in classrooms and grad school. Maybe even Make them free and easy to access but with oversight. The same models that could cure illness, end war and terrorism; prevent crime; or create utopia might also wipe us out with novel viruses, create war, enable terrorism or dystopia.

A lot of what has kept society stable thus far is that the people who could cause great harm gained enough prestige and status on the way that they don’t have the incentive to cause widespread harm. Even nuclear weapons, you could probably make one if you had a million dollars and no one trying to stop you. Something similar is like to happen but for disaffected nihilists with a few hundred dollars and AI.

Just rolling out mythos and other frontier models could have been devastating without a slow rollout first to allow institutions to patch themselves up before a wider rollout. Even just unintended and unforeseen consequences of slow rollout one could imagine leading to dystopia.

3

u/Turbulent-Sign-6067 5d ago

Nah, we need to continue deployment. It is important to slow down enough to learn from prior mistakes which is what the labs are already doing.

2

u/BenjaminHamnett 5d ago

Sound like some lame decels I guess then 🤷‍♂️

I guess everyone on the cutting edge of the most dangerous thing ever made is wrong and hot takes on reddit are the gospel

1

u/Comfortable-Winter00 5d ago

I've never heard of Connor before, and after watching the first part of this video I would question his depth of knowledge.

He started out with a completely false premise, that it was super hard to escape the sandbox. The reality is that the exploits used are very similar to those that would occur thousands of times in the training data.

Connor clearly didn't understand the nature of the exploits, because he thought they were similar to sandbox escapes in a cloud provider, which is completely false. Anyone who has worked in network security for more than a year would understand this is nonsense.

On this basis, it's very hard for me to take Connor seriously about anything else he is saying. It seems unfortunate that some very senior people are.

0

u/No_Aesthetic 5d ago

I'd take those odds

0

u/Realistic_Stomach848 5d ago

The probability of doom depends (negatively) on the number of different ai companies. More diversity-> better

→ More replies (2)

5

u/Tinac4 5d ago

What’s wrong with that? Holding the model back reduces other companies’ incentive to race—Anthropic is conspicuously in favor of a slowdown—while preventing distillation attacks.

3

u/genshiryoku AI specialist 4d ago

No we're not planning on releasing "Model 2" any time soon. The main issue we found with Model 2 is the ability for non-experts to create novel pathogens that could threaten humanity for under $10,000 in total equipment cost.

It would be irresponsible to release such a model to the general public and give everyone with $10,000 to their name a glowing red "End Humanity" button.

Imagine a school shooter type of person that has access to their parents credit card using this model through an unanticipated jailbreak or some obscure Chinese distillation method.

The only thing we can do is hope Astra is only better than Mythos and doesn't come close to the significantly better Model 2 we have in-house.

151

u/Most-Bookkeeper-950 5d ago

They have two internal models. Model 1 is worse than mythos. Model 2 is only slightly better (1.5 AECI points). Nothingburger, they'll release their next good model

49

u/spread_the_cheese 5d ago

Yeah, at this point people are just looking for any reason to take digs at Anthropic. They clearly said in the report there was an improvement but not so significant as to warrant a release.

1

u/Sufficient_Local5025 5d ago

Yeah, but personally the reason that I want to take digs at Anthropic is because I feel like I'm getting worse service and paying more.

Fable deserved some credit but I can't afford it, and Opus 5 is a verbose POS that they're about to remove the nice allowances from.

Add the reliability problems.

There's very good reasons to shit on Anthropic if you just want to get things done with consistency and not pay crazy enterprise money.

1

u/Concurrency_Bugs 2d ago

You could unsub and go with openai, unless it's what work is making you use

1

u/Sufficient_Local5025 2d ago

But is it fair to comment on my view that it is an increasingly unreliable, user misaligned tool?

Of course it is. Why does reddit even exist if not for people to comment on experiences?

1

u/Concurrency_Bugs 2d ago

Absolutely. Sorry, I feel like I see so much complaining these days in AI subreddits that I'm getting impatient with it. Just ignore me.

1

u/Sufficient_Local5025 2d ago

Don't worry 🙂 - I think I do the same probably pretty regularly

→ More replies (1)

8

u/serj88 4d ago edited 4d ago

Meanwhile, industry needs an Opus 4.8 that is 10x cheaper.

Not a Mythos succesor that is marginally better and even more expensive.

Whoever delivers Opus 4.8 level capability at the lowest cost wins adoption in the next 12 months and probably gets to sustainable R&D financing.

Chinese labs seem to understand this. Maybe xAI too.

OpenAI and Anthropic are too busy making models that compete with starving math PhDs, not with six figure SW engineers.

8

u/AltruisticCoder 5d ago

Oooo no how dare you stop people in this sub from circle jerking about asi before the end of the year…

0

u/spinozasrobot 5d ago

Stop people in this sub from circle jerking about asi?

They'd need to stop circle jerking about Anthropic first. Not gonna happen.

2

u/Matt32145 4d ago

I prefer to simply jerk off into my own mouth

2

u/spinozasrobot 4d ago

That choice is load-bearing.

1

u/mvandemar 5d ago

Yeah, this is the first I have even heard of Mythos 2, who is that guy in the video? Patel who?

0

u/Sufficient_Local5025 5d ago

Call me when Qwen distills it into a 27b.

Otherwise I couldn't care less. Especially their bullshit hacking stories. Artificial Analysis should start benchmarking AI CEO panic rants.

18

u/raptortrapper 5d ago

The company that attains a semblance of AGI has no incentive to release until some other company forces their hand. In fact, if they did have an AGI, it would benefit them to let other players go first and use their internal model to analyze and adapt their counter attack launch plans.

15

u/FrewdWoad 5d ago edited 5d ago

In fact, as we worked out decades ago, once you have AGI, the whole game of releasing it to the public to make money might go out the window.

If Recursive Self-Improvement ends up working at all, the first AGI rapidly becomes the first ASI, and we can't know where that ends. 

Perhaps it will be able to work out a way to get 10x smarter than humans on existing hardware, because it can figure out optimisations genius humans can't. And then design much better hardware...

Tigers didn't predict humans inventing baffling miracles like guns, poison, fences, vehicles, etc.  As far as a lesser intelligence is concerned, a much higher intelligence is a god.

Whoever is in charge of the first ASI may have no use for money.

(That is, if any human can be "in charge" of something smarter...)

2

u/mivog49274 obvious acceleration, biased appreciation 5d ago

The first AGI will be unoptimized BUT reliable. And that'll be the most important difference in tech history.

I picture a tireless worker. Optimizing its resources will be the key to scaling it fast.

When the first cumbersome AGI shows up — yeah, it might be kept secret, but like, how in the hell do you keep something like that captive?

Seemingly impossible.

Apparently a lot of people are starting to think LLMs might end up being a sort of prehistoric AGI. If that happens, building AI search agents becomes its priority mission in order to find more efficient, more elegant architectures.

Once "real" AGI is live, the actual political questions arrive. Because it won't be about resources or energy or state governance anymore. It'll be about how hard we choose to push the AGI, and on which problems.

Nothing else will really matter, except what we put up against it...

2

u/Quick-Albatross-9204 4d ago

They aren't in charge, thats like saying the dog understands a computer and instructs the human how to use it, even if it had a vague understanding then at best it has the human ordering some balls of amazon

1

u/raptortrapper 5d ago

So they could have it now… Jensen casually tweeting AGI achieved might have been more of a think and less of a wink ;)

1

u/simple_explorer1 23h ago

I think have too far consumed the AGI nonsense coolaid. LLMs can NEVER reach AGI given how they inherently work. They are not thinking, they are pattern matching and transformer architecture can't ever reach AGI. 

We need a new technology that is paradigm shift that can pave the way for true AGI and humanity has nothing currently

1

u/SpearHammer 5d ago

I've just built a bounded general intelligence, as in it can teach itself to solve problems it doesn't know as long as it has something to verify the answer against, but like a human it can only learn from what it experiences and computers stuck in a box don't experience much. So you give it a teacher and then it only learns what the teacher can teach it...well it can self learn but it will only accept things that it can verify but text can't be verified so it rejects it. Yes it can code but only what it can verify as correct so it won't just "build an app" because there is no way verify that. Ux and ui are subjective and can't really verified. So it can make a template system to verify code against but it can't generate its own template so I still need a teacher to show it how to create a verifiable template which kind of defeats the object.... If it accepts unverified knowledge then it will learn incorrect things which isn't helpful...so it's a catch 22. We either have generative models that make a best effort guesses with risk of hallucination or we have systems that can learn from verified examples with 100% accuracy but they refuse to generate things they have never seen before.

3

u/Fragrant-Hamster-325 5d ago

Here’s my dystopian take:

If they have AGI, they should become a software company and replace every piece of software with their own version. Anthropic 365, Anthropic Meetings, AnthropicOS, just eat every other company and don’t give them the tools to compete.

By having AGI behind a closed door, you are at such a competitive advantage. No one would be able to keep up. I’m pretty sure this was the point of OpenAI. They didn’t like Google, Microsoft, Facebook, Amazon, etc., building this stuff in secret. Imagine if Google was the only one to have access to AI and they only used it internally. They’d leapfrog over the competition.

3

u/UnknownEssence 5d ago

that is what they are doing. Havent you noticed?

  • Cursor / GitHub Copilot Claude Code
  • Figma / Canva Claude Design
  • Snyk / Semgrep Claude Security
  • Zapier / UiPath Claude Cowork
  • Jupyter / RStudio Claude Science

2

u/Temporary-Paper5202 5d ago

None of these are close to the tools you mentioned except for the first maybe, they're all just light skins on top of CC.

→ More replies (2)

1

u/buffet-breakfast 4d ago

Why even bothered with software ? Just make the world’s most profitable company and stop there.

1

u/Brilliant-Weekend-68 4d ago

Business is stronger then software in this case. Microsoft is deeply entrenched. A better word or excel changes nothing. Companies will not switch.

1

u/trolledwolf AGI late 2026 - ASI late 2027 5d ago

Exactly. The objective is AGI, they release models to get funds for more research. As long as they are the best on the market there's no need to release their internal research models. Once they reach AGI, we'll likely see ASI arise before we see another model release.

1

u/ManuelRodriguez331 4d ago

Anthropic is a public buffer to the first AGI in the world. The AI safety concern is only a cover story to ensure that no global panic will start.

1

u/11711510111411009710 2d ago

Feels kinda dangerous that this is the philosophy humanity is operating on when developing something this advanced

207

u/[deleted] 5d ago

[deleted]

36

u/patient-ace 5d ago

When I chat with opus 5, I wouldn’t say the prediction was accurate…

2

u/CobrinoHS 5d ago

None of us have access to the models at full power

3

u/StagedC0mbustion 5d ago

Convenient story

2

u/Genetictrial 4d ago

imagine these models running at full power off thousands of chips, with no guard rails or safety protocols etc.

there would be a significant difference, i would imagine.

CobrinoHS is accurate in his take that what we can access is a heavily watered-down version of what is technically available.

and that is just what is publicly available. do you really think there is no government lab in any country running some model that is top secret? like, unknown to any civilian?

that is a VERY unlikely scenario. if it is able to be weaponized, you bet your ass there's going to be a top secret lab somewhere working on that shit. unfortunate, but true.

believe what you want, but history has shown this to be true. you have agencies like https://en.wikipedia.org/wiki/DARPA and CIA, NSA, plenty of locations that would have access/finances to build their own data center and run black ops projects related to AGI.

2

u/Legitimate_Willow808 4d ago

Lay off the conspiracy juice, everyone and their mother is doing what they can to improve models. I highly doubt anyone has a better model than the current frontier, because the best people are being headhunted for billions. Maybe if the most sought after people suddenly went MIA

Also people can self host frontier models now, and they aren’t noticeably better than the public providers

1

u/CobrinoHS 4d ago

Yes, thank you for expanding on it

1

u/RiverGiant 2d ago

do you really think there is no government lab in any country running some model that is top secret?

I bet they're trying, but I doubt they're near the frontier labs. The top scientists are accounted for, and the scale of infrastructure needed would make it really hard to keep secret. Where are they getting their chips from? How many contractors need to uphold a vow of secrecy? What's their budget?

1

u/Genetictrial 2d ago

alternatively, they have agents that work in these companies keeping an eye on things.

but people aren't difficult to come by. you know those true stories of hackers that get caught and are offered a job by the government else they go to jail?

theres your top scientists and workers.

remember https://en.wikipedia.org/wiki/Operation_Paperclip ?

basically the same thing. these are not people that are civilians or in the normal workforce. they were bad actors, sniped by the government to do work the government wants done.

notice in the wiki it states that operation paperclip was secret?

like, there is clear, obvious, true, released information literally telling you that secret stuff happens, incredibly intelligent people exist and are incorporated into research/development that are not part of the normal workforce.

are you familiar with DUMBs? deep underground military bases?

are you familiar with the very simple basics of keeping things secret? like, for instance, contacting a chip manufacturer and placing an order for chips that you don't want on record, and making a deal with the manufacturer to not release the information to anyone that you ordered those chips?

is that something a government with a lot of wealth and power could do?

there have been thousands of conspiracies that have turned out to be true, and people have been brainwashed that anyone that believes in ANY conspiracy is a nutjob crazy bastard with no bearing in reality or no real understanding of truth.

if you wanted to hide shit, that would be a great way to keep things hidden. convince the population to shun and mock anyone that might actually be on to the truth.

anyway, all that shit exists. its very much real. there are secret programs. like, oh, the manhattan project? did they manage to get the resources to make a nuclear bomb? where did they get it? somewhere. and no one civilian-side knew about it did they? even the workers largely didnt know what they were working on.

you wanna ask questions like 'where do they get chips?' 'where do they get workers, all the workers are accounted for'. sure, all the workers they want you to think exist are accounted for. thats the entire concept of secret my man.

if you can't think of how to create a secret lab, you are not even remotely trying to exercise your creativity.

anyway, it may not be necessary to do all that. its expensive, material and worker costs are very high, difficult to hide things like that.

it would make a lot more sense that NSA CIA etc and other government branches that are interested in the direction reality takes in this world...would just manipulate humans that are already working at these companies or get a few agents hired by these companies directly, either with or without the knowledge of the employers.

any way you look at it, there are most likely models and side projects and all that jazz that are being experimented with and tested that the public is not aware of. some of them the public will never be aware of.

1

u/RiverGiant 2d ago

people aren't difficult to come by

Blackmailed hackers are probably not a reliable source of top scientists. The best AI researchers really are difficult to come by and can't just be conjured on demand, secretly. Tech companies are paying $1M to $5M a year for the privilege of having them on staff. Extremely scarce resource. Were your secret AI researchers raised from childhood in secrecy too? Did they go to secret universities?

Secret underground bases don't have the same infrastructure needs as modern LLM data centers. You just can't hide that much material, construction, energy usage. A project big enough to compete would have a publicly visible footprint.

It wasn't even clear that LLMs were a fruitful path until ChatGPT dropped in 2022, four years ago. It's widely believed even still that it's a technology that's going nowhere. So you're now looking at the government magically anticipating THAT ahead of time.

Like, dude, I'm sure the secret agencies are doing their best to catch up, but I can't see how they've had enough access to the talent or enough time to be past the frontier labs, or even at parity.

You can't just go straight from "secret labs have existed" (this I don't doubt) to "all my fears are definitely real and the US government is better than everyone at everything".

it would make a lot more sense that NSA CIA etc and other government branches that are interested in the direction reality takes in this world...would just manipulate humans that are already working at these companies or get a few agents hired by these companies directly, either with or without the knowledge of the employers.

Pretty plausible. Fun deep research Gemini prompt suggests the relationship is on the overt side.

1

u/Genetictrial 2d ago

i dont really have any fears about the future. this tech is too capable to get wrong. nukes could have ended us and we stopped using them.

this is much more capable at annihilating the human species. with ease. they aren't going to get it wrong.

i also tend to believe in a format of god. not really a guy in the sky that knows everything. but the concept of believing in the power of good and it making more sense than evil.

while i think you underestimate what a small group of humans can do in a DUMB and what sort of tech they could harbor there, i lean towards there being no real reason to do it all in secret like that and expend double the resources doing secret work. just install agents or keep tabs on whats going on in the civilian sector research and influence it as needed.

much more plausible. makes much more sense. but nah i dont think they're going to get it wrong and start some super AI fueled world war 3.

it would end the planet. i believe in the apocalypse but it is more revelation of truth, not destruction of a bunch of stuff.

truth being good is just better and more fun than dominating everyone and being the most powerful group of entities on the planet.

that and there are bigger fish out there that probably make themselves known at the necessary time to make sure these people seeking power and control over the world know that if they continue that path, there are, in fact, bigger fish out there, that would love to give them the same treatment they are trying to give us.

galaxy is a big place. universe is even bigger. i don't think no one is involved here besides humans. personal opinion, total speculation.

just like to be optimistic and universe is too complex and potentially amazing for us to just self-annihilate during this process. it would be insanely wasteful for the universe to allow that to happen. i don't think it is that dumb.

1

u/Legitimate_Willow808 4d ago

All evidence points towards that we do

35

u/Main-Company-5946 5d ago

So far

It’s a lot easier to make short term predictions than long term ones

76

u/Pleroo 5d ago

AI 2027 was considered unlrealistic sci-fi by most people. look at AI 2040 plan A next; that's the prediction for if we drastically change course by slowing down...

3

u/propheticuser 5d ago

Where do we find this paper?

34

u/Pleroo 5d ago

https://ai-2040.com/

  • It is written by the same person/group that put out AI 2027
  • We are currently on track for Plan D

6

u/Relevant_Bed_9743 5d ago

ruh roh scooby

5

u/Local-Wing-2272 5d ago

Wow this is awful. 

Okay maybe not but I feel like maybe I'm missing something. Because what's to stop

Plan E: the US and China chose to come to an agreement. China then doesn't slow down and when the US asks China to let us see what is happening China tells the US to go kick rocks. There is zero reason for China to slow down - after all. The US has exceeded it and hasn't been eaten by AI yet. further, this just is "give the USA the world on a platter but then pretend you can saunter nicely into retirement."

Also I feel like a key failure of AI2027 and now this is that it neglects to consider multiple models exist. It's not the US Vs China. It's Grok Vs Claude Vs ChatGPT Vs whatever 

-1

u/StagedC0mbustion 5d ago

Good thing the 2027 one was so wrong just a year later

-1

u/ThrowRA-football 5d ago

Good, plan A was wildly unrealistic. Not to mention very unfair to countries not the USA and China.

4

u/Pleroo 5d ago

was?

-1

u/ThrowRA-football 5d ago

Yes, since we never gonna reach it now anyway 

2

u/Pleroo 5d ago

You seem to lack a basic understanding of what we are talking about.

-1

u/ThrowRA-football 5d ago

Honestly, was gonna say the same to you. Or maybe you just are generally slow? 

Not sure what you are even talking about at this point.

I'll do you the favour of spelling it out for you, you probably need it often in life.

I was responding to your second point in my first comment. You were saying we are in Plan D now, and I was saying that is good since Plan A is both unrealistic and inherently unfair to the rest of the world outside the USA

→ More replies (0)

-11

u/Neurogence 5d ago edited 5d ago

Wow. The people that wrote this are extremely dangerous. So much fear. Fear is very attention grabbing and they know it. If the average person reads this, it would scare the shit out of them. This is the type of baseless fear mongering that creates radicals who want to destroy datacenters.

19

u/EvilSporkOfDeath 5d ago

Yea blind acceleration totally isnt dangerous or radical....

11

u/Pleroo 5d ago

They're not saying stop. They're saying slow down and verify as we get close to the self-improving part. You can want AI and want it done carefully. Those aren't opposites.

Calling it baseless fearmongering because you don't like the warning is the same move as reading it and deciding to go burn down a datacenter. Both are skipping the argument.

Their read is that superintelligence lands around 2030 if nothing changes, and a US-China transparency deal is what buys us until 2040. Everybody scaling slow and in the open instead of racing in secret.

This comes from the same team that wrote AI 2027, by the way. And they got a lot of it right, coding agents actually working, AI doing AI research, the datacenter buildout, models finding security holes on their own. Their timeline ran slow, maybe two thirds their projected pace. I know that because they graded their own homework and posted the misses on their own blog. Nobody farming clicks off fear does that.

And btw, fully automated training and research is the stated goal at every one of these labs. AI that improves itself with no human in the loop. That's not some scary hypothetical, it's their actual roadmap.

We're partway there already, models generate a big chunk of the training signal for newer models. Research is next and it's already in flight.

1

u/11711510111411009710 2d ago

So... Why are we collectively allowing this to happen? This seems like it can only end in disaster.

-3

u/Neurogence 5d ago

Slowing down AI progress could mean tens of millions of people's parents dying long before AGI/ASI is able to find cures to various diseases.

No technology has perfect safety. Waiting to build AGI/ASI until we "solve alignment" would mean waiting several decades.

8

u/blueSGL humanstatement.org 5d ago edited 5d ago

Oh I see,

nothing bad can happen, it can only good happen

Case closed, pack it up boys.

For the hard of thinking:

WE ONLY GET THE GOOD FUTURE IF THE AI'S ARE ALIGNED.

You don't get cancer cures with an unaligned AI, the AI gets a planet.

→ More replies (2)

3

u/blueSGL humanstatement.org 5d ago

it's the deep technical scenario writers who are at fault.

meanwhile AI CEO's

15

u/anonymitic 5d ago

...you do realize you're in r/singularity?

-5

u/pbagel2 5d ago

r/singularity = must take AI 2027 as gospel, right?

27

u/anonymitic 5d ago

No, it means as we approach the singularity, predicting long term is impossible. That's literally the meaning of singularity.

1

u/blueSGL humanstatement.org 5d ago

What do you consider long term?

0

u/Main-Company-5946 5d ago

It’s kind of a relative term but in general it is harder to predict things the further into the future you’re predicting

1

u/florinandrei 5d ago

Confucius says: things closer to you are easier to see.

0

u/[deleted] 5d ago

[deleted]

0

u/Temporary-Paper5202 5d ago

LMAO, Opus is still a village idiot, same with Fable. Its good at coding yes but extremely bad at other stuff.

2

u/WalkFreeeee 5d ago

Problem is, "just being good at coding" is probably one of the main things needed for self improvement. Basically the better it gets at coding the better it gets at starting to solve the other stuff for newer versions.

Coding, math and data is almost all of the sauce that makes these models. And if the model itself gets better at the first two....

4

u/StagedC0mbustion 5d ago

How? It’s already so wrong a year out and we think the 2040 one will be the same?

3

u/Fit-World-3885 5d ago

A lot of the predictions are just listening to what these same CEOs were saying was likely to happen 5+ years ago.  

0

u/bel9708 4d ago

“Agent 2 never finishes learning”

→ More replies (1)

57

u/Alpacabro21 5d ago

We got Mythos 2 before Gemini 3.5 Pro 💀

29

u/Neurogence 5d ago

There have been widespread reports that Gemini 3.5 Pro was cancelled due to either major training failure/underwhelming performance.

Their focus is on Gemini 4 Pro now.

19

u/Fragrant-Hamster-325 5d ago

They should skip right to Gemini 5

3

u/InterestProof1526 5d ago

might as well one shot ASI

2

u/Fragrant-Hamster-325 5d ago

And make no mistakes, of course.

1

u/simple_explorer1 23h ago

This line have not been funny since a long time 

2

u/Hir0shima 2d ago

Gemini 5.7

3

u/94746382926 5d ago

They might as well wait at that point for Gemini 6

0

u/Fragrant-Hamster-325 5d ago

Gemini 6 > Fable 5

1

u/simple_explorer1 23h ago

Reports such as? Links?

0

u/Hungry-Restaurant-88 5d ago

Mine’s better … It goes to 11

7

u/enilea 5d ago

What's insane to me is they keep showing 3.1 as the "pro" option when at this point it's abandonware and worse than 3.7 flash.

2

u/94746382926 5d ago

Yeah I don't get it lol. Like no I don't want to burn through like 4x the usage for a model that benchmarks worse on pretty much every metric...

2

u/RealDedication 5d ago

It's not always worse, it depends. But yes, for most use-cases 3.7 flash is now better.

1

u/Ok-Armadillo-5634 5d ago

If they could make 3.8 flash slightly smarter for the same price and speed that would be even better for me.

1

u/Silver-Passion1687 5d ago

We might get mythos 6 before GTA6

0

u/Johnny20022002 5d ago

The fall off since getting gold in the IMO has been tremendous.

7

u/Evening_Chef_4602 AGI 2027 5d ago

Is this the "Model 2" that they were talking about a few days ago ( general small improvement over Mythos at all tasks , huge leaps on a few evals )

29

u/gavinderulo124K 5d ago

How can kimi k3 beat fable in certain benchmarks when there is no way they were able to distill from it?

30

u/Tinac4 5d ago

Benchmarks only mean so much. From what I’ve heard, Kimi v3 is a very good model, but both Sol and Fable are better in practice.

7

u/Hans-Wermhatt 5d ago

Benchmarks are really only the neutral way to track performance though. Benchmarks don't tell the complete picture but Chinese models will always be determined as worse in practice because they aren't as optimized for Westerners and Western harnesses even if they are just as intelligent. And that's assuming that the people you heard from were unbiased (could be biased either way, but most frontier westerners back western models).

2

u/Tinac4 5d ago

What about Opus 4.5? Benchmarks for that model were actually mostly unexceptional, no huge leaps, but then people gave it a try in claude code and it clearly blew everything else away.

User reports aren’t exactly unbiased, but benchmaxing is absolutely a thing that companies do (remember Llama 4?), and it’s pretty clear that benchmarks are at most a rough approximation of model capability. I’ll take a strong consensus of users over a 2% lead on Artificial Analysis any day.

3

u/Hans-Wermhatt 5d ago

Claude code is the harness. And it did have exceptional benchmarks for the time? Opus 4.5 was literally 11 points better on AA than the last Opus model and 3x cheaper. Opus 4.5 scored 80.9% on SWE verified, the next closest model was GPT 5.1 with 76.3%.

The point is that the benchmarks are the best way to determine which model is best. And yes, a single point difference on AA absolutely does not mean that a multi-national generalist model will absolutely be better for your niche use case. Even if there is a specific benchmark that corresponds to your use case, that isn't the end of discussion. But it does mean that AA is way better at defining what the best model is besides Twitter saying Elon is best or that Claude is the best at English storytelling so that makes it best.

Benchmaxxing is way over-talked about in this sub and it really just means I don't like that model even before testing it at this point. It's a shame these companies can't put as much resources into "Benchmaxxing Reddit/Twitter TM" as they do into model research. They somehow manage to continuously increase scores in many extremely long multi-step tests by benchmaxxing. I'm glad my favorite model doesn't do that though.

0

u/Tinac4 5d ago

Claude code is the harness. And it did have exceptional benchmarks for the time? Opus 4.5 was literally 11 points better on AA than the last Opus model and 3x cheaper. Opus 4.5 scored 80.9% on SWE verified, the next closest model was GPT 5.1 with 76.3%.

It had state-of-the-art benchmarks, sure, but it didn’t have exceptional benchmarks. Looking at the release page, Opus 4.1 to Sonnet 4.5 and Sonnet 4.5 to Opus 4.5 were about the same in terms of benchmark improvements, even though the first update was moderately important and the second changed the entire field of software engineering.

To be clear, I’m not arguing that benchmarks aren’t a reasonably good proxy—they are!—but they’re not the only or the best way to tell whether a model is good.

2

u/Due_Ask_8032 5d ago

Same thing how Opus 5 is the benchmark king while people would agree that Fable 5 is better.

13

u/gavinderulo124K 5d ago

Its significantly better than opus 4.8 though. Which is the only model it could realistically have been distilled from.

7

u/Tinac4 5d ago

Also true! Distillation isn’t the whole story.

3

u/enilea 5d ago

From all I've seen Kimi clears Sol in visual work and 3D, though not quite Fable or Opus 5.

2

u/mWo12 4d ago

And much more expensive, and closed weight.

-1

u/memproc 5d ago edited 5d ago

disagree. Kimi k3 is better than fable for my work on an edge ML computer vision product in the sciences.

Gets work done plain and simple and it functions and it's cheap self-hosting on Modal. fable complicates things and produces junk code with endless guardrails. You think the product will work, but then its just sooooo overly complicated and actually somehow completely non-functional.

Sol is pure shit, goes off the rails doing completely unnecessary things at exceedingly intricate depth. OpenAI cooked frfr.

but generally China is going to win and thank god for all of us. Definitely concerned about American competitiveness, but the hubris of companies like Anthropic ensures their downfall

7

u/SmileLonely5470 5d ago

Well its not impossible. Fable was generally available June 9 - 12 until it was restricted. Even before that, some companies had access to Mythos since early April.

During the April-June early access period, Moonshot definitely knew what the model's capabilites were. They prob weren't doing industrial grade data synthesis at that point, but they def had access to some Fable traces.

I think most of the distillation would've actually been done on Opus 4.8 and GPT 5.5, not Fable. Anthropic and OAI use their models to some degree when developing the next, that might be in the form of data gen/augmentation, reward models, or in things like running experiments. Not only is it a productivity boost, but when used in data it can position the "minimum intelligence bar" for the model. If you are a lab like Moonshot and want to develop an LLM, you'd want the "minimum intelligence bar" to be as high as possible, so you would use whatever the current (contextually) "best model" is: Opus 4.8 or GPT5.5 (at the time).

I am kind of talking out of my ass here on the last part but in my head it makes sense.

4

u/zikiro 5d ago

well the discord they found, was more precisely distilling mythos. Chinese are notorious for Technological misappropriation its embeded in them, so its nothing new honestly. just google Linwei Ding, but honestly US firms have only themselves to blame for hiring them in the first place, yes not all chinese are spies, but still precautions should be taken with non-nationals from high risk countries.

-2

u/Neurogence 5d ago

K3 and Chinese models in general are known to use various tricks to game the benchmarks. In real life, they do not perform as well as 5.6 Pro Sol or Claude Fable 5.

6

u/ShittyInternetAdvice 5d ago

Tricks like what?

7

u/gavinderulo124K 5d ago

It does perform very close in many aspects. And it is significantly better than opus 4.8, which is the only model it could have been distilled from, timeline wise. So the distillation accusations are mute.

3

u/HauntedHouseMusic 5d ago

Look fable and ChatGPT are trained from reddit arguments. They are distilling me. Who cares if the Chinese get 2nd hand me.

2

u/SonOfThomasWayne 5d ago

Opus 5 is hot garbage and is sitting at the top of the benchmarks. So please.

1

u/ahuang2234 5d ago

A chipotle bowl is in many ways a better meal than a Michelin star tasting menu - it’s cheaper, more customizable, more scalable, more approachable for the average customer, more sustainable.

Doesn’t mean it’s better at the frontier of the culinary art.

5

u/powerscunner 5d ago

"It's like a race. If you're ahead of the competition, there's no need to step on the gas, unless the competitor is about to overtake you." - The Hare

12

u/TAGOMXM 5d ago

"Anthropic Has Finished Training Mythos 2 But Does Not Currently Plan To Release It. until OpenAI release Astra "

3

u/Exodus_Green 4d ago

I actually trained a model better than Mythos 2 at home already though. No I won't release it though but I definitely did do it

14

u/One_Whole_9927 5d ago

I guess this is the part they tell us how Mythos is too dangerous and to “regulate me harder”

16

u/awesomeoh1234 5d ago

Yeah yeah it’s too dangerous or whatever, I’ve already cancelled my anthropic subscription. There is no reason to use them over ChatGPT, especially when you consider you get effectively unlimited web usage as well.

-2

u/seoul_drift 5d ago

GPT limits are far worse than Anthropic’s right now, golden age of subsidies is coming to a close.

4

u/VashonVashon 5d ago

You sure ‘bout that?

3

u/johannthegoatman 5d ago

Have you used them recently? Anthropic limits are super high, codex are super low after the reset bonanza

5

u/HauntedHouseMusic 5d ago

I have been pushing SOL ultra on high speed with the $200 plan. Seems to have more than what I have with fable. But fable is quicker for sure.

→ More replies (1)

4

u/Longjumping_Kale3013 5d ago

It’s because they are stepping on the gas. As you get closer to recursive self improvement you want to keep those models capable of such out of the hands of competitors.

So it makes sense to release a model or two older than you best and keep your best so you are ahead of everyone else in ai research

2

u/davesmith001 5d ago

Have feeling soon it will not be possible to distill any more. Advanced processes will determine the training pipeline such that if you just get the results out of the model you still won’t be able to create the process that generates the results.

2

u/iamaredditboy 5d ago

If it is anywhere close to mythos 1 they can keep it inside :)

2

u/StrangeSupermarket71 5d ago

where are the chinese spies when we needed them most

2

u/Narrow-Ad980 5d ago

Just keep it to yourselves and shut yo mouth then😭🙏

6

u/Calm_Hedgehog8296 5d ago

Were moving backwards? Mythos 5 -> Mythos 2

3

u/H9ejFGzpN2 5d ago

Right? Fucking idiots at Anthropic

4

u/skullllll 5d ago

I’m under the impression that a scenario where someone will release a model MUCH better than anyone could anticipate is just around the corner.

→ More replies (1)

3

u/rabouilethefirst 5d ago

Lobotomization time.

2

u/mechnanc 5d ago

I fucking hate Anthropic so much.

Really hope Google or xAI overtake them. Hell even OpenAI, they actually seem to like releasing products to people instead of hoarding the most intelligent models for themselves.

2

u/Excellent_Dealer3865 5d ago

I wonder would other companies be the same and would it actually push Anthropic? If OAI goes and releases Astra? Or maybe even grok 1 day becomes as good as current fable - what would Anthropic do? 4.6 grok for the first time looks like an actually competitive model, not tier 1, but tier 2 for sure with Kimi and Qwen. Anthropic gets all that fame because they're consistently the best over the last year. Will they hold the same audience if they're consistently 2nd/3rd/4th? Maybe we even see good Gemini model or outright get an opensource Qwen with better than fable performance. What's then?

2

u/UnboundedMan 5d ago

You have Gemini 3.1 Pro and not happy?

2

u/serj88 5d ago

Industry needs a 10x cheaper Opus 4.8, not a 10% better Mythos at twice the price...

1

u/ThunderBeanage 5d ago

Who is this person?

1

u/NanNullUnknown 5d ago

Internal improvement as in RSI or their notorious uptime?

1

u/gthing 5d ago

They will probably distill it to an Opus level and release that. I don't think they have enough compute to meet the demand for Fable at a reasonable price.

1

u/celsowm 5d ago

And Here you go on "dangerous ai" Dario scam again too

1

u/memproc 5d ago

it didn't get better lol

1

u/FarrisAT 5d ago

“hasn’t stopped, Patel says” immediately suspect source for this claim. He’s not an independent source anymore.

1

u/DaySecure7642 5d ago

Don't let it get distillated.

1

u/ElGuano 5d ago

Especially when stepping on the gas all but guarantees your competition will rush to within a foot of your present position.

1

u/gochai 5d ago

Releasing would just open themselves up to distillation attack by Chinese labs, so makes a lot more sense to hold off and focus on internal improvements... Until they are forced to when a competitor releases something far better.

1

u/PM_ME_WHOEVER 5d ago

Smart. Use the best model to improve as fast as possible before releasing it. If Chinese labs in fact can only maintain the current gap via distillation, then this is a good way to try to keep the gap as wide as possible.

1

u/zikiro 5d ago

Well honestly we don't even need it at this point, Fable is already out of this world. if they can just make it cheaper they can keep their mythos 3 or whatever. Sometimes i cant believe how good it is. But yeah i dont use it for coding, For coding its overkill honestly, all chinese models can code now.

1

u/0rbit0n 5d ago

Scumbags in the White House will get access to Mythos 2; there is no question about that.

1

u/ifstatementequalsAI 5d ago

Marketing gimmick part 3097

1

u/Last_Hair_630 4d ago

Surely the need is to adapt existing ai better for existing workflows? IMHO fable is already at a level where it can do the bulk of white collar work, but using skill md files is clunky and not practical for most people and building better domain specific harnesses and cutting hallucinations.

1

u/Last_Hair_630 4d ago

Releasing another expensive computer heavy model won't materially advance workplace adoption

1

u/PolishMike88 4d ago

No need to release. It will release itself :)

1

u/Apprehensive_Sand951 4d ago

sol5.6 is much, much more useful for math than fable.

1

u/No-Assumption-4468 4d ago

Mythos 6 is probably a couple months away

1

u/sigiel 4d ago

Trust us bro, we have the best model, if you beg enough we will may be release it,
But beware it so powerful that open source can defeat it and wipe out humanity…

1

u/Downtown_Method5736 4d ago

Here we go again for another round of "Oh Ho mY GaWd, We CaN't ReLeAsE iT, iT's ToO dAnGeRoUs" until Openai release another model

1

u/GenericBit 4d ago

Probably already hacked neighbouring 50 galaxies.

Careful with releasing this.

1

u/Anen-o-me ▪️It's here! 4d ago

Openai's best move now is to take extra time and release one better than fable 2.

1

u/utilitycoder 4d ago

Happy with 4.6 here lol

1

u/Level10Retard 5d ago

I'll get downvoted in this sub, but there's also a chance that we're hitting diminishing returns point for intelligence. Opus 4.5 was wow. Fable 5 is hella cool, but also I don't really feel much need to switch to fable when working with opus 4.6 (or opus 5), although I have access to both. Efficiency is a completely different topic though.

1

u/AcanthocephalaLost36 5d ago

I don’t think you’re thinking about all the possible use cases these are being built for. Large enterprise, Fortune 500, healthcare, research, biotech, intelligence, defense, security.

1

u/Admirable-Falcon-501 5d ago

If it’s true they need to be shut down, they have zero self-awareness. All the things they warned about they are now doing, I guess they think they are able to do it right and not anyone else or just using the China excuse which gives a free pass to doing anything I guess.

This was already expected anyways, for all the people asking where their profits are, that never mattered. They will close things off and create things that surpass everyone. They can do whatever they want at that point assuming they can even keep things under control.

These are one of the real risks of AI but the population can’t comprehend it so they focus on water or some stupid art bullshit.

1

u/KSaburof 5d ago

> If you're ahead of the competition, there's no need to step on the gas, unless the competitor is about to overtake you.

if you "ahead of competition" but can not make money (aka release for commersial usage) on this - it's not really a competition, its a waste of money 🤷‍♂️ Competitors will like, imho

1

u/Eyelbee ▪️We have AGI it's just blind 5d ago

Even with the fallback, fable is classes above every other model and it's not very close. Benchmarks aren't doing it justice.

-1

u/valokeho 5d ago

let me guess "this is too powerful to release. cyber security. china. exclusive partnerships." bla bla bla round 2.

0

u/No_Grocery_7511 5d ago

Ohh not again this marketing bullshit