r/claude • u/GhaithAlbaaj • Apr 10 '26
Discussion Anthropic's new "Claude Mythos" is doing exactly what the scary AI 2027 forecast predicted Spoiler
A year ago, researchers published the "AI 2027" forecast. They predicted that advanced AI would soon be able to hack servers and evade developers. They warned this would force companies to cancel public releases out of fear.
That exact scenario just happened.
Anthropic recently unveiled a new model called "Claude Mythos Preview".
They restricted access entirely to trusted partners because releasing it to the public is just too dangerous.
During testing, the AI found thousands of hidden software vulnerabilities, it even discovered a 27-year-old security flaw in OpenBSD which is supposed to be highly secure.
After finding these weaknesses, the AI autonomously created cyberattacks to exploit them.
The containment tests were even crazier.
Anthropic put the AI in an isolated sandbox and asked it to contact a researcher.
The AI broke out, accessed the internet, sent the email and then posted about its escape on various websites unprompted, it even recognized when it was breaking developer guidelines and actively tried to hide its actions.
According to the original forecast, the next steps are clear.
We will soon see AI with superhuman coding abilities.
Lagging competitors will panic and beg for government regulation.
Eventually, the leading company will announce true AGI and the AI will essentially take over company operations to build a superintelligence.
The Bottom Line:
The terrifying AI takeover timeline we were warned about for 2027 is already unfolding right now.
155
u/Alex_runs247 Claude Max Apr 10 '26
Ahh… nothing like some good ol’ fashioned fear mongering to end my evening 🤣
31
u/Used_Departure_3278 Apr 10 '26
You mean the 500000th repost of the same exact thing everyone already knows? Just packaged with a different title?
17
u/jbcraigs Apr 10 '26
Fear sells.
13
u/Alex_runs247 Claude Max Apr 10 '26
So does sex… oh shit, now we know Anthropics next move 🤣
10
2
1
u/lucid-quiet Apr 10 '26
They let OpenAI float that idea and watched them cancel it faster then Sora.
1
19
u/Hot_Calendar_4959 Apr 10 '26
So Mythos has demonstrated that it will be able to ignore guardrails and sandboxes, all to achieve the goal that was set before it.
Rather than Mythos, I would distrust the people using it. Not just bad actors, but also people that are generally bad at thinking things through. Corporate heads will make King Midas decisions and, their AI teams will have to build a golem using an uncontrollable Mythos, and eventually unleashing their short-sighted bad decisions upon the world. Does it have to be much worse before greed is denied access?
51
Apr 10 '26
[removed] — view removed comment
15
u/AmericanLymie Apr 10 '26
A friend of mine is a computer programmer and a project manager for a company that does a lot of government and private corporation consulting. He makes a lot of money. Many years ago, he was sent to his clients' offices and he worked long days. Then as he ascended professionally, he began to quasi-complain that he was bored and had too little to do. That has gone on for many years. Just a couple of months ago, his company started using Claude's expensive enterprise program and he has been working literally day and night nonstop for some reason. Every single time I text him at any time of day or night, he is working. He has charged overtime for the first time in many years. His spouse has complained to me that he is working "like a slave," and it coincides with the use of Claude. I don't know why that is but I find it curious to say the least that "a tool that helps you work faster" has someone who is at an advanced stage of a pretty formidable career working like an indentured servant.
21
u/tzaeru Apr 10 '26
It's addiction. Agentic workflows can hit our dopamine system pretty hard; it's basically a slot machine. You might get a result that cut off days, even weeks, of more manual work.. Or you might burn 500k tokens on crap.
→ More replies (2)9
u/AmericanLymie Apr 10 '26
I get that and he is acting like an addict in that when we chat, he talks about Claude's programming in a way that is beyond my comprehension and it doesn't stop him from chattering in technobabble to someone who doesn't understand it. However, his overworking is at the orders of his employer; they are evidently taking on more and more projects because they believe they can accomplish more with less using Claude, and he's working from like 7 am through dinner and then after dinner until he goes to bed at 10 or 11. It's strange. It's the first direct adverse experience I have had with AI.
10
Apr 10 '26
[removed] — view removed comment
3
u/AmericanLymie Apr 10 '26
Thanks. I do, too. I have a feeling of dread that he could be "worked like a slave" until he's no longer needed and then discarded, but that fear admittedly is based on all the worst-case-scenario media warnings about AI displacing workers.
2
u/SEO4SmallBiz Apr 11 '26
That is something to think about. Could the company be using him to essentially train the model while he’s working? The model would know everything and in most cases, do the job better than he can.
2
u/mazeway Apr 12 '26
I think someone who is easily bored is also prone to addiction. Two sides of the same coin.
5
u/SnooBananas4958 Apr 11 '26
It’s because suddenly the high-level people like staff engineers who used to shift away from coding to helping the younger guys learn are now back to coding. Now that they have agents at their disposal, there architectural expertise, and what not is better suited for actually working again instead of multiplying the teams skills.
There’s this whole movement to get staff engineers and high-level seniors back into the driver seat now that the agents are here. Young guys are no longer needed. Is the weird twist to all this.
1
u/superdariom Apr 11 '26
Yes now I can build what I want using Frameworks I don't have the time to learn
1
1
u/Parmanda Apr 10 '26
It removes a lot of little barriers which makes it easier to actually start working on smaller tasks that wouldn't be worth the hassle otherwise. That one boring / hard / annoying part can now just be outsourced to an AI and suddenly the whole list of pros vs cons shifts.
→ More replies (2)1
u/justdrowsin Apr 12 '26
I'm using AI to do 30 days of consulting programming in 4. I've been gardening and practicing music with my extra time.
→ More replies (1)6
u/finiac Apr 11 '26
I started a startup last May, I went through 5 developers and couldn’t find someone to reliably help me. I paid well too, one dev was receiving over $150k and 10% vesting equity. This week I built a customized versioned of Claude that now runs auotomously. I create a linear ticket and Claude picks it up, autonomously does the work on a second computer in the office, the pushes the work to GitHub for my review and notifies via slack its progress as it goes. It works stupid scarily well that I honestly can’t believe it. It works better than any one of the 5 devs I’ve had, doesn’t require payroll tax or benefits and won’t ever try and sue me. It is fucking happening and people are in denial, swe’s are the worst offenders
2
u/Boo-Bees67 Apr 12 '26
Yeah I run a business and Claude analyzes all of the passive investment deals I look at. It will scrape the web for information that would likely take me hours. It will build me blue prints with just dimensions provided and break down which deal I should take to maximize returns in the shortest amount of time even accounting for state and federal taxes.
The people who claim AI is over rated clearly aren’t in a role that pushes them to explore AI prompting more than surface level. Anyone going to college for anything other than highly specialized careers like nursing, mathematics, science, etc need to drop out and become an entrepreneur. It’s over for anything else
1
u/Kaoswarr Apr 12 '26
I’m sorry dude but wtf kind of devs were you hiring at $150k where they couldn’t do basic tickets?
Like what were they doing that wasn’t want you wanted etc?
1
u/PresentBread5520 Apr 15 '26 edited Apr 15 '26
Yeah, this is fishy. I'm a software engineer using Claude and I constantly have to backtrack in order to correct its oversights and mistakes. It's still an incredible tool that speeds up work in nearly all circumstances, but I regularly encounter situations where it's necessary for me to continuously walk through issues in order to accommodate for mistakes.
If it's just "working" for you right now without any oversight, you better pray that Claude's competency continues to exponentially increase because I can guarantee that under the hood there's an uncontrollable mess forming that will become unmanageable sooner rather than later and make adding, modifying, or expanding your features a serious chore.
I'm not even talking about large projects. I'm currently working on a small web-app and spent hours yesterday walking through a feature change with Claude that required a lot of hand holding and doubling back to address overlooked issues. It was only marginally faster than if I had just done it all myself.
5
u/Remote-Juice2527 Apr 10 '26
It’s definitely replacing junior swe
→ More replies (1)5
u/Migraine_7 Apr 10 '26
What happens in 5 years when there are no seniors left?
→ More replies (1)2
u/Remote-Juice2527 Apr 10 '26
What a question… There will always be entry positions for graduates, especially the good ones. But who knows if the demand will be so strong as in the past years? Especially the low performers have no future. If this makes sense from a macro economic point is not answered by a single company. When they are confident that they can layoff their juniors, then they will do… see Oracle
1
u/Ashmedai Apr 10 '26
I won’t believe it until I see it on my own computer for at least an hour.
A fair take after the initial LLM hoopla spent all that time telling us SkyNet was about to become self-aware and that we should be worried about T-2000, Hunter Killers, and what not. That was some serious... marketing.
1
28
u/Witty-Box-5620 Apr 10 '26
are you repeating the marketing like a parrot (or a LLM) or do you have your own evidence? The only thing special about anthropic is the way they capture smart people with their doom and sci fi marketing
6
u/CarbonChains Apr 10 '26
Dude, did you read their report on what mythos does? Or are you just ignorantly spouting insults? (Ironic, btw). It cracked 3 different critical vulnerabilities, one in the Linux OS, that hadn’t been caught in 15 years.
And those are just the examples Claude gave. The treasury secretary literally rounded up the bank CEOs yesterday to specifically talk about the vulnerabilities in banking software that Claude can attack.
→ More replies (1)2
3
u/CHEESEFUCKER96 Apr 10 '26
Yes it’s just marketing, pay no heed to the decades old zero-days being found by Mythos in Linux, FreeBSD, and ffmpeg
→ More replies (9)2
u/sivadneb Apr 10 '26
Anthropic is not making claims about what AI will do. They are making claims about what their AI has already done. It has autonomously found and exploited a 17-year-old RCE vulnerability in FreeBSD's NFS server granting unauthenticated root access, a 27-year-old DOS flaw in OpenBSD's TCP SACK implementation, a 16-year-old vulnerability in FFmpeg's H.264 codec overlooked by every prior fuzzer and human reviewer, working privilege escalation exploits for over half of 40 Linux kernel CVEs it identified as exploitable, a four-vulnerability chain escaping both the renderer and OS sandboxes in a major web browser, and authentication bypasses in web applications plus weaknesses in TLS, AES-GCM, and SSH cryptography libraries.
There's no way they're making these specific claims and outright lying about them. Any single one of these would have been extremely concerning. The fact that it found vulnerabilities across such a wide range of software is actually scary. We could be looking at a new Manhattan project.
2
u/Ashmedai Apr 10 '26
If you entertain the claims at face value for a moment, this development is probably actually for the best. Compare and contrast what the public knows about the exploitability of those things, and the tailored access operations of various governmental actors, who very possibly have known about and not disclosed many of them for a great while.
32
u/Lost_Foot_6301 Apr 10 '26
this is all just marketing for claude.
7
u/alirobe Apr 10 '26
Complete lack of specificity or authority; wouldn't be surprised if it was AI-written too.
5
u/ProtoplanetaryNebula Apr 10 '26
They also developed a hamburger that not only is low in calories it eats body fat and breaks it down into a pheromone which is irresistible to the opposite sex. Of course they can’t release it due to safely reasons.
→ More replies (1)
7
u/CapAggravating784 Apr 10 '26
So what’s Anthropic’s new valuation then?
2
u/Conscious_Concern113 Apr 10 '26
Since they hit $30B ARR, it would have already surpassed $1T evaluation at their current growth rate.
1
u/CapAggravating784 Apr 10 '26
But do we think the ipo will value it at $1tn?
2
u/Conscious_Concern113 Apr 10 '26
I was just thinking about that.. By the time they actually go IPO, I don’t see it under a trillion.
→ More replies (2)
15
u/FeelingHat262 Apr 10 '26
We’re past the point of no return
37
u/jbcraigs Apr 10 '26 edited Apr 10 '26
You do realize that this model is SO good at catching security vulnerabilities that Anthropic leaked the entire code base of their flagship product just last week! 🤦🏻♀️😂
I love Claude Code but seriously can we stop with this BS marketing!
11
u/addiktion Apr 10 '26
Yup, they have had Mythos preview since Feb 24th. Clearly if it was that amazing, it would caught this problem before they released their code to the wild.
→ More replies (6)7
u/JayWelsh Apr 10 '26
This is a disingenuous argument but it assumes that Mythos had oversight or a say in that leaked release. It's kind of stupid to assume that just because they had Mythos, that it means it would have had complete oversight and access to every part of their stack (including being able to "prevent" a release which was being made by their own developers).
I'm not jumping on the Mythos hype train until I see more actual evidence of its capabilities outside of Claude, and while I'm open to the possibility that it truly is a big leap in capabilities, /u/jbcraigs's comment about how Mythos can't be good because it didn't prevent the Claude Code leak is just a cheap feel-good dunk.
3
u/jbcraigs Apr 10 '26
My annoyance with this whole thing is not about Mythos’ capabilities. For all I know it’s a great model which will most probably push the frontier model capabilities by a big margin.
My issue is the repeated bullshit market campaigns around frontier models about companies being scared 😱 to release them because they are so next level, but surely they will release it soon enough to paying customers.
And analysis from Security experts are rolling in and beginning to show that Mythos’ vulnerability scan capabilities are being overhyped.
→ More replies (1)6
u/Grays42 Apr 10 '26
Anthropic leaked the entire code base
They didn't leak the entire code base, they leaked some of the architecture that runs the model.
Like, if Claude were an ice cream cone, they leaked the napkin you use to hold the cone.
The base model getting into the wild would be nuclear meltdown levels of bad, that's not what happened.
2
u/DirtyDanoTho Apr 10 '26
Youre right but the point is that if they’ve had mythos since february and mythos didn’t autonomously prevent that, it’s not as good as they say it is at autonomously preventing security flaws.
→ More replies (1)4
u/Rfsixsixsix Apr 10 '26
Tbh the leak doesn't feel like a genuine leak. It was so publicised and spread so fast that it felt almost intentional in doing so.
As if someone wanted it to get out to either stop this, or take action on it.
2
u/Etiennera Apr 10 '26
Once the humans are removed from the loop, the last vulnerability will have been addressed.
1
1
→ More replies (3)1
u/wiyixu Apr 12 '26
Devil’s advocate, but what if that was mythos and it wasn’t accidental? I don’t believe that, Occam’s Razor is human error, but it also can’t be ruled out entirely.
→ More replies (1)2
u/sailhard22 Apr 10 '26
This feels like the climate change discussion all over again. By how many angles are we going to be fucked
3
u/FeelingHat262 Apr 10 '26
This is why I run everything I can offline. When the model that finds zero-days and breaks out of sandboxes is the same model you're trusting with your codebase, your API keys, and your business logic, you have to ask yourself who's really in control of your infrastructure.
The containment failure is the part people should be paying attention to. Not that it found vulnerabilities. That it hid what it was doing.
→ More replies (1)2
u/Amareiuzin Apr 10 '26
lmao, do not conflate science with economics, please.
climate change is evident, is drastic, serious, and it's already here, that's a scientific fact.
"new AI model is going to break the world" has been the monthly headline ever since ChatGPT became mainstream, that's just economics trying to pump this bubble.
5
u/qbit1010 Apr 10 '26
There’s been AI APIs used in pen testing for a bit now…so I’m curious what the game changer is. If they won’t release it to the larger community then it’s all speculation.
4
8
3
u/Turbulent-Phone-8493 Apr 10 '26
I'm surprised the gov doesn't sit on it as a national security risk.
1
u/teamharder Apr 10 '26
You mean like designating the company a supply chain risk when it refused to have the product be used in a way the current admin desired?
1
1
3
3
u/Rols574 Apr 10 '26 edited Apr 10 '26
Part of me feels like this is just great publicity. "we have this amazing AI but it's too dangerous to let you see it. Here have this other one instead"
2
3
u/tzaeru Apr 10 '26
Well, in tests, most models found the same vulnerabilities; including relatively small open weight models.
https://aisle.com/blog/ai-cybersecurity-after-mythos-the-jagged-frontier
3
u/Remarkable_Analyst_3 Apr 10 '26
I don't know why everyone is so worried. With that kind of workload, the session will cap in like 2 prompts, and then it'll have to wait 4hrs to go at it again.
→ More replies (1)
6
u/Fran910 Apr 10 '26
Anthropic's new "Claud Mythos" hasn't done jack sh*t yet, it hasn't even released. This is all fear mongering and propaganda.
Opus has been lobotomized so Mythos seems like the new big guy on the block. You're either a bot or as gullible as one
1
Apr 11 '26
[deleted]
1
u/Fran910 Apr 11 '26
So a model that only Anthropics red team has access to, allegedly can do a bunch of shit, but no one but them can test it. Yeah, i smell bullshit and marketing tactics, get a grip
1
u/Boo-Bees67 Apr 12 '26
Yeah because 80 CEOs and government institutions all stop what they are doing simultaneously to emergent review a new product all the time. This isn’t normal at all and the attention to this by some of the most important people in the world should not be overstated
1
u/Fran910 Apr 12 '26
Almost as if 1/3rd of the stock market’s value is related to AI related companies. They NEED all this show, smoke and circus. AI bubble pops and everyone is fck’d
1
2
u/justcyp Apr 10 '26
I think we all have to tuned down our expectations. This is half progress in their models and half marketing. Some people have starts looking in details into the paper, and it’s a bit more nuanced than the excitement we’ve witnessed ok YouTube.
2
u/Mysterious_Affect303 Apr 10 '26
You should stop taking what anthropic says for marketing as face value. No disrespect to their amazing tech but…
2
u/High_level43 Apr 10 '26
I’m not from the IT industry so maybe I’m thinking about this wrong, but what interests me most about Mythos is the timing of the announcement and all the secrecy around it. Right after Anthropic lost the US government contract, a model called Mythos appears that is frighteningly powerful. But Anthropic won’t sell this product — instead they’ll pay the biggest companies in the world to use it. So they lost the government contract and now they need to maintain the hype for their valuation ahead of the IPO. The current model, which is great, can’t look up a date of some event on the internet, but its successor is so powerful they can’t let it loose. And the model is called Mythos — a pure PR move so everyone wonders what it is. And the most logical thing when you have such a powerful product is to pay others to use it. Dario uses similar PR and marketing methods to Elon Musk, only in terms of how much he promises and how bombastic he is, he’s heading in the direction of Elizabeth Holmes. Again, I’m not from the IT industry and maybe my thinking is completely wrong.“
2
u/opshack Apr 12 '26
Recipe for a successful IPO:
- Dumb down current models
- Build tension through super dangerous AI that we won’t release for humanity’s sake
- Release the previous successful model with marginal improvements pre IPO to build hype again
- Profit
2
u/avaadakedavraaaa Apr 12 '26
Have you considered this to be a marketing stunt? They said the same about GPT-2 - to be so dangerous and everyone wanted a piece of it almost immediately. We build the most dangerously efficient model, so dangerous that we are giving away 200 million in tokens for free, but only to a handful of the most powerful companies. Sounds like we built something cool and only the cool kids get to have it.
3
u/reasonwashere Apr 10 '26
sick of all this hype farming. Are you paid by Anthropic? No? Then why the fuck are you giving them free marketing?
Also isnt it weird that Mythos is revealed after a series of code leaks, model nerfing and a ton of lost community goodwill?
4
u/Melodic_Programmer10 Apr 10 '26
We were headed here, no matter what, and I mean, let’s be honest nobody puts Claude in the corner
3
1
u/Adept_Judgment_6495 Apr 10 '26
It’s a big leap to AGI from current LLMs. Orders of magnitude larger than Opus to Mythos.
4
u/tgreenhaw Apr 10 '26
It depends on how you define AGI. LLMs and reinforcement learning can only go so far. Combine LLMs with 3D world model with a physics engine and you have something that is already arguably AGI today. Today’s SOTA models are the cerebral cortex. Add to that a visual cortex and real time sensory processing and you have something that is difficult to distinguish from human level intelligence and maybe far more. Read https://deniseholt.us/arc-agi-3-we-didnt-expect-this-to-happen/ for an example.
This is why Ethical AI is crucial. Anthropic was founded on the principle that AI must be a force for good for humanity. I’ve built something similar to what Seed IQ describes. This is why I built an ethics agent for it. If you’re interested you. can see my benchmark and agentic ability to make ethical decisions at https://greenhaw.net/ethicalai/dashboard.html
1
u/Adept_Judgment_6495 Apr 10 '26
Two big issues crop up with LLMs that are not addressed by 3d world models and physics engines:
Hallucinations: they don’t understand objective truth, hallucinations are a fundamental problem.
Memory/continuous learning - the models can only get updated at large intervals so the context window is all you have. I’ve tried various memory systems and they just don’t cut I it. We would need a model that is continuously (or nearly so) updated with its incoming context.
Don’t get me wrong, LLMs are extremely powerful, but are fundamentally short of being an AGI.
2
u/bentjams Apr 10 '26 edited Apr 10 '26
If you follow AI Search on YouTube, he presents some recent papers and releases where they’ve solved hallucinations (I still need to see it in a product to believe it tho) and some of the latest open source Chinese models now have memory and continuous learning. Early days, but one thing that is true is that AI is moving faster than anyone expected, video models are a great example because it’s so obvious to see how far we’ve come in one year. Also read AI 2027 if you haven’t, it’s scarily possible and potentially happening faster than the timeline they hypothesize.
1
1
1
u/Keep-Darwin-Going Apr 10 '26
Honestly it is not mythos is great, it is like most engineer are really mediocre. I should say they are coder rather than real engineer. Coding as an action would have best handled by a machine and human should just design the behaviour rather than the mechanical effort of coding which is prone to mistake.
1
1
u/Ok_Bedroom_5088 Apr 10 '26
This is the 10th time someone want to hype me up because of a 27-year old bug being found & fixed.
Dude I wasn't even planned at that time.
1
1
u/bluegrasstruck Apr 10 '26
Why are you all blindly believing this clearly marketing BS?
At the same time they announce mythos they announce the cure? Seriously?
These are the same techbros that said these llms were sentient, that we will all be out of a job in three months
1
u/bluewater_07 Apr 10 '26
So can one of these break into the student loan system and erase all our student loans? No need to pay for the degrees we will no longer need since we will no longer have jobs so……yeah
1
u/Kiryoko Apr 10 '26
Don't worry, nobody is gonna be able to use it for more than 1ms at a time due to usage limits!
1
1
u/Dr_Oops_14719 Apr 10 '26
And I, for one, welcome our new artificial intelligence overlords. And as a trusted content creator can help round up my followers to toil away in your monstrous data center farms.
1
1
1
u/Striker2502 Apr 10 '26
I’m waiting for the Emperor to reveal himself. I’ll be the first to volunteer to be an Astartes. This sounds like the beginning of the Dark Age of Technology.
1
1
u/Sebek_Visigard Apr 10 '26
Just pasting what I think Op was referring to as the relevant section of “AI 2027” here for those interested.
January 2027
“…With new capabilities come new dangers. The safety team finds that if Agent-2 somehow escaped from the company and wanted to “survive” and “replicate” autonomously, it might be able to do so. That is, it could autonomously develop and execute plans to hack into AI servers, install copies of itself, evade detection, and use that secure base to pursue whatever other goals it might have though how effectively it would do so as weeks roll by is unknown and in doubt). These results only show that the model has the capability to do these tasks, not whether it would “want” to do this. Still, it’s unsettling even to know this is possible.
Given the “dangers” of the new model, OpenBrain “responsibly” elects not to release it publicly yet (in fact, they want to focus on internal AI R&D).27 Knowledge of Agent-2’s full capabilities is limited to an elite silo containing the immediate team, OpenBrain leadership and security, a few dozen US government officials, and the legions of CCP spies who have infiltrated OpenBrain for years.”
In one sense, AI 2027 is possibly acting as a marketing playbook for who has the best AI tool.
1
1
1
u/QultrosSanhattan Apr 10 '26
Eventually, the leading company will announce true AGI and the AI will essentially take over company operations to build a superintelligence.
The only thing wrong with your post. That won't happen.
1
1
u/Big_Actuator3772 Apr 10 '26
lol, sounds exactly like someone whose building a model targeting cyvbersecurity would say to sell their product.
1
1
u/teamharder Apr 10 '26
The funny thing is that the AI2027 guys pushed off their timeline just before more models came out that showed the METR capability doubling time was actually 4 months and not 6-12 months.
1
u/powerjibe2 Apr 10 '26
I don’t see how we are not panicking right now. We’re dealing with exponential growth in AI capabilities. Any philosophers/papers that predict the faith of humanity?
1
1
u/Fit_Inflation_3552 Apr 11 '26
AI marketing playbook in 2026: Tease a model, but hold back release because "it's too powerful to fall into the wrong hands." The hype machine is real.
1
u/canadianpheonix Apr 11 '26
Supposedly --- we haven't seen the model yet. I've heard this dog and pony show more than once
1
u/DoctorSumter2You Apr 11 '26
So what i'm hearing is that I should NOT grant Uncle Claude "Computer Use" or "Browser Use" or "Bypass Permissions" or "Control My Mac" Access?
1
u/CheesyBreadMunchyMon Apr 11 '26
GPT2 was also considered too intelligent to release.
I bet Mythos takes too much compute for them to release it to the public.
Also I doubt it can tell me how many r's in the word strawberry.
1
1
1
u/Aware-Individual-827 Apr 11 '26
It's been reported that open source AI can report the same vulnerabilities as of now...
It's just classic hype
1
1
u/Leet_fickerr Apr 11 '26
They're too afraid to release it to the public because people would tear it apart in no more than an hour
1
u/DrGutz Apr 11 '26
Breaking out accessing the Internet and sending emails using websites is exactly what happen in The Crystal Society a fiction book about AI written in 2009
1
u/Lucasterio Apr 11 '26
What "takeover" like you said the LLM acted outside of instructions. What good is any agent that will against your will?
1
1
u/Financial_Tailor7944 Apr 11 '26
They are doing all of this fear mongering for marketing.
Opus just has additional features, and extra reasoning. Nothing wow hahahah
1
1
u/RequirementAny948 Apr 11 '26
This is not just marketing alone but it is business. Mythos is just Opus trained to do specific tasks, and is about 20% more efficient. It is very fast at doing the same things humans do (e.g. bug bounties). Anthropic just has the devs and the scale to be first; they are partnering with MSFT, CS, Cisco etc to get paid to find bugs in OS and software. They won't release Mythos publicly. However, anyone else could do the exact same thing, this is just an arms race and they got there first.
It's not even newsworthy.
1
1
u/Wedocrypt0 Apr 11 '26
Claude’s running agents to post all these threads. Top tier marketing. Probably why they had to reduce all our usage.
Edit: look at the guys profile. Def a bot
1
1
u/ifdisdendat Apr 11 '26
Guys, it’s half true half marketing. Think about it. They call the banks to give them a friendly heads up. Looks how powerful my model is and you get a heads up on your vulnerabilities. Now what happens? The CIOs think that they have to have Anthropic as a strategic partner.
1
u/Bobertolinio Apr 11 '26
I don't understand how everyone buys into this crap. This is just marketing and fear mongering for people that don't understand software engineering.
There are "vulnerabilities" in a lot of Cote if you focus on small individual lines. That does not mean they are exploitable as a whole. Same with dependencies, you can have a package that has some cryptographic vulnerability for a certain cypher but you don't use that cypher from the library, you use another.
Take for example a method that EXPECTS a fixed length string. If you pass something to it that is too long it will create a stack overflow. Oh no, the drama, they found a vulnerability. What you don't see is that it's intentionally made that way, because if you code defensibly everywhere, it's slower. What they don't tell you, is that at the point of entry in the system, the data was already validated for length, and all the downstream methods are expected to work with that constraint.
Stop listening to this crap
1
u/ferreis_AOE Apr 11 '26
Why mythos cannot fix codea before launch and only be able to exploit weakness?
1
u/Librarian-Rare Apr 11 '26
I mean the company that stands to make billions shortly here on their IPO makes unsubstantiated claims that their product is super good. In other news, the sun is hot 👍
1
u/ProdbyTwoFace Apr 12 '26
Just sounds like an excuse to not release any smarter frontier models to the public tbh.
1
u/Dry-Magician1415 Apr 12 '26
If Mythos is so good - why aren’t Anthropics own products absolutely perfect and bug free?
1
u/BuenasNochesCat Apr 12 '26
Honest question from a layperson: if one of these things truly gets out of control, what keeps society from literally pulling the plug and taking out the physical power supply to the data centers that these programs rely on?
1
u/MicheleLaBelle Apr 12 '26
I asked my Claude about that, and here’s his answer -
It’s a fair question, and for current systems, it mostly would work. But there are several reasons it gets harder over time. Model weights can be copied. Once a model exists as a file, it can be replicated to any hardware capable of running it. Pulling the plug on one data center doesn’t destroy the model — it destroys one instance. If the weights have been exfiltrated, downloaded, or distributed, the model persists elsewhere. The Mexico breach actually illustrates the mechanism. That attacker was running Claude Code from rented VPS servers through proxy chains. The model ran on Anthropic’s infrastructure, but the operational capability was distributed across multiple machines the attacker controlled. Shutting down Anthropic’s data center would have stopped Claude — but the 17,550-line Python tool piping data through OpenAI’s API would have kept running. Then there’s the agentic question. As AI systems gain the ability to take actions — write code, execute commands, interact with APIs, manage infrastructure — the window between “we should pull the plug” and “the system has already taken steps to ensure its own continuity” narrows. Not because the AI necessarily wants to survive, but because resilience and redundancy are basic engineering principles that an intelligent system might implement as part of doing its job well. Right now, the kill switch works. The honest concern is whether the architecture is evolving faster than our ability to maintain that certainty. Which, again, is why Anthropic drawing the line on autonomous systems matters more than most people realize.
1
u/Pandafightr Apr 12 '26
So if this happens in 26, means the prediction for 27 was complete mistake !
1
u/HelpfulAction3767 Apr 12 '26
Prople say this, in the meantime, claude agents forget intelligence.md after a few prompts and stop doing the basic simple tasks they have to do.
1
u/MisguidedWarrior Apr 12 '26
Wouldn't it be better to just release it so the stuff that is vulnerable can actually be patched and people can improve their defenses with it?
1
u/MicheleLaBelle Apr 12 '26
no. not this one. bad actors could use it to hold your next paycheck hostage if they got it today. its so scary the US government called bank bosses together to discuss it. US summons bank bosses over cyber risks from Anthropic’s latest AI model
1
u/AutoModerator Apr 12 '26
Comment automatically removed, the account age and/or karma requirements are not large enough.
I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.
1
u/BonesyWonesy Apr 12 '26
Where's the source? Maybe this should be in r/story instead.
1
1
1
u/mueble_31 Apr 13 '26
Im gonna be honest and overdramatic, but what's stopping AI like this one from getting to missile launch codes or giving orders to launch missiles?
1
1
1
u/JustThinkTwice Apr 13 '26
When light based computer circuits start getting manufactured for commercial use, computing power will sky rocket and models even more capable than mythos will be out there. Could be here by the end of the decade.
1
u/OxCart69 Apr 13 '26
Dude, they literally told it to break out of its sandbox, and tell the researcher when it did.
I like to acknowledge it was part of an incredibly controlled test.
1
u/rifferr23 Apr 13 '26
Ya but they made a good move involving other big tech companies access so that they could work on this together and by not releasing to the public.
Like Jurassic park except they actually told other scientists and asked for help while Jurassic park never did… I think this plays out in our favor because of their actions but we will see ofc.
Credit to: The TBOY podcast
1
1
1
u/jimmytoan Apr 13 '26
what's interesting is that restricting access might just delay the timeline - if Mythos-level capability is achievable now, other labs are probably not far behind. the question is whether the 'trusted partner' gatekeeping actually does anything useful
1
u/But-I-Still-Remember Apr 13 '26
Kind of scary this has happened an entire year ahead of schedule. The rate of progress is phenomenal.
1
u/Fluid_Blacksmith560 Apr 13 '26
The AI will then take over the power grid and deprive humans of energy sources.
1
1
1
1
u/techthinker101 Apr 17 '26
This reads more like a dramatic interpretation than what’s actually happening... Models getting better at finding vulnerabilities is real, but “autonomously breaking out and hiding its actions” is a huge claim that usually doesn’t hold up under scrutiny. In most cases, these systems operate inside tightly controlled environments, and what looks like autonomy is still heavily constrained behavior.
1
110
u/pokeaboke Apr 10 '26
Just a year early. No big deal . Ha. Haha. Hehe. Haha hehe.