back to top
HomeTechAnthropic's Mythos Just Helped Find macOS vulnerability That Could Break Apple's Security...

Anthropic’s Mythos Just Helped Find macOS vulnerability That Could Break Apple’s Security Protections

- Advertisement -

Anthropic has been explicit about why Mythos isn’t public. The model is too good at finding security flaws repeatedly, in production systems that some of the best engineers in the world have been maintaining for years.

So instead of a public release, Anthropic built Project Glasswing. Around 40 organizations get controlled access. Anthropic committed $100 million in usage credits to support the effort. The list includes Apple, Google and Microsoft, companies that aren’t exactly short on security talent themselves.

One of those organizations is Calif, a Palo Alto cybersecurity firm. In April their researchers used techniques derived from Mythos to find two previously undocumented vulnerabilities in macOS. They chained them together into a privilege escalation exploit capable of bypassing Apple’s memory integrity enforcement, the part of the system that’s supposed to be completely off-limits to normal processes. Then they flew to Cupertino and handed Apple a 55-page report in person.

Apple is reviewing it. Patches are expected. And Mythos just added macOS to a list that already includes a 27-year-old OpenBSD bug and multiple Linux vulnerabilities nobody had caught before.

Not Mythos alone

Calif CEO Thai Dong was direct with the Wall Street Journal: the attack “couldn’t have been pulled off by Mythos alone and leveraged the very human cybersecurity expertise of some of Calif’s hackers.”

That distinction matters in both directions. It’s not a story about AI replacing security researchers, the exploit required serious human expertise layered on top of what the model produced. But it’s also not nothing. Mythos narrowed the search space, surfaced the vulnerabilities, and gave researchers a starting point that would have taken significantly longer to reach on their own. The combination found something Apple missed.

An unavoidable track record

Before Calif walked into Apple’s headquarters, Mythos had already surfaced a vulnerability in OpenBSD that had gone undetected for 27 years. It found exploitable weaknesses in Linux that human researchers had walked past for years without noticing.

That’s not a coincidence or a lucky find. That’s a pattern. And it’s exactly why Anthropic won’t release the model publicly, because it works consistently enough that putting it in the wrong hands is a genuinely serious consideration.

The $100 million in usage credits Anthropic committed to Project Glasswing starts to make more sense in that context. It’s an attempt to extract real defensive value from a capability that exists whether Anthropic monetizes it or not. Better to have Apple and Google finding their own flaws with it than to wonder who else might find those flaws first.

You May Like: OpenMythos: The Closest Thing to Claude Mythos You Can Run (And It’s Open Source)

The question this raises for the rest of the industry

Forty organizations have controlled access to Mythos. The rest of the security research industry doesn’t.

That gap is going to widen. If a model under strict access controls is already finding decades-old bugs in the most scrutinized codebases on the planet, the natural question is what happens when similar capabilities become more broadly available, whether through Anthropic eventually loosening access, a competitor releasing something comparable, or less scrupulous actors building toward the same destination through different means.

Calif’s CEO thinks Apple will patch these bugs quickly. That’s probably true. But the more durable question isn’t about these two specific vulnerabilities. It’s about what the security research industry looks like when this kind of capability stops being rare.

The 55-page report is sitting on a desk in Cupertino right now. Full technical details won’t be released until Apple ships patches. When they do, this will become one of the cleaner documented examples of what AI-assisted vulnerability research actually produces in practice, two real bugs in a real operating system that nobody found until a model that Anthropic won’t let the public touch helped look for them.

Don’t miss any Tech Story

Subscribe To Firethering NewsLetter

You Can Unsubscribe Anytime! Read more in our privacy policy

LEAVE A REPLY

Please enter your comment!
Please enter your name here

YOU MAY ALSO LIKE
someone build an ai generated watermark open source remover after claude watermark

Anthropic Added Invisible Watermarks to Claude. Someone Already Built a Tool to Remove Them.

0
It hasn't even been a week since Anthropic started putting invisible watermarks into Claude's text. Now there's an open-source tool built to remove them. The project, watermarks-remover, has already exploded on GitHub, passing 8.8K+ stars and nearly 900+ forks in a matter of days. The numbers are impressive. But they're not the most important part. What's more revealing is how the tool works, what kinds of AI signals it targets, and how quickly a community-built project appeared around a system designed to make AI-generated content easier to identify. Because this is the uncomfortable reality of building anything in software, companies can spend months designing a new system, but once that system reaches the public, someone can start looking for a way around it.
Claude Will Soon Leave a Hidden Mark on Everything It Writes

Claude Will Soon Leave a Hidden Mark on Everything It Writes

0
Anthropic is adding invisible watermarks to text generated by Claude, and unlike a visible label, the marking is designed to travel with the text when users copy and paste it elsewhere. The move comes as AI-generated content becomes harder to distinguish from human writing and as the European Union begins requiring AI companies to make generated or manipulated content machine identifiable. Anthropic says the marking happens at the model level, meaning it can follow Claude-generated text across different products and surfaces rather than being tied to a particular app. The company also says it may survive some editing. But it raises a question, can AI-generated text actually be made traceable once it leaves the model that created it? And Anthropic's approach suggests the answer may be more complicated than simply adding a hidden signature to every sentence.
Zuckerberg Wrote 14 Pages About Open AI. His Best AI Model Is Still Closed

Zuckerberg Wrote 14 Pages About Open AI. His Best AI Model Is Still Closed.

0
Mark Zuckerberg published a 14-page essay today about why open-source AI is the path forward for humanity. Distribute intelligence rather than centralize it. Put the power in everyone's hands. A new era of personal empowerment. On the same day, Meta released Muse Glimmer, an open-source version of its most powerful model, Muse Spark, that anyone can download, modify, and build on for free. But the interesting part is, Muse Spark itself stays closed. You still pay to access it. The open version is nearly identical, Meta says, but the model that actually competes at the frontier, the one Zuckerberg's essay is implicitly defending remains behind a paywall. That gap between the philosophy and the product decision is what makes today's announcement interesting.