AI News Today
← All episodes
Episode 1 · May 19, 2026 · 12:26

NEW Claude Mythos LEAKS + GPT 5.6!

Claude Mythos Leaks: Regulators Briefed, Cloudflare 2-Hour Patches, Apple M5 Bypassed (What It Means for Business)The script claims new Claude Mythos leaks indicate a major capability jump, alongside Sam Altman’s post that ChatGPT has improved and rumors of a GPT 5.6 release. It cites the UK AI Security Institute’s updated assessment of Mythos as a notable leap, including Mythos solving the previously unsolved “Cooling Tower” cybersecurity evaluation in 3/10 attempts and doubling autonomous task length in months. Cloudflare reports Mythos performs at a senior researcher level, prompting some teams to adopt two-hour patch SLAs. A Tom’s Hardware story says Mythos Preview helped find a local privilege escalation bypassing Apple’s memory integrity enforcement on the M5 chip. The Financial Stability Board and other finance regulators are being briefed, and the narrator argues businesses should build Claude “agent operating system” workflows now rather than wait, promoting the AI Profit Board/Ballroom offering.00:00 Mythos Leaks Shockwave00:57 Sam Altman Teases Upgrade01:49 Mythos AISI Breakthrough02:52 Cloudflare Two Hour Patches03:37 Apple M5 Exploit Found04:44 Finance Regulators Briefed05:48 What It Means For Business06:44 Agent OS Pitch07:46 Three Beliefs Holding You Back10:11 Claude Agent OS In Practice10:56 Final Signals And Next Steps

Full transcript

New Claude Miphos leaks just dropped and what's coming out is going to change how you think about AI forever. Plus Sam Orman just posted the chat GPT has gotten so much better. His exact words GPT 5.6 is reportedly dropping this week and Anthropic is sitting on a model so powerful they won't even let the public use it yet. But here's what nobody's really talking about.

The new Miphos leaks this week showed that AI just hit a level that has global financial regulators being briefed. That has Cloudflare moving to two hour response windows. That has Apple's most advanced chip protection getting bypassed. And there's one thing I'm going to show you later in this video that tells you exactly what this means for your business and how to make sure you're on the right side of the shift before everyone else figures it out.

Because the businesses that understand what these leaks are actually signaling and act on it now are going to be way ahead. I'll show you exactly what to do in a minute. Stick with me to the end. Let's get into it.

So with Sam's tweet this morning he posted he's really proud of the team for this one, right? No change log, no specifics, just the CEO telling the world that something has just shifted. His replies are split. Some people immediately notice it, other people saying the personality is flattened from the previous version.

Whether this is GPT 5.6 rolling out or something similar, we don't know yet. But that's where the 5.6 speculation is coming from and Sam's tweet this morning lines up with something that's happening. The rumor profile on 5.6 is consistent, better coding, finally stronger front end, more capable agents, cheaper to run. If those things are accurate this is a real jump not just a patch.

But here's the thing, whilst OpenAI is shipping updates and the rumor mill is running on 5.6, Anthropic is sitting on something they won't even release publicly. And the new information coming out about Mythos this week is the real story. Let's go through what actually leaked and has been reported in the last few days. So the UK's AI Security Institute, the government body that independently tests frontier AI models, just published an updated assessment of Mythos.

They called it a notable capability jump and this is notable because they already tested a preview version a month earlier and were impressed then. The version now being shared with companies like Apple and JP Morgan is a meaningful step beyond even that. Here's a specific thing they found. Mythos completed a cyber security evaluation called Calling Tower.

This test had never been solved by any AI model they'd ever tested. Mythos completed it in three out of ten attempts. First time in history any model has done that. And the AISI added something that should stop you in your tracks.

They said the length of tasks that frontier AI can now complete autonomously has doubled in a matter of months not years. They're now building entirely new harder tests just to keep pace with what these models can do. That's how fast this is moving. Then there's the Cloudflare report.

Cloudflare is one of the companies inside Project Glasswing, Anthropix initiative that gives select organizations access to Mythos to use it defensively. Cloudflare published their findings this week. They said Mythos is performing at the level of a senior researcher. It can look at a complex system, identify multiple small weaknesses, figure out how they connect to each other, and work out how they chain together into something serious all autonomously.

Because of this some teams inside Cloudflare have moved to two hour patch SLA windows. What does that mean? Well it means when Mythos finds something they've decided they need to fix it within two hours. That's how seriously they're taking the capability.

And the Apple story from Tom's Hardware is the one that hit the hardest this week. A security research team called Caliph used Mythos Preview, the version Anthropix has been sharing, to find a local privilege escalation exploit on Apple's M5 chip. Specifically it bypassed Apple's memory integrity enforcement. MIE is a hardware level security feature built into M5 and A19 chips.

It's Apple's most advanced memory protection. It operates at the hardware level. It has almost no performance at cost. It was designed to block the most common classes of attacks.

Mythos helped crack it. The result, run one command as a standard user, get full root access on the machine. Caliph published this as part of what they're calling the month of AI discovered bugs. That name alone tells you something.

This isn't a one-off. AI assisted security research is now producing fast enough to run a monthly series. And this is all coming from the preview version of Mythos, the one that's being shared externally. Not even the full model.

Meanwhile, the FSB, the Financial Stability Board, which is chaired by Bank of England Governor Andrew Bailey and includes senior officials from the US, UK, Australia, and China, is now being briefed directly by Anthropix. This is the body that monitors global financial stability. The IMF published a blog post this month saying cyber risk does not respect borders. As AI capabilities spread across countries, inconsistent oversight could weaken a globally interconnected system.

Goldman Sachs CEO David Solomon said he's hyper aware of Mythos. JP Morgan's Jamie Dimon said AI has made certain kinds of defense harder. Though he added it will ultimately help companies protect themselves too. The FCA's chief Nikhil Rafi confirmed the Bank of England is engaged on this and that UK and US authorities are already cooperating.

So this is not hype. This is the institutions that run global finance paying serious attention to one AI model. Now, what does this actually mean for you? Here's the thing that most people miss when they see news like this.

They either panic or they dismiss it. Neither is actually useful. What Mythos actually tells us is where the capability curve is going. And the direction is clear here.

AI models are getting faster at expert level work. The tasks that used to require a team of senior people are being compressed. The things that took days are taking hours. The things that took hours are taking minutes.

That curve doesn't just affect security research. It affects every knowledge-based business. Content, marketing, sales, agencies, client delivery, research. Every area where the output has historically been gated by human hours is being affected by the same curve.

And the businesses that set up proper AI systems right now, before this becomes a standard expectation, are going to be compounding whilst everyone else catches up. Now, if you want to get the most out of Claude right now, not just use it as a chatbot, but actually run your whole business with it, that's where the agent operating system comes in. It's the system that gives Claude a shared memory of your business, connects it to your actual workflows, and turns it from something you occasionally prompt into something that works in the background every day. Every time Anthropic ships a new capability, every time Claude gets stronger, which is happening fastest you've seen, your agent OS absorbs it and your whole system gets more powerful automatically.

That's a compounding effect. If you want to get that agentic AOS system, we've got it inside the iProfitBoarding, we've built it for you, we've given a little prompts, all the zip files to start improving it, and we walk you through the whole setup. So you get coaching calls every week, where we go deep on your specific Claude setup. For e-files members already running Claude agents in their business right now inside there, daily tutorials to full agent OS sips file.

And if you want Claude actually working for your business, instead of just answering questions, the link's in the comments description, or go to the AIProfitBoarding.com. Now let me address the beliefs that keep most people stuck when watching stuff instead of using it, right. So some people watch this stuff and not implement anything. That'll be 99% of people, but if you're watching this and you actually want to make it useful, here's what I'd recommend.

The first one is, you know, some people say this is all too advanced for me, right. I understand why it feels that way when you read about Miphos cracking Apple's M5 chip, or completing tests no AI's ever solved. It sounds like deep technical territory, but the same underlying capability AI doing complex expert level work autonomously is what makes Claude useful for your business right now. When Claude researches a prospect and drafts a personalized outreach email, it's using the same direction of capability.

When it processes a long brief and produces a structured content plan, same thing. The advanced version of this technology is being used by Cloudflare and Apple. The version available to you today is already powerful enough to save you hours every week. The gap between too technical for me and useful for my business is your setup.

That's it. And the second belief that I see people struggling with is, I should wait for Miphos to be publicly available before I invest time into this. Here's why that logic actually costs you. The businesses building Claude agent systems today will have months of compounding.

By the time Miphos drops publicly, they'll have memory systems trained on their business. They'll have workflows that are already tuned, prompts that are already optimized. When Miphos becomes available, the businesses plug it into their agent OS and instantly get a more powerful version of a system that already works. The businesses that waited start from zero.

So waiting for a better model is totally the wrong frame, right? Build the system now. Better models make the system better. They don't replace the need for a system.

And the third belief I see is, OpenAI is winning, so I should just use ChatGPT and not bother with Claude. And you know, for example, Sam Altman's tweet this morning is genuine, right? ChatGPT has improved a lot this year. But Anthropic locked a model because it's too capable to release.

The AISI called it a notable capability jump over the version they'd seen a month before. Cloudflare has run into our patch windows because of it. These are not the signals of a company that's fallen behind. They're the signals of a company that built something so far ahead, they're managing the rollout carefully.

Both Claude and ChatGPT are both capable of running serious agent systems right now. The question isn't which one is winning or which one to use. You know, the question is really, how are you going to implement both of them into your business? How can you use the power of both of them to get even more leverage?

Let me give you the practical picture of what a Claude agent OS system looks like in your business. You set it up once, Claude has a memory file that knows your business, your offer, your tone, your clients, your goals. You have workflows built around that memory. One workflow, for example, can monitor your inbox, identify new leads, research them and draft a personalized response ready for you to approve.

One workflow can take your raw ideas and turn them into structured content. One workflow could even track your pipeline and flag anything that's gone client in terms of follow-up. And none of these run because you remember to prompt Claude. They run because the agent operating system is set up to run them.

You check in, review, approve and move on. That's what AI working for your business actually looks like. Not a chatbot you open when you remember to. A system that runs whilst you do other things.

The mythos leaks this week are a signal, not a reason to panic, a signal that the capability curve is still steep. The pace is still accelerating and the window to get ahead of all of this is still open but not permanent. Cloudflare is at a two-hour patch window. Apple's most advanced chip protection just got bypassed.

The FSB is being briefed. The AISI is building new tests just to keep pace. All of this from a model that isn't even publicly available yet. The version available to you, Claude Sonic, Claude Opus, is already capable of doing very meaningful work in your business.

The direction this is going means that capability will only grow. The agent operating system is how you make sure you set up to benefit from every step of that growth and not scramble to catch up when the next version drops. The pace right now is unlike anything we've seen before. If you're still on the sidelines, the cost of waiting went up again.

Inside the AI Profitable Boarding we've got the full agent operating system for Claude, the zip file, every prompt, the memory system, the mission control dashboard. And every week we run four live coaching calls where we go deep on your specific Claude setup. How to connect your workflows, how to build the memory layer, how to get Claude actually generating leads and saving you time every day. There's 3,000 business owners in there running Claude agents right now.

Daily tutorials as new features drop, a 30-day roadmap so you know exactly what to build first, and a member map so you can connect with other Claude users near you who are already figuring this out. Link in the comments description or go to the AIProfitableBoarding.com. The leaks are out, the signals are clear, the only question is what are you going to build next?

More episodes

Browse all episodes →