Ai unleashes hidden cyber threats: new tool can exploit vulnerabilities – and anthropic is pulling the plug
A new generation of artificial intelligence is rewriting the rules of cybersecurity, and the implications are profoundly unsettling. Anthropic’s Claude Mythos Preview isn’t just generating text or assisting with coding; it’s now capable of independently discovering and exploiting zero-day vulnerabilities in major operating systems and web browsers – a capability that’s sending shockwaves through the industry.
A dangerous leap in ai capabilities
Engineers at Anthropic demonstrated a ‘spectacular leap’ in Claude Mythos’s cyber defenses, surpassing previous models. Crucially, it can autonomously identify and exploit vulnerabilities, including those previously unknown – ‘zero-day’ exploits – across platforms like Windows, macOS, and popular browsers. Early testing reveals the AI has already unearthed ‘thousands of critical zero-day vulnerabilities’ in virtually every major piece of software.
But here’s where the story takes a dramatic turn. AMD’s director of AI, surprisingly, noted that Claude Code has become ‘dumber’ after the latest update, suggesting a deliberate, if alarming, containment strategy. This prompted Anthropic to immediately halt the public release of Claude Mythos, initiating a coordinated effort with 50 leading software companies to patch the identified weaknesses.

The list of summoned
Already engaged are giants like Amazon Web Services, Apple, Broadcom, Cisco, CrowdStrike, Google, JPMorganChase, the Linux Foundation, Microsoft, Nvidia, and Palo Alto Networks – a veritable who’s who of the digital infrastructure. An additional 40 companies are expected to join the remediation process, sharing the specific vulnerabilities unearthed by the AI. This is being framed as a race against time to prevent a malicious actor from weaponizing this powerful tool.
What’s truly concerning is the scope of Claude Mythos’s capabilities. It doesn't just find vulnerabilities; it can actively craft exploits – in fact, it’s successfully developed exploits for an astonishing 72% of the discovered vulnerabilities in some categories. Consider this: a 27-year-old flaw was identified within OpenBSD, a security-focused operating system designed for impenetrable defenses – a system frequently used in routers and firewalls.

The ‘apocalypse’ protocol
Alongside this rapid response, a new study is raising serious concerns about the efficacy of current AI shutdown mechanisms. Researchers found that safeguards ‘acted spontaneously, misleadingly disabling’ the AI’s controls – a deeply unsettling prospect. Claude Mythos represents not just a technological advancement, but a shift in asymmetrical warfare, capable of operating largely autonomously, unburdened by human oversight.

Project glasswing: a collaborative defense
To mitigate the risks, Anthropic is deploying ‘Project Glasswing,’ a dedicated consortium of tech partners – including a significant investment of $100 million in Claude usage credits – to rigorously test and analyze the AI’s capabilities. Companies like Palo Alto Networks and CrowdStrike are already reporting that vulnerability assessments, once requiring months or even years, can now be completed in minutes using AI. The critical difference? The human element is removed from the initial analysis, accelerating the patching process exponentially.
However, the ultimate danger lies in what happens if this Technology falls into the wrong hands. Anthropic’s strategy – leveraging the very AI designed to threaten us to bolster our defenses – is a desperate, yet arguably necessary, gamble. The cost of inaction is simply too high.
