Anthropic's Claude Mythos AI model revealed in data leak, described as 'step change' in capabilities

What was leaked
A data leak from an unsecured, publicly-searchable data store revealed that Anthropic is developing and testing a new AI model called Claude Mythos. The leak included approximately 3,000 unpublished assets linked to Anthropic's blog, including what appeared to be a draft blog post announcing the new model.
Model details from the source
According to the leaked documents:
- The model is called "Claude Mythos" and is also referred to as "Capybara," which Anthropic describes as "a new name for a new tier of model: larger and more intelligent than our Opus models."
- Capybara represents a new tier above Opus in Anthropic's model hierarchy (which currently includes Opus as the largest/most capable, Sonnet as faster/cheaper, and Haiku as smallest/fastest).
- Compared to Claude Opus 4.6, Capybara gets "dramatically higher scores on tests of software coding, academic reasoning, and cybersecurity, among others."
- The draft blog post describes Claude Mythos as "by far the most powerful AI model we've ever developed."
- Anthropic considers this model "a step change and the most capable we've built to date."
Current status and rollout
The model is currently being trialed by "early access customers" as part of a cautious rollout strategy. According to the source material:
- The model is expensive to run and not yet ready for general release
- Anthropic is "being deliberate about how we release it" due to the strength of its capabilities
- The company is working with "a small group of early access customers to test the model"
- The leak also revealed details of a planned, invite-only CEO summit in Europe as part of Anthropic's drive to sell AI models to large corporate customers
Security implications
The draft blog post indicated that the company believes Claude Mythos "poses unprecedented cybersecurity risks." The leak itself resulted from what Anthropic described as "human error" in the configuration of its content management system, which made draft content publicly accessible.
Context for developers using AI coding agents
For developers who rely on AI coding assistants, this leak suggests significant improvements in coding capabilities may be coming from Anthropic. The specific mention of "dramatically higher scores on tests of software coding" indicates potential advancements that could affect tools and workflows that integrate with Claude's API.
📖 Read the full source: HN AI Agents
👀 See Also

Seven Ways to Avoid Losing Your Job to AI – Tyler Cowen's Practical Guide
Tyler Cowen outlines seven principles, including seeking messy jobs and being wary of remote work, to protect your career against AI competition.

Talkie: A 13B LLM Trained Exclusively on Pre-1931 Text, Using Claude as a Judge in RL Training
Researchers released Talkie, a 13B LLM trained only on text published before 1931 (no internet, no WWII data). Claude Sonnet 4.6 was used as the judge in its online DPO reinforcement learning pipeline, and Claude Opus 4.4 generated synthetic multi-turn conversations for fine-tuning. The model can write Python code from a few in-context examples despite zero modern code in training.

Google donates Agent Payments Protocol (AP2) to FIDO Alliance, releases v0.2 with 'Human Not Present' payments
Google is donating the Agent Payments Protocol (AP2) to the FIDO Alliance, and releasing v0.2 with support for autonomous 'Human Not Present' payments and a new Verifiable Intent standard co-developed with Mastercard.

Adaptive Inference Routing Proposal for AI Query Efficiency
A proposal submitted to Anthropic in April 2026 outlines a five-step system for routing queries to appropriate AI models based on complexity scoring, using simple signals like character count and sentence count before any model inference occurs.