Anthropic Discloses 'Model 2' Outperforming Claude Mythos 5 in Risk Report
Anthropic's August 2026 risk report has disclosed a new, unreleased AI model named "Model 2" that slightly outperforms their current strongest model, Claude Mythos 5, on internal benchmarks like "CoBench v2". Users speculate that Model 2 and its predecessor Model 1 are further iterations of the Claude Mythos Preview model. This disclosure provides a rare glimpse into Anthropic's internal model roadmap and safety evaluations, showing steady incremental progress beyond their highly restricted frontier models. It highlights how AI developers are managing risk assessments for next-generation models before public deployment. While Model 2 shows superior performance across multiple tasks on the "CoBench v2" benchmark compared to Claude Mythos 5, the performance gains are relatively modest and focus on improving general capabilities.
## BACKGROUND
Claude Mythos is Anthropic's most powerful series of large language models, with Claude Mythos 5 released in June 2026 as a restricted-access version for vetted partners due to its advanced capabilities in cybersecurity and biology. Unlike its counterpart Claude Fable 5, Mythos 5 has fewer safety classifiers, allowing it to perform deep security scans without triggering standard refusals.