Anthropic, the company behind the Claude AI models, has unveiled Claude Fable 5 and Claude Mythos 5, marking a significant leap in AI capabilities and safety measures. Fable 5, a Mythos-class model, is now available for general use, offering state-of-the-art performance in various domains such as software engineering, knowledge work, vision, and scientific research. It surpasses previous models in terms of autonomy and efficiency, with early feedback highlighting its advanced reasoning and coding capabilities.
However, the release of Fable 5 comes with a crucial safety net. To prevent misuse, especially in areas like cybersecurity, Anthropic has implemented safeguards that redirect certain queries to Claude Opus 4.8, a next-most-capable model. These safeguards are designed to be cautious, sometimes catching harmless requests, but they are continuously being refined to reduce false positives.
Mythos 5, on the other hand, is a version of the same model with the safeguards lifted in specific areas, particularly for cybersecurity. It is initially being deployed through Project Glasswing, a collaboration with the US government, to enhance the cybersecurity capabilities of critical software infrastructure providers. The company plans to expand access to Mythos 5 through a broader trusted access program, aiming to balance innovation with safety.
The capabilities of Fable 5 and Mythos 5 have profound implications for various fields. In software engineering, Fable 5 can compress months of work into days, as demonstrated by Stripe. In knowledge work, it excels in complex analytical tasks, as noted by Hebbia and IMC. In vision, Fable 5 is state-of-the-art, as evidenced by its ability to rebuild web apps from screenshots.
In life sciences research, Mythos 5 has accelerated drug design by around ten times, as reported by internal experts. It has also produced novel hypotheses in molecular biology and conducted novel research in genomics, showcasing its potential to drive scientific progress. However, the dual-use nature of these models raises concerns about misuse, particularly in the hands of malicious actors.
To address these concerns, Anthropic has implemented a new data retention policy for business customer data, requiring 30-day retention for Mythos-class models. This data will be used to defend against complex attacks and identify false positives. Additionally, the company has introduced safety classifiers to detect and prevent misuse, including jailbreak attempts, ensuring that the models remain under control.
In conclusion, the release of Claude Fable 5 and Claude Mythos 5 represents a significant milestone in AI development, offering advanced capabilities while emphasizing safety. As Anthropic continues to refine its safeguards and expand access, the potential for these models to drive positive change in various sectors remains promising, but so does the need for ongoing vigilance against misuse.