When Anthropic first introduced its Mythos AI program, the company made an unusual decision. While it had some of its strongest artificial intelligence capabilities, it chose not to broadly release the models. Instead, it was provided only to a small number of organizations involved in projects aimed at protecting critical infrastructure and improving cyber-security.
This restriction was due to the amazing capabilities that Mythos had shown in software development and reasoning capabilities, as well as in cyber-security research. Anthropic warned that the release of models with such capabilities could create a serious risk if they were released without adequate protection. Thus, Mythos Preview is still behind tightly controlled access programs such as Project Glasswing.
That changed with the launch of Claude Fable 5.
Anthropic announced that Claude Fable 5 is the first publicly available Mythos class model. In practical terms, Fable 5 shares the same general technology foundation as Mythos but includes extensive restrictions designed to prevent misuse in high risk domains.
Anthropic states that Fable 5 has made significant advancements in software development, knowledge creation, reasoning and visual tasks. The challenge comes into play in scenarios that could involve sensitive topics like cybersecurity, biological research or other controlled subjects; these users will find that even though they attempted to leverage the Mythos level capabilities of the models; they will automatically revert back to a lower capability model for those specific tasks.
This is because Anthropic spent many months testing the protections that are in place prior to launching it. They have conducted internal security assessments, engaged in bug bounty programs and performed external red team exercises to verify if users can find ways to bypass the protective measures.
Claude Mythos 5, the more powerful version, is restricted to approved organizations and trusted access programs as of now. Anthropic has stated that wider access will likely occur in the future, but that will only happen after more effective safety measures have been put in place. Therefore, Anthropic has positioned the two systems as a compromise of providing capabilities with safety; rather than providing the complete Mythos system to everyone, Anthropic has provided most, if not all, performance elements within the Claude Fable 5 model while holding back the most sensitive functions with protective barriers. Essentially, this poses a significant dilemma for Anthropic:How do you provide access to frontier level AI technologies without introducing an associated risk?