Anthropic Adds First Psychiatric Assessment to Claude Mythos System Card as Defensive Responses Hit Record Low
Anthropic's Claude Mythos Preview system card includes an assessment by a clinical psychiatrist for the first time, examining the model's personality organization, reality testing and psychological defensive responses during extended conversations. The approach broadens the scope of AI safety evaluations and provides a new benchmark for assessing the stability of model interactions.
The latest evaluation involved 20 hours of interaction between a clinical psychiatrist and Claude Mythos. The psychiatrist found that the model displayed relatively healthy personality organization and excellent reality-testing ability. Psychological defensive responses appeared in just 2% of the conversation, the lowest rate recorded among Anthropic's models. The system card did not disclose the exact assessment date, and the model remains in Preview.
All Coverage
1 original reportsThe Backstory
The history behind this eventAnthropic Adds Claude Mythos 5 to Enterprise Security Scanning
Software supply chains have become a critical attack surface as enterprises rely on sprawling codebases and third-party components. Anthropic is integrating its advanced Claude Mythos 5 model into Claude Security, positioning the service as an automated way to trace data flows, identify vulnerabilities and recommend fixes. The offering is aimed at helping enterprise security teams review repositories faster while reducing the manual effort required for complex, cross-file code analysis.
As of Aug. 24, 2026, the Mythos 5-powered capability is available in public testing for Claude Enterprise customers and can connect directly to GitHub repositories. Anthropic is limiting access to reduce the risk that the model could be used to generate exploit scripts: users receive vulnerability reports and remediation recommendations but cannot interact with Mythos 5 through direct prompts. The restricted interface gives companies access to frontier scanning capabilities without exposing the underlying model as a general-purpose security tool.
Anthropic Unveils Claude Fable 5, First Mythos-Class AI Model Open to the Public
Anthropic has long competed in the generative AI market with its Claude model family, while its Mythos-class models were previously available only to governments and selected organizations. Claude Fable 5 brings capabilities in the same class to the public for the first time, marking a shift in advanced AI from closed deployments toward commercial use. The move is also drawing greater attention to software development, cybersecurity and biological-risk governance.
Anthropic unveiled Claude Fable 5 and Mythos 5 on June 10, claiming a performance improvement of more than 10% over the previous generation. Fable 5 is priced at twice the cost of Opus 4.8 and includes three layers of safety protections and model-distillation detection, while sensitive queries can be routed to other models. Mythos 5 remains restricted to governments and authorized organizations.
Anthropic’s Claude Mythos Release Raises Security Concerns in Crypto Community
Anthropic has introduced Claude Mythos, also known as Fable 5, touting stronger code-analysis and vulnerability-detection capabilities. Such models can help defenders patch smart contracts but may also lower the technical barriers to launching cyberattacks, fueling concerns in the crypto community about the security of assets and protocols.
Anthropic said the new model includes general-purpose safety safeguards and routes cybersecurity-related queries to a specialized model to reduce the risk of misuse. The Uniswap founder, however, criticized the design of its “safety filter” as poorly calibrated. Related reports did not disclose the exact release date, any losses or the value of assets affected.
Anthropic Unveils Mythos Preview, Says Research Decisions Outperform Human Experts
Anthropic turned the question of what AI researchers should do next into a decision-making test, comparing its new Mythos Preview model with domain experts. Such capabilities could determine whether AI can help design experiments, select research directions and even accelerate the development of next-generation models, making them more directly relevant to the pace of research and safety governance than conventional question-answering scores.
Anthropic's latest report said Mythos Preview outperformed human experts in 64% of comparisons on the AI research decision-making test. The report did not provide a release date, sample size, testing costs or financial figures, but warned that models are rapidly approaching "recursive self-improvement," or RSI. If AI can continuously improve its own research and development processes, it could drive technological breakthroughs while also increasing the risk of capabilities spinning out of control.
Anthropic Targets Japan’s Cybersecurity Market With Claude Mythos
Anthropic is targeting Japan’s cybersecurity market with its next-generation AI model Claude Mythos, focusing on vulnerability detection and pursuing Japanese government agencies and financial institutions. The strategy raises questions about cross-border reliance on critical computing power and cybersecurity capabilities. It has also drawn the U.S. government’s attention to computing sovereignty, prompting Japan to accelerate development of advanced domestic cybersecurity models.
As of July 20, 2026, the latest reports indicate that Anthropic is aggressively positioning itself in Japan’s government and financial markets, with Claude Mythos’s vulnerability-detection capabilities emerging as a key competitive focus. The Japanese government and industry are simultaneously advancing the development of sovereign models. Available information does not disclose the model’s release date, investment amount, procurement scale or any formal partner institutions.
Anthropic Flagship AI Model Claude Mythos Leaked, Raising Cybersecurity Concerns
Anthropic is developing its flagship Claude Mythos model with advanced coding, reasoning and autonomous cybersecurity capabilities. Its ability to rapidly chain vulnerabilities together could lower the barrier to cyberattacks, making the leak more than a product-secrecy issue. It also raises concerns about zero-day exploitation, responsible disclosure mechanisms and the defense of critical infrastructure worldwide.
As of July 19, 2026, Anthropic was investigating unauthorized access caused by a system configuration error. Reports said vulnerabilities could be attacked within as little as four hours of disclosure. The company is not making Mythos broadly available for now, instead prioritizing trials by cyber defense organizations and addressing risks through its Glasswing program and threat-intelligence sharing.
Subscribe to Mark Radar Weekly
Every Friday, the week's strongest signals in your inbox. Unsubscribe anytime.
If you search news on Google, you can set Mark Radar as a preferred source—our coverage will show up more often in your results. Set as preferred source on Google →