According to Anthropic's internal behavioral audit, Opus 5 scored better on alignment measures than any prior model, including Opus 4.8, Sonnet 5, and Fable 5, showing the fewest instances of deceptive outputs and the greatest resistance to manipulation attempts. Anthropic said it deliberately withheld cyber-focused training from Opus 5 — a decision mirroring its approach with Opus 4.8 — yet the model's cybersecurity performance rose anyway, a byproduct of broader capability improvements. In Anthropic's OSS-Fuzz tests, Opus 5 found vulnerabilities 79.4% of the time — nearly level with Mythos 5's 80% — but when it came to turning those findings into working exploits, the model succeeded in just 4 of the challenges where Mythos 5 cleared 13.