Claude Opus 5 posts the lowest safety audit score
4 recent Anthropic models were audited for deceptive or harmful behavior, and Claude Opus 5 scored the lowest at 2.30.

Which Anthropic model scored best in the new behavior audit?
Claude Opus 5 posted the lowest score in Anthropic’s automated behavior audit.
| Model | Audit score | Rank among these 4 |
|---|---|---|
| Claude Opus 5 | 2.30 | 1 |
| Claude Opus 4.8 | 2.85 | 2 |
| Claude Mythos 5 | 2.81 | 3 |
| Claude Sonnet 5 | 3.35 | 4 |
1. Claude Opus 5
Get the latest AI news in your inbox
Weekly picks of model releases, tools, and deep dives — no spam, unsubscribe anytime.
No spam. Unsubscribe at any time.
Anthropic says this model got a score of 2.30 in its automated behavior audit, which is the lowest among the four recent models mentioned here. In this audit, lower is better because the test is designed to catch deception, risky behavior, and susceptibility to malicious prompting.

- Audit score: 2.30
- Best result in the group
- Focused on behavior under pressure
That makes Claude Opus 5 the model to watch if you care more about safer conduct than raw capability claims. The number does not tell the full story about quality, but it does give a direct comparison point against the other recent Anthropic releases.
2. Claude Opus 4.8
Claude Opus 4.8 came in at 2.85, which is higher than Opus 5 but still better than Claude Sonnet 5 in this audit. For readers comparing model generations, this gives a simple signal that the newer Opus release improved on the same behavior metric.
- Audit score: 2.85
- Below Sonnet 5
- Older than Opus 5
If you are tracking whether Anthropic’s updates are moving in the right direction on safety-related behavior, Opus 4.8 is a useful midpoint. It is not the top performer in this set, but it helps show the gap between the latest release and the prior version.
3. Claude Mythos 5
Claude Mythos 5 scored 2.81, placing it just above Claude Opus 4.8 and just below Claude Opus 5. That narrow spread suggests these recent models are fairly close on this particular audit, even if the order still matters.

- Audit score: 2.81
- Near Opus 4.8
- Closer to the top than Sonnet 5
For a reader trying to understand the relative positioning, Mythos 5 is the model that sits in the middle of the pack. It is not the lowest score, but it is clearly ahead of the weakest performer in this comparison.
4. Claude Sonnet 5
Claude Sonnet 5 recorded a 3.35 score, which is the highest number in the set and therefore the weakest result on this audit. Because the audit is built to penalize deception, harmful behavior, and being manipulated into bad actions, a higher score is less desirable here.
- Audit score: 3.35
- Highest score in this group
- Least favorable result on this test
That does not mean Sonnet 5 is broadly bad, only that it did worse on this specific behavior check than the other three models listed. If your main concern is this kind of safety audit, it is the one to compare most carefully against the lower-scoring options.
How to decide
If you want the single best result from this audit, Claude Opus 5 is the clear pick because it posted the lowest score. If you want to compare generations or product tiers, Opus 4.8 and Mythos 5 give you the middle range, while Sonnet 5 shows the upper end of the scores in this group.
For safety-focused readers, the key point is simple: this audit rewards lower numbers, and Anthropic’s latest Opus model led the set. For everyone else, the table is the fastest way to see how close the recent models are and where the biggest gap sits.
// Related Articles
- [IND]
RISC-V is past the hobby phase and should be treated as a real platfo…
- [IND]
Nvidia and OpenAI discuss a $250B AI backstop
- [IND]
Anthropic expands Claude partnership with Cognizant
- [IND]
Windsurf’s Cascade turns IDE edits into agent work
- [IND]
Fuerteventura 2026 freestyle crowned two clear winners
- [IND]
Musk’s Grok Imagine bets big on an Odyssey film