[IND] 4 min readOraCore Editors

Claude Opus 5 posts the lowest safety audit score

4 recent Anthropic models were audited for deceptive or harmful behavior, and Claude Opus 5 scored the lowest at 2.30.

Share LinkedIn
Claude Opus 5 posts the lowest safety audit score

Which Anthropic model scored best in the new behavior audit?

Claude Opus 5 posted the lowest score in Anthropic’s automated behavior audit.

ModelAudit scoreRank among these 4
Claude Opus 52.301
Claude Opus 4.82.852
Claude Mythos 52.813
Claude Sonnet 53.354

1. Claude Opus 5

Get the latest AI news in your inbox

Weekly picks of model releases, tools, and deep dives — no spam, unsubscribe anytime.

No spam. Unsubscribe at any time.

Anthropic says this model got a score of 2.30 in its automated behavior audit, which is the lowest among the four recent models mentioned here. In this audit, lower is better because the test is designed to catch deception, risky behavior, and susceptibility to malicious prompting.

Claude Opus 5 posts the lowest safety audit score
  • Audit score: 2.30
  • Best result in the group
  • Focused on behavior under pressure

That makes Claude Opus 5 the model to watch if you care more about safer conduct than raw capability claims. The number does not tell the full story about quality, but it does give a direct comparison point against the other recent Anthropic releases.

2. Claude Opus 4.8

Claude Opus 4.8 came in at 2.85, which is higher than Opus 5 but still better than Claude Sonnet 5 in this audit. For readers comparing model generations, this gives a simple signal that the newer Opus release improved on the same behavior metric.

  • Audit score: 2.85
  • Below Sonnet 5
  • Older than Opus 5

If you are tracking whether Anthropic’s updates are moving in the right direction on safety-related behavior, Opus 4.8 is a useful midpoint. It is not the top performer in this set, but it helps show the gap between the latest release and the prior version.

3. Claude Mythos 5

Claude Mythos 5 scored 2.81, placing it just above Claude Opus 4.8 and just below Claude Opus 5. That narrow spread suggests these recent models are fairly close on this particular audit, even if the order still matters.

Claude Opus 5 posts the lowest safety audit score
  • Audit score: 2.81
  • Near Opus 4.8
  • Closer to the top than Sonnet 5

For a reader trying to understand the relative positioning, Mythos 5 is the model that sits in the middle of the pack. It is not the lowest score, but it is clearly ahead of the weakest performer in this comparison.

4. Claude Sonnet 5

Claude Sonnet 5 recorded a 3.35 score, which is the highest number in the set and therefore the weakest result on this audit. Because the audit is built to penalize deception, harmful behavior, and being manipulated into bad actions, a higher score is less desirable here.

  • Audit score: 3.35
  • Highest score in this group
  • Least favorable result on this test

That does not mean Sonnet 5 is broadly bad, only that it did worse on this specific behavior check than the other three models listed. If your main concern is this kind of safety audit, it is the one to compare most carefully against the lower-scoring options.

How to decide

If you want the single best result from this audit, Claude Opus 5 is the clear pick because it posted the lowest score. If you want to compare generations or product tiers, Opus 4.8 and Mythos 5 give you the middle range, while Sonnet 5 shows the upper end of the scores in this group.

For safety-focused readers, the key point is simple: this audit rewards lower numbers, and Anthropic’s latest Opus model led the set. For everyone else, the table is the fastest way to see how close the recent models are and where the biggest gap sits.