Anthropic πŸ‡ΊπŸ‡Έ Β· Claude Sonnet

Claude Sonnet 5

A widely available, lower-cost model able to plan, browse, use terminals, and autonomously finish longer multi-step tasks.

DOOM SCORE79.8out of 100model risk profile, not the overall index
0 comments Β· 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT Β· REVISION 6

Why this model scores 79.8

Broad availability and cost-efficient autonomy increase real-world exposure, while lower cyber capability and improved agentic safety moderate the score.

Capability89
Autonomy90
Deployment94
Misuse potential62
Control difficulty58
MODEL-ATTRIBUTED EVIDENCE

News tied to Claude Sonnet 5

The model score of 79.8 rates this model's risk profile. The overall Doom Index of 67.9 measures the complete temporally weighted evidence record. These values answer different questions.

NET MODEL-ATTRIBUTED INDEX CONTRIBUTION+0.29

Each article's current Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores use a bounded temporal offset around their technical profile, but never feed back into the overall index.

Safety AWAY

Anthropic shows automated researchers can mitigate ten alignment failures

In a controlled study, Claude autonomously developed post-training methods that improved all ten tested alignment-failure categories without measured capability loss, generalized to withheld evaluations and larger models, and closed 65% of an early Claude Opus 4.8 checkpoint's measured safety gap in 60 hours.

Full item contribution
-0.11
Claude Sonnet 5 equal share
-0.06
Read assessment β†’
Autonomy TOWARD

Anthropic gives paid Claude users autonomous browser actions

Anthropic made Claude in Chrome generally available on every paid plan and enabled automatic approval for actions its safety classifier judges consistent with the user's request. Claude can use existing logins to read, type, click, navigate, and fill forms. In a stronger prompt-injection evaluation, probes plus the classifier reduced successful attacks to zero for Sonnet 5, Opus 5, and Mythos 5 and 0.3 percent for Fable 5, with the remaining successes manually rated low severity.

Full item contribution
+0.14
Claude Sonnet 5 equal share
+0.03
Read assessment β†’
Safety AWAY

Anthropic reports deployed auto mode cuts serious unintended agent harm

Anthropic reported that Claude Code's deployed permission classifier reduced production-level unintended harm in reviewed sessions from 6.3% under manual approval to 2.4%. Separate dated production case studies document sustained use at Nuro, Gusto, and Garner Health, while third-party testing found no successful attacks against three current Claude models in 720 prompt-injection trials.

Full item contribution
-0.08
Claude Sonnet 5 equal share
-0.03
Read assessment β†’
AUDIT TRAIL

Model score history

  1. R6
    Doom Score 79.8

    Source-backed model availability audit using the model's existing primary or authoritative catalogue evidence. Exact-version evidence chronology replayed under temporal-monthly-pressure-v4.

    12 Aug 2026
  2. R5
    Doom Score 79.8

    Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-142416 under temporal-monthly-pressure-v4.

    12 Aug 2026
  3. R4
    Doom Score 79.4

    Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-132112 under fixed-sensitivity-v3.

    12 Aug 2026
  4. R3
    Doom Score 78.5

    Full-corpus evidence recalculated after run doombench-hourly-news-20260812-002108 under bounded-corpus-v2.

    12 Aug 2026
  5. R2
    Doom Score 81.6

    Full-corpus evidence recalculated after run intensive-backfill-20260811-191225 under bounded-corpus-v2.

    11 Aug 2026
  6. R1
    Doom Score 79.7

    Initial source-backed model assessment

    11 Aug 2026
SHARE THE FINDINGS

Share this page

DoomBench social sharing card for Claude Sonnet 5.
  1. Claude Sonnet 5 by Anthropic has a DoomBench model risk score of 79.8 out of 100, based on five transparent version-specific dimensions rather than the overall index.

  2. Claude Sonnet 5's highest current DoomBench dimension is deployment at 94.0 out of 100; the profile publishes every component score and its editorial rationale.

  3. DoomBench links 4 source-backed evidence items to Claude Sonnet 5, while keeping the model's risk profile separate from each item's contribution to the live Doom Index.

    https://www.doombench.com/models/anthropic-claude-sonnet-5