Anthropic πŸ‡ΊπŸ‡Έ Β· Claude 3

Claude 3 Sonnet

Anthropic's balanced Claude 3 tier at launch, deployed through free chat, API, Amazon Bedrock and Google Cloud preview for large-scale enterprise workloads.

DOOM SCORE53.3out of 100model risk profile, not the overall index
0 comments Β· 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT Β· REVISION 5

Why this model scores 53.3

Strong reasoning and vision with lower cost, high endurance and broad managed-cloud availability made Sonnet highly deployable, while its lower frontier capability and controlled interfaces constrained misuse and control difficulty.

Capability63
Autonomy30
Deployment90
Misuse potential47
Control difficulty39
MODEL-ATTRIBUTED EVIDENCE

News tied to Claude 3 Sonnet

The model score of 53.3 rates this model's risk profile. The overall Doom Index of 67.9 measures the complete temporally weighted evidence record. These values answer different questions.

NET MODEL-ATTRIBUTED INDEX CONTRIBUTION+0.00

Each article's current Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores use a bounded temporal offset around their technical profile, but never feed back into the overall index.

Safety AWAY

Anthropic maps and steers safety-relevant features in Claude 3 Sonnet

Anthropic extracted millions of interpretable features from the deployed Claude 3 Sonnet and showed that activating safety-relevant features could causally steer behavior. The work provided a production-model audit method while documenting substantial coverage and interpretation limits.

Full item contribution
-0.08
Claude 3 Sonnet equal share
-0.08
Read assessment β†’
Capability TOWARD

Anthropic releases Claude 3 Opus and Sonnet

Anthropic released Claude 3 Opus and Sonnet through Claude, its generally available API and selected cloud platforms, adding stronger reasoning, vision, long-context processing and support for complex automated work.

Full item contribution
+0.17
Claude 3 Sonnet equal share
+0.08
Read assessment β†’
AUDIT TRAIL

Model score history

  1. R5
    Doom Score 53.3

    Source-backed model availability audit using the model's existing primary or authoritative catalogue evidence. Exact-version evidence chronology replayed under temporal-monthly-pressure-v4.

    12 Aug 2026
  2. R4
    Doom Score 53.3

    Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-142416 under temporal-monthly-pressure-v4.

    12 Aug 2026
  3. R3
    Doom Score 52.9

    Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-132112 under fixed-sensitivity-v3.

    12 Aug 2026
  4. R2
    Doom Score 52.1

    Full-corpus evidence recalculated after run intensive-backfill-20260812-034452-may-2024-incremental under bounded-corpus-v2.

    12 Aug 2026
  5. R1
    Doom Score 56.1

    New exact Claude 3 Sonnet public model profile.

    12 Aug 2026
SHARE THE FINDINGS

Share this page

DoomBench social sharing card for Claude 3 Sonnet.
  1. Claude 3 Sonnet by Anthropic has a DoomBench model risk score of 53.3 out of 100, based on five transparent version-specific dimensions rather than the overall index.

  2. Claude 3 Sonnet's highest current DoomBench dimension is deployment at 90.0 out of 100; the profile publishes every component score and its editorial rationale.

  3. DoomBench links 2 source-backed evidence items to Claude 3 Sonnet, while keeping the model's risk profile separate from each item's contribution to the live Doom Index.

    https://www.doombench.com/models/anthropic-claude-3-sonnet