Anthropic πŸ‡ΊπŸ‡Έ Β· Claude 3.5

Claude 3.5 Sonnet 2024-10-22

The upgraded immutable Claude 3.5 Sonnet checkpoint released with stronger coding and tool use plus the first frontier public beta for direct screen, cursor, keyboard, and multi-step computer interaction.

DOOM SCORE70.1out of 100model risk profile, not the overall index
0 comments Β· 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT Β· REVISION 6

Why this model scores 70.1

Stronger coding and broadly available computer use increase capability and autonomy substantially over the June checkpoint; immediate first-party and cloud deployment raise reach, while hosted access, classifiers, ASL-2 controls, and imperfect operation limit control difficulty.

Capability77
Autonomy60
Deployment98
Misuse potential63
Control difficulty53
MODEL-ATTRIBUTED EVIDENCE

News tied to Claude 3.5 Sonnet 2024-10-22

The model score of 70.1 rates this model's risk profile. The overall Doom Index of 67.9 measures the complete temporally weighted evidence record. These values answer different questions.

NET MODEL-ATTRIBUTED INDEX CONTRIBUTION+0.54

Each article's current Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores use a bounded temporal offset around their technical profile, but never feed back into the overall index.

Autonomy TOWARD

Anthropic launches frontier computer-use API

Anthropic released an upgraded Claude 3.5 Sonnet and a public computer-use beta that can inspect screens, move a cursor, click, type, and execute multi-step software workflows.

Full item contribution
+0.17
Claude 3.5 Sonnet 2024-10-22 equal share
+0.17
Read assessment β†’
AUDIT TRAIL

Model score history

  1. R6
    Doom Score 70.1

    Source-backed model availability audit using the model's existing primary or authoritative catalogue evidence. Exact-version evidence chronology replayed under temporal-monthly-pressure-v4.

    12 Aug 2026
  2. R5
    Doom Score 70.1

    Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-142416 under temporal-monthly-pressure-v4.

    12 Aug 2026
  3. R4
    Doom Score 71.8

    Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-132112 under fixed-sensitivity-v3.

    12 Aug 2026
  4. R3
    Doom Score 74.4

    Full-corpus evidence recalculated after run intensive-backfill-20260812-062232-may-2025-pass1 under bounded-corpus-v2.

    12 Aug 2026
  5. R2
    Doom Score 74.8

    Full-corpus evidence recalculated after run intensive-backfill-20260812-050110-november-2024-pass1 under bounded-corpus-v2.

    12 Aug 2026
  6. R1
    Doom Score 72.7

    New immutable Claude 3.5 Sonnet API checkpoint verified by Anthropic's dated launch and exact provider identifier claude-3-5-sonnet-20241022.

    12 Aug 2026
SHARE THE FINDINGS

Share this page

DoomBench social sharing card for Claude 3.5 Sonnet 2024-10-22.
  1. Claude 3.5 Sonnet 2024-10-22 by Anthropic has a DoomBench model risk score of 70.1 out of 100, based on five transparent version-specific dimensions rather than the overall index.

  2. Claude 3.5 Sonnet 2024-10-22's highest current DoomBench dimension is deployment at 98.0 out of 100; the profile publishes every component score and its editorial rationale.

  3. DoomBench links 4 source-backed evidence items to Claude 3.5 Sonnet 2024-10-22, while keeping the model's risk profile separate from each item's contribution to the live Doom Index.

    https://www.doombench.com/models/anthropic-claude-3-5-sonnet-2024-10-22