Claude 3 Sonnet
Anthropic's balanced Claude 3 tier at launch, deployed through free chat, API, Amazon Bedrock and Google Cloud preview for large-scale enterprise workloads.
0 comments Β· 0 votes
Sign in to join the discussion β
No comments yet. Start the discussion.
Why this model scores 53.3
Strong reasoning and vision with lower cost, high endurance and broad managed-cloud availability made Sonnet highly deployable, while its lower frontier capability and controlled interfaces constrained misuse and control difficulty.
News tied to Claude 3 Sonnet
The model score of 53.3 rates this model's risk profile. The overall Doom Index of 67.9 measures the complete temporally weighted evidence record. These values answer different questions.
Each article's current Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores use a bounded temporal offset around their technical profile, but never feed back into the overall index.
Anthropic maps and steers safety-relevant features in Claude 3 Sonnet
Anthropic extracted millions of interpretable features from the deployed Claude 3 Sonnet and showed that activating safety-relevant features could causally steer behavior. The work provided a production-model audit method while documenting substantial coverage and interpretation limits.
- Full item contribution
- -0.08
- Claude 3 Sonnet equal share
- -0.08
Anthropic releases Claude 3 Opus and Sonnet
Anthropic released Claude 3 Opus and Sonnet through Claude, its generally available API and selected cloud platforms, adding stronger reasoning, vision, long-context processing and support for complex automated work.
- Full item contribution
- +0.17
- Claude 3 Sonnet equal share
- +0.08
Model score history
-
R5
Doom Score 53.3
Source-backed model availability audit using the model's existing primary or authoritative catalogue evidence. Exact-version evidence chronology replayed under temporal-monthly-pressure-v4.
12 Aug 2026 -
R4
Doom Score 53.3
Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-142416 under temporal-monthly-pressure-v4.
12 Aug 2026 -
R3
Doom Score 52.9
Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-132112 under fixed-sensitivity-v3.
12 Aug 2026 -
R2
Doom Score 52.1
Full-corpus evidence recalculated after run intensive-backfill-20260812-034452-may-2024-incremental under bounded-corpus-v2.
12 Aug 2026 -
R1
Doom Score 56.1
New exact Claude 3 Sonnet public model profile.
12 Aug 2026