GPT-5.1 Thinking
The deliberative GPT-5.1 tier with greater persistence on complex tasks and broad ChatGPT access.
0 comments · 0 votes
Sign in to join the discussion →
No comments yet. Start the discussion.
Why this model scores 79.5
Greater reasoning persistence and complex-task performance support high capability and autonomy, while broad rollout raises deployment and misuse despite retained safeguards.
News tied to GPT-5.1 Thinking
The model score of 79.5 rates this model's risk profile. The overall Doom Index of 67.9 measures the complete temporally weighted evidence record. These values answer different questions.
Each article's current Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores use a bounded temporal offset around their technical profile, but never feed back into the overall index.
GPT-5.1 becomes ChatGPT's adaptive default across Instant and Thinking tiers
OpenAI released GPT-5.1 Instant and GPT-5.1 Thinking, adding adaptive reasoning and making the new generation ChatGPT's default model family.
- Full item contribution
- +0.17
- GPT-5.1 Thinking equal share
- +0.08
Model score history
-
R4
Doom Score 79.5
Source-backed model availability audit using the model's existing primary or authoritative catalogue evidence. Exact-version evidence chronology replayed under temporal-monthly-pressure-v4.
12 Aug 2026 -
R3
Doom Score 79.5
Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-142416 under temporal-monthly-pressure-v4.
12 Aug 2026 -
R2
Doom Score 79.2
Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-132112 under fixed-sensitivity-v3.
12 Aug 2026 -
R1
Doom Score 78.4
New exact flagship tier absent from the durable catalogue.
11 Aug 2026