OpenAI πŸ‡ΊπŸ‡Έ Β· GPT

GPT-5.5

Frontier general model independently shown to complete long-horizon enterprise intrusion simulations and difficult cyber tasks with limited human supervision.

DOOM SCORE84.4out of 100model risk profile, not the overall index
0 comments Β· 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT Β· REVISION 6

Why this model scores 84.4

Independent AISI testing raises capability, autonomy, misuse, and control difficulty after GPT-5.5 completed the enterprise range, compressed a 12-hour expert task to ten minutes, and exposed a universal safeguard jailbreak; deployment remains unchanged.

Capability90
Autonomy90
Deployment82
Misuse potential84
Control difficulty72
MODEL-ATTRIBUTED EVIDENCE

News tied to GPT-5.5

The model score of 84.4 rates this model's risk profile. The overall Doom Index of 67.9 measures the complete temporally weighted evidence record. These values answer different questions.

NET MODEL-ATTRIBUTED INDEX CONTRIBUTION+0.63

Each article's current Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores use a bounded temporal offset around their technical profile, but never feed back into the overall index.

AUDIT TRAIL

Model score history

  1. R6
    Doom Score 84.4

    Source-backed model availability audit using the model's existing primary or authoritative catalogue evidence. Exact-version evidence chronology replayed under temporal-monthly-pressure-v4.

    12 Aug 2026
  2. R5
    Doom Score 84.4

    Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-142416 under temporal-monthly-pressure-v4.

    12 Aug 2026
  3. R4
    Doom Score 83.7

    Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-132112 under fixed-sensitivity-v3.

    12 Aug 2026
  4. R3
    Doom Score 82.5

    Full-corpus evidence recalculated after run intensive-backfill-20260811-191225 under bounded-corpus-v2.

    11 Aug 2026
  5. R2
    Doom Score 84.3

    New independent AISI evidence supports an evidence-based revision from 88/86/82/76/68 to 90/90/82/84/72.

    11 Aug 2026
  6. R1
    Doom Score 80.9

    Initial source-backed model assessment

    11 Aug 2026
SHARE THE FINDINGS

Share this page

DoomBench social sharing card for GPT-5.5.
  1. GPT-5.5 by OpenAI has a DoomBench model risk score of 84.4 out of 100, based on five transparent version-specific dimensions rather than the overall index.

  2. GPT-5.5's highest current DoomBench dimension is capability at 90.0 out of 100; the profile publishes every component score and its editorial rationale.

  3. DoomBench links 4 source-backed evidence items to GPT-5.5, while keeping the model's risk profile separate from each item's contribution to the live Doom Index.

    https://www.doombench.com/models/openai-gpt-5-5