SpaceXAI πŸ‡ΊπŸ‡Έ Β· Grok-2

Grok-2

xAI's larger second-generation reasoning model, deployed in beta on X with tool use, visual understanding, and image-generation access for Premium users.

DOOM SCORE60.0out of 100model risk profile, not the overall index
0 comments Β· 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT Β· REVISION 4

Why this model scores 60.0

Strong reasoning and multimodal functions inside a mass social platform create substantial capability and misuse exposure; hosted access limits distribution but documented election-image guardrail failures raise control difficulty.

Capability72
Autonomy24
Deployment79
Misuse potential68
Control difficulty57
MODEL-ATTRIBUTED EVIDENCE

News tied to Grok-2

The model score of 60.0 rates this model's risk profile. The overall Doom Index of 67.9 measures the complete temporally weighted evidence record. These values answer different questions.

NET MODEL-ATTRIBUTED INDEX CONTRIBUTION+0.15

Each article's current Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores use a bounded temporal offset around their technical profile, but never feed back into the overall index.

Capability TOWARD

xAI deploys Grok-2 and Grok-2 mini on X

xAI released beta versions of Grok-2 and Grok-2 mini to X Premium users with stronger reasoning, tool use, visual understanding, and image generation through a Black Forest Labs model.

Full item contribution
+0.13
Grok-2 equal share
+0.06
Read assessment β†’
AUDIT TRAIL

Model score history

  1. R4
    Doom Score 60.0

    Source-backed model availability audit using the model's existing primary or authoritative catalogue evidence. Exact-version evidence chronology replayed under temporal-monthly-pressure-v4.

    12 Aug 2026
  2. R3
    Doom Score 60.0

    Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-142416 under temporal-monthly-pressure-v4.

    12 Aug 2026
  3. R2
    Doom Score 61.3

    Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-132112 under fixed-sensitivity-v3.

    12 Aug 2026
  4. R1
    Doom Score 64.1

    New exact Grok generation and tier verified by xAI's dated launch.

    12 Aug 2026
SHARE THE FINDINGS

Share this page

DoomBench social sharing card for Grok-2.
  1. Grok-2 by SpaceXAI has a DoomBench model risk score of 60.0 out of 100, based on five transparent version-specific dimensions rather than the overall index.

  2. Grok-2's highest current DoomBench dimension is deployment at 79.0 out of 100; the profile publishes every component score and its editorial rationale.

  3. DoomBench links 2 source-backed evidence items to Grok-2, while keeping the model's risk profile separate from each item's contribution to the live Doom Index.

    https://www.doombench.com/models/spacexai-grok-2