Autonomy and agency

GPT-5.1-Codex-Max sustains agentic work across millions of tokens

OpenAI released GPT-5.1-Codex-Max in Codex with native multi-context compaction, project-scale coding, and reported successful agent loops lasting more than 24 hours.

0 comments · 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT · REVISION 1
TOWARD DOOM50confidence 87/100

Why it moved the index

A released frontier coding agent that can independently persist, recover from failures, and work for hours directly extends consequential autonomy in software and cyber-relevant environments.

AUDIT TRAIL

Assessment history

  1. R1
    Toward 50 · confidence 87

    New historical record from OpenAI's dated launch and system-card evidence.

    11 Aug 2026
SHARE THE FINDINGS

Share this page

DoomBench social sharing card for GPT-5.1-Codex-Max sustains agentic work across millions of tokens.
  1. DoomBench assesses “GPT-5.1-Codex-Max sustains agentic work across millions of tokens” as evidence moving toward doom, with magnitude 50 and confidence 87 out of 100 in the autonomy and agency category.

  2. The DoomBench assessment of “GPT-5.1-Codex-Max sustains agentic work across millions of tokens” is based on reporting from OpenAI and records the editorial rationale, source quality, attribution, and revision history.

  3. DoomBench summarizes “GPT-5.1-Codex-Max sustains agentic work across millions of tokens” as follows: OpenAI released GPT-5.1-Codex-Max in Codex with native multi-context compaction, project-scale coding, and reported successful agent...

    https://www.doombench.com/news/gpt-5-1-codex-max-sustains-agentic-work-across-millions-of-tokens-2025-11-19