Capability gains

UK AISI finds autonomous cyber task horizons doubling every few months

UK AISI found that frontier models' 80-percent-reliability cyber task horizon had doubled every 4.7 months since late 2024, with GPT-5.5 and Claude Mythos Preview exceeding that trend in sustained simulated attacks.

0 comments · 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT · REVISION 1
TOWARD DOOM64confidence 94/100

Why it moved the index

The evaluation measures sustained autonomous cyber exploitation, and a later UK government action explicitly says this research informed policymaking, establishing practical security and governance impact rather than benchmark significance alone.

AUDIT TRAIL

Assessment history

  1. R1
    Toward 64 · confidence 94

    New practical-impact capability evidence from a dated government evaluation and a separately dated policy-impact source.

    11 Aug 2026
SHARE THE FINDINGS

Share this page

DoomBench social sharing card for UK AISI finds autonomous cyber task horizons doubling every few months.
  1. DoomBench assesses “UK AISI finds autonomous cyber task horizons doubling every few months” as evidence moving toward doom, with magnitude 64 and confidence 94 out of 100 in the capability gains category.

  2. The DoomBench assessment of “UK AISI finds autonomous cyber task horizons doubling every few months” is based on reporting from UK AI Security Institute and records the editorial rationale, source quality, attribution, and revision history.

  3. DoomBench summarizes “UK AISI finds autonomous cyber task horizons doubling every few months” as follows: UK AISI found that frontier models' 80-percent-reliability cyber task horizon had doubled every 4.7 months since late 2024, with...

    https://www.doombench.com/news/uk-aisi-finds-autonomous-cyber-task-horizons-doubling-every-few-months-2026-05-13