GLM-5.1
Flagship API model designed for sustained planning, execution, testing, repair, and long-horizon coding workflows.
0 comments · 0 votes
Sign in to join the discussion →
No comments yet. Start the discussion.
Why this model scores 78.3
Developer-reported eight-hour closed-loop execution suggests high autonomy, but capability and persistence lack independent validation; public API access still creates meaningful deployment.
News tied to GLM-5.1
The model score of 78.3 rates this model's risk profile. The overall Doom Index of 67.9 measures the complete temporally weighted evidence record. These values answer different questions.
Each article's current Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores use a bounded temporal offset around their technical profile, but never feed back into the overall index.
Z.ai releases GLM-5.1 for eight-hour autonomous task execution
Z.ai released GLM-5.1 through its API with long-horizon planning, tool use, iterative testing, repair, and claimed autonomous execution lasting up to eight hours.
- Full item contribution
- +0.13
- GLM-5.1 equal share
- +0.13
Model score history
-
R4
Doom Score 78.3
Source-backed model availability audit using the model's existing primary or authoritative catalogue evidence. Exact-version evidence chronology replayed under temporal-monthly-pressure-v4.
12 Aug 2026 -
R3
Doom Score 78.3
Exact-version evidence chronology replayed after run doombench-hourly-news-20260812-132112 under fixed-sensitivity-v3.
12 Aug 2026 -
R2
Doom Score 78.1
Full-corpus evidence recalculated after run intensive-backfill-20260811-191225 under bounded-corpus-v2.
11 Aug 2026 -
R1
Doom Score 78.3
New exact version verified from dated official release notes and documentation.
11 Aug 2026