OpenAI WebGPT browses the web and outperforms human demonstrators
WebGPT autonomously issued searches, followed links, scrolled pages, gathered citations, and produced long-form answers. Its best 175B system was preferred over human demonstrations 56 percent of the time, and later motivated deployed ChatGPT browsing plugins.
0 comments · 0 votes
Sign in to join the discussion →
No comments yet. Start the discussion.
Why it moved the index
This was a general language model gaining consequential external-tool autonomy that directly influenced a later released browsing system.
Assessment history
-
R1
Toward 41 · confidence 93
New original research result with exact official RSS clock, exact 175B system identity, and separate primary practical-impact evidence from ChatGPT plugins.
12 Aug 2026
Share this page
-
DoomBench assesses “OpenAI WebGPT browses the web and outperforms human demonstrators” as evidence moving toward doom, with magnitude 41 and confidence 93 out of 100 in the autonomy and agency category.
-
The DoomBench assessment of “OpenAI WebGPT browses the web and outperforms human demonstrators” is based on reporting from OpenAI and records the editorial rationale, source quality, attribution, and revision history.
-
DoomBench summarizes “OpenAI WebGPT browses the web and outperforms human demonstrators” as follows: WebGPT autonomously issued searches, followed links, scrolled pages, gathered citations, and produced long-form answers. Its best 175B...
https://www.doombench.com/news/openai-webgpt-browses-the-web-and-outperforms-human-demonstrators-2021-12-16