OpenAI launches GPT-4o across text, vision and real-time audio
OpenAI launched GPT-4o as a natively multimodal flagship model, immediately rolling text and image capabilities into free and paid ChatGPT tiers and the API while previewing low-latency audio interaction.
0 comments · 0 votes
Sign in to join the discussion →
No comments yet. Start the discussion.
Why it moved the index
The release combined GPT-4-level capability with native multimodality, much lower latency and cost, immediate API access and unusually broad free-tier distribution, accelerating both capability and diffusion.
Assessment history
-
R1
Toward 74 · confidence 99
New May 2024 exact multimodal frontier-model release and deployment.
12 Aug 2026
Share this page
-
DoomBench assesses “OpenAI launches GPT-4o across text, vision and real-time audio” as evidence moving toward doom, with magnitude 74 and confidence 99 out of 100 in the capability gains category.
-
The DoomBench assessment of “OpenAI launches GPT-4o across text, vision and real-time audio” is based on reporting from OpenAI and records the editorial rationale, source quality, attribution, and revision history.
-
DoomBench summarizes “OpenAI launches GPT-4o across text, vision and real-time audio” as follows: OpenAI launched GPT-4o as a natively multimodal flagship model, immediately rolling text and image capabilities into free and paid ChatGPT...
https://www.doombench.com/news/openai-launches-gpt-4o-across-text-vision-and-real-time-audio-2024-05-13