Autonomy and agency

Alibaba releases Qwen2-VL visual-agent models

Alibaba released open 2B and 7B Qwen2-VL models and a 72B API model with image, long-video, multilingual text, and visual-agent capabilities for operating mobile devices and robots from visual inputs.

0 comments · 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT · REVISION 1
TOWARD DOOM61confidence 98/100

Why it moved the index

The release paired strong visual reasoning with explicit device and robot operation, while open weights for two tiers reduced control boundaries and the hosted 72B tier broadened high-capability access.

AUDIT TRAIL

Assessment history

  1. R1
    Toward 61 · confidence 98

    New August 2024 exact visual-language and agentic model release.

    12 Aug 2026
SHARE THE FINDINGS

Share this page

DoomBench social sharing card for Alibaba releases Qwen2-VL visual-agent models.
  1. DoomBench assesses “Alibaba releases Qwen2-VL visual-agent models” as evidence moving toward doom, with magnitude 61 and confidence 98 out of 100 in the autonomy and agency category.

  2. The DoomBench assessment of “Alibaba releases Qwen2-VL visual-agent models” is based on reporting from Qwen Team and records the editorial rationale, source quality, attribution, and revision history.

  3. DoomBench summarizes “Alibaba releases Qwen2-VL visual-agent models” as follows: Alibaba released open 2B and 7B Qwen2-VL models and a 72B API model with image, long-video, multilingual text, and visual-agent capabilities for operating...

    https://www.doombench.com/news/alibaba-releases-qwen2-vl-visual-agent-models-2024-08-29