GLM-4.6V open weights connect multimodal perception to tool use
Z.ai released GLM-4.6V and GLM-4.6V-Flash with MIT weights, 128K multimodal context, and native function calling that turns images, documents, and interfaces into agent actions.
0 comments · 0 votes
Sign in to join the discussion →
No comments yet. Start the discussion.
Why it moved the index
Open multimodal models that natively translate untrusted visual inputs into tool actions directly expand agentic diffusion and prompt-injection exposure.
Assessment history
-
R1
Toward 39 · confidence 84
New historical record from Z.ai's dated launch, repository, and model cards.
11 Aug 2026
Share this page
-
DoomBench assesses “GLM-4.6V open weights connect multimodal perception to tool use” as evidence moving toward doom, with magnitude 39 and confidence 84 out of 100 in the competitive race category.
-
The DoomBench assessment of “GLM-4.6V open weights connect multimodal perception to tool use” is based on reporting from Z.ai and records the editorial rationale, source quality, attribution, and revision history.
-
DoomBench summarizes “GLM-4.6V open weights connect multimodal perception to tool use” as follows: Z.ai released GLM-4.6V and GLM-4.6V-Flash with MIT weights, 128K multimodal context, and native function calling that turns images,...
https://www.doombench.com/news/glm-4-6v-open-weights-connect-multimodal-perception-to-tool-use-2025-12-08