News
AI Summary
5 Oct 202624 Rabiʻ II 1448 AH
Dalian University of Technology's VA-Bench: Top Multimodal Models Finish Only Half of Robot Tasks

Dalian University of Technology's VA-Bench: Top Multimodal Models Finish Only Half of Robot Tasks

A research team from Dalian University of Technology has launched VA-Bench, a benchmark that assesses the ability of multimodal large language models to translate visual input into successful robotic arm actions. The best-performing model, Alibaba's Qwen3.8-max, achieved an average task completion rate of 53.93%, indicating that models capable of 'understanding' a scene still struggle with reliable execution.

Follow these topics

Sign in to follow the topics that matter to you

Sign in to follow

This summary is generated with AI and receives periodic editorial review. Refer to the original source for full details.

0
0 reading now

Insight Score

Rate to unlock

Sign in to react, rate, and save. Sign In