China’s Startups Race to Dominate Humanoid Robot Market


AGIBOT’s WITA-Omni PreviewTops Daily-Omni Audio-VisualReasoning Benchmark

【SHANGHAI, CHINA, July 28, 2026】AGIBOT today announced that WITA-Omni Preview, itsmultimodal foundation model for embodiedinteraction, has ranked first on the Daily-Omniaudio-visual reasoning benchmark.

According to the latest published results, WITAOmni Preview achieved an average accuracy of85.21%, outperforming models including Qwen3.5-Omni-Plus, Gemini 3.1 Pro Preview and DoubaoSeed 2.0 Lite.

The model recorded the highest scores in audiovisual alignment, comparison, event sequencing,and the benchmark’s 30-second and 60-secondvideo subsets. It also tied for first place in inference,ranking first or joint first in six of the eight metricsreported on the leaderboard.

1785231502806809.jpg
 

Users who are viewing this thread

Back
Top