PulseAugur
EN
LIVE 04:28:51

AgiBot WITA-Omni leads DailyOmni leaderboard, surpassing Google Gemini

AgiBot WITA-Omni, a new full-modal AI model, has achieved the top position on the DailyOmni Global Leaderboard for embodied cross-modal understanding. The model outperformed competitors such as Google Gemini, ByteDance Doubao, and Alibaba Qwen. AgiBot WITA-Omni utilizes a Thinker-Talker-Actor architecture, which synchronizes speech, action, and expression on a unified timeline, and secured first place in six out of eight performance indicators. AI

IMPACT Sets a new benchmark for embodied cross-modal understanding, potentially influencing future multimodal AI development.

RANK_REASON The cluster reports on a new AI model achieving a top score on a specific benchmark, surpassing competitors. [lever_c_demoted from research: ic=1 ai=1.0]

Read on Pandaily →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

AgiBot WITA-Omni leads DailyOmni leaderboard, surpassing Google Gemini

COVERAGE [1]

  1. Pandaily TIER_1 English(EN) · [email protected] (Pandaily) ·

    AgiBot WITA-Omni Full-Modal Model Tops DailyOmni Global Leaderboard: Beating Google Gemini, ByteDance Doubao, and Alibaba Qwen at Embodied Cross-Modal Understanding

    AgiBot WITA-Omni scores 85.21 on DailyOmni benchmark, 6 of 8 indicators first place, using Thinker-Talker-Actor architecture that synchronizes speech, action, and expression on a single timeline.