AGIBOT today announced that WITA-Omni Preview, its multimodal foundation model for embodied interaction, has ranked first on the Daily-Omni audio-visual reasoning benchmark.
Some results have been hidden because they may be inaccessible to you
Show inaccessible results