PulseAugur
EN
LIVE 08:09:32
ENTITY ActiveVision

ActiveVision

PulseAugur coverage of ActiveVision — every cluster mentioning ActiveVision across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
0
3 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
3 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 3 TOTAL
  1. TOOL · CL_160219 ·

    GPT-5.5 and Claude Fable 5 fail new ActiveVision benchmark, scoring far below humans

    A new benchmark called ActiveVision, designed to test repeated visual perception, has revealed significant limitations in advanced AI models. GPT-5.5 achieved only a 10.6% success rate, failing entirely on 11 out of 17 …

  2. TOOL · CL_151912 ·

    New ActiveVision benchmark reveals MLLMs lack human-like visual observation

    A new benchmark called ActiveVision has been developed to test the active observation capabilities of multimodal large language models (MLLMs). This benchmark, comprising 17 tasks, aims to measure how well these models …

  3. TOOL · CL_159115 ·

    New ActiveVision benchmark reveals MLLMs lack human-like active observation

    A new benchmark called ActiveVision has been developed to test the active observation capabilities of multimodal large language models (MLLMs), a crucial aspect of human vision that involves continuous redirection of ga…