PulseAugur
EN
LIVE 07:05:00
ENTITY Vision-Large-Language Models

Vision-Large-Language Models

PulseAugur coverage of Vision-Large-Language Models — every cluster mentioning Vision-Large-Language Models across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
0
1 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
0 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 4 TOTAL
  1. TOOL · CL_171272 ·

    Vision LLMs Revolutionize Web Automation with Sight-Driven Approach

    Web automation is undergoing a significant shift from brittle, hardcoded locators like XPath and CSS selectors to a more resilient, sight-driven approach. This new paradigm leverages multimodal Vision Large Language Mod…

  2. RESEARCH · CL_90999 ·

    New framework enables self-supervised concept discovery in VLLMs

    Researchers have developed S$^2$COPE, a novel framework for self-supervised concept discovery in vision-large language models (VLLMs). This label-free approach uses VLLMs as active participants in a preference optimizat…

  3. RESEARCH · CL_53932 ·

    New METATR benchmark evaluates multilingual ATR systems

    Researchers have introduced METATR (v1.0), a new multilingual benchmark designed to evaluate Automatic Text Recognition (ATR) systems, particularly Vision-Large Language Models (vLLMs). Unlike existing benchmarks that f…

  4. TOOL · CL_29293 ·

    New benchmark tests VLLMs on historical Chinese character evolution

    Researchers have introduced Chronicles-OCR, a new benchmark designed to test the cross-temporal perception abilities of Vision Large Language Models (VLLMs) on Chinese characters. This benchmark covers the complete evol…