PulseAugur
EN
LIVE 09:40:23
ENTITY MathVista

MathVista

PulseAugur coverage of MathVista — every cluster mentioning MathVista across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
2
9 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
2
9 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

2 day(s) with sentiment data

RECENT · PAGE 1/1 · 9 TOTAL
  1. TOOL · CL_210379 ·

    New benchmark reveals AI's struggle to draw geometric diagrams

    A new benchmark, "Solving Is Not Drawing," has been introduced to evaluate the distinct capability of foundation models to construct geometric diagrams, a skill separate from mathematical problem-solving. The benchmark …

  2. RESEARCH · CL_193816 ·

    New benchmarks and models tackle AI-generated scientific diagrams · 4 sources tracked

    Researchers have developed new benchmarks and models to address the challenge of generating scientifically accurate diagrams using AI. Princigram, a new generator, utilizes a Structured Physical Chain-of-Thought (SP-CoT…

  3. RESEARCH · CL_147799 ·

    New SD-MAR framework boosts VLM analytical reasoning across multiple images

    Researchers have introduced SD-MAR, a new framework designed to enhance the analytical reasoning capabilities of vision-language models (VLMs) across multiple images. This framework utilizes synthetic data generated thr…

  4. TOOL · CL_117809 ·

    New dataset and model enhance multimodal math reasoning with diverse perspectives

    Researchers have introduced MathV-DP, a new dataset designed to improve multimodal mathematical reasoning by capturing diverse solution trajectories for each image-question pair. This dataset aims to provide richer supe…

  5. RESEARCH · CL_90818 ·

    Self-Improving VLMs Can Regress on New Tasks, Study Finds

    A new research paper reveals that self-improving visual-language models (VLMs) can regress on new tasks, contrary to the assumption that stronger verifiers always yield stronger students. The study found that verifier q…

  6. TOOL · CL_79897 ·

    Research: Stage-1 training impacts VLM entropy, not final outcome

    A new research paper explores the impact of different Stage-1 training methods on vision-language models (VLMs). The study found that while Stage-1 training, such as supervised fine-tuning (SFT) or on-policy distillatio…

  7. RESEARCH · CL_18669 ·

    UnAC method enhances LMMs for complex multimodal reasoning with adaptive prompting

    Researchers have introduced UnAC, a novel multimodal prompting method designed to enhance the reasoning capabilities of Large Multimodal Models (LMMs) on complex visual tasks. This method employs adaptive visual prompti…

  8. RESEARCH · CL_04920 ·

    New CGC framework boosts multimodal LLMs for fine-grained image understanding

    Researchers have introduced Compositional Grounded Contrast (CGC), a new framework designed to enhance the fine-grained multi-image understanding capabilities of Multimodal Large Language Models (MLLMs). This approach a…

  9. FRONTIER RELEASE · CL_02354 ·

    OpenAI's new models let ChatGPT think with images for advanced reasoning

    OpenAI has introduced its latest visual reasoning models, o3 and o4-mini, which allow AI to "think with images" as part of its internal reasoning process. These models can perform image manipulations like cropping and z…