PulseAugur
EN
LIVE 12:54:33
ENTITY Omni-MATH

Omni-MATH

PulseAugur coverage of Omni-MATH — every cluster mentioning Omni-MATH across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
3 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
3 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 3 TOTAL
  1. TOOL · CL_142786 ·

    Researchers explore diminishing returns in LLM benchmark size using IRT

    Researchers explored the diminishing returns of increasing benchmark size for Large Language Models (LLMs) using Item Response Theory (IRT). They found that while IRT provides a theoretical framework for measuring the i…

  2. TOOL · CL_119403 ·

    Research probes how language agents effectively use feedback for improvement

    A new research paper investigates the effectiveness of feedback in improving language agent performance. The study introduces a controlled student-teacher protocol across multiple benchmarks, comparing external feedback…

  3. RESEARCH · CL_104766 ·

    New decoding strategy bypasses LLM alignment tax for better reasoning

    Researchers have introduced a novel decoding strategy called Confident Decoding, which aims to mitigate the "alignment tax" in large language models. This tax occurs when final layers of LLMs, after being fine-tuned for…