PulseAugur
EN
LIVE 21:57:52
ENTITY Elo

Elo

PulseAugur coverage of Elo — every cluster mentioning Elo across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
2
12 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
2
11 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

2 day(s) with sentiment data

RECENT · PAGE 1/1 · 12 TOTAL
  1. RESEARCH · CL_193054 ·

    New benchmarks and training methods for LLM social reasoning unveiled

    Researchers have introduced Social Gym, a new environment featuring 21 multi-agent social games designed to objectively benchmark and improve LLM social reasoning. The system uses an Elo tournament to rank models, revea…

  2. TOOL · CL_171870 ·

    AI system predicts football outcomes using tactical profiles, outperforming traditional methods

    Researchers have developed Sim2Win, a novel framework for predicting football match outcomes and profiling team tactics without relying on team names or identities. This system utilizes event-based data to construct tac…

  3. TOOL · CL_133582 ·

    New ELO algorithm enhances learned optimizers for long-horizon tasks

    Researchers have developed a new meta-training algorithm called Efficient Long-Horizon (ELO) learning to address limitations in current learned optimizers (LOs). ELO efficiently scales meta-training to long-horizon inne…

  4. TOOL · CL_138256 ·

    New ELO algorithm enhances learned optimizers for long-horizon tasks

    Researchers have developed a new meta-training algorithm called ELO (Efficient Long-hOrizon) to improve learned optimizers (LOs). ELO addresses the challenges of scaling meta-training to long-horizon problems and compet…

  5. TOOL · CL_126174 ·

    Football prediction engine Model90 uses Bayesian methods for 2026 World Cup forecasts

    A bioreactor engineer has developed Model90, a statistical forecasting engine for football matches, including the 2026 FIFA World Cup and major European competitions. The engine uses an eight-stage pipeline that incorpo…

  6. TOOL · CL_111648 ·

    New chess rating system uses cognitive model to track skill changes

    Researchers have developed a new skill assessment framework for chess called the Drift-Diffusion-Enhanced Elo Rating System (DD-Elo). This system draws inspiration from cognitive neuroscience's drift diffusion model to …

  7. TOOL · CL_104712 ·

    New methods assess physical consistency in AI-generated videos

    Researchers have developed new methods to evaluate the physical consistency of videos generated by world models, addressing a gap in current simulation tools. These reference-free measures combine relative and absolute …

  8. TOOL · CL_93296 ·

    New framework uses AI to guide human comparisons for efficient ranking

    Researchers have developed a novel human-in-the-loop ranking framework called Surprise-Guided MergeSort (SGS). This system uses a Vision-Language Model (VLM) to identify comparisons that genuinely require human judgment…

  9. RESEARCH · CL_79519 ·

    New research validates pairwise comparisons for AI model accuracy

    A new research paper proposes that pairwise comparisons, commonly used to evaluate generative models, align well with accuracy-based rankings. The study converted five benchmarks into generative evaluations and found th…

  10. RESEARCH · CL_91476 ·

    New methods improve LLM evaluation accuracy with AI and human insights

    Researchers have developed new methods to improve the accuracy and calibration of Large Language Model (LLM) evaluations. One approach, Conformal Elo Estimation, uses LLM judgments to estimate Elo ratings, achieving res…

  11. RESEARCH · CL_22018 ·

    Study finds global LLM leaderboards misleading, proposes portfolio rankings

    A new research paper argues that current leaderboards for large language models (LLMs) are misleading due to significant heterogeneity in user preferences across languages and tasks. The study analyzed approximately 89,…

  12. TOOL · CL_17792 ·

    Chess-GPT model learns world model, can be manipulated to change skill

    Researchers have explored interventions on a language model trained to play chess, dubbed Chess-GPT. By manipulating the model's internal representations of the board state and player skill, they demonstrated a causal l…