PulseAugur
EN
LIVE 04:49:26
ENTITY Process Reward Model

Process Reward Model

PulseAugur coverage of Process Reward Model — every cluster mentioning Process Reward Model across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
1 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
1 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 4 TOTAL
  1. RESEARCH · CL_233341 ·

    New training method improves web agent world models

    Researchers have developed a new training objective called predicted-state matching for world models used in web agents. This method aims to improve the ability of these models to distinguish between different web state…

  2. TOOL · CL_94209 ·

    New Framework Enhances LLMs for Program Synthesis from Examples

    Researchers have developed a novel framework called PRM-PBE to enhance the ability of large language models (LLMs) in Programming-by-Example (PBE) tasks. This method addresses the limitation of current LLMs in PBE, whic…

  3. RESEARCH · CL_44959 ·

    New VRPRM model enhances LLM reasoning with visual cues

    Researchers have developed VRPRM, a novel process reward model that utilizes visual reasoning to enhance the fine-grained evaluation of Large Language Model (LLM) reasoning steps. This approach significantly reduces the…

  4. RESEARCH · CL_15892 ·

    New method debiases LLMs at decoding time, improving fairness without model retraining

    Researchers have developed a novel method to mitigate biases in large language models during the decoding phase, without altering the model's weights. This approach uses a separate Process Reward Model (PRM) to score to…