ResearchQA
PulseAugur coverage of ResearchQA — every cluster mentioning ResearchQA across labs, papers, and developer communities, ranked by signal.
-
New SARA method improves LLM judge consistency by mitigating rubric interference
Researchers have developed a new method called Self-Anchored Rubric Alignment (SARA) to address rubric interference in large language model (LLM) judges. This interference occurs when LLMs evaluate multiple rubrics in a…
-
New AI training method uses rubrics as privileged information for open-ended generation
Researchers have developed a new method called On-policy self-distillation (OPSD) that utilizes rubrics as privileged information (PI) for open-ended text generation. This approach enhances the training signal for model…
-
New Deep Research Pretraining framework enhances AI agent training
Researchers have developed Deep Research Pretraining (DRP), an offline framework designed to improve the training of deep research agents. DRP derives supervision from existing evidence structures like citation graphs a…
-
New benchmark evaluates LLMs on citation-grounded scientific paper Q&A
Researchers have introduced ResearchQA, a new benchmark designed to evaluate how well large language models can answer questions based on scientific papers while ensuring answers are supported by verifiable citations. T…