SciFact
PulseAugur coverage of SciFact — every cluster mentioning SciFact across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
New framework prioritizes LLM answer review under budget constraints
Researchers have developed a new framework for prioritizing answers from large language models (LLMs) for human review, particularly when review budgets are limited. The proposed method, termed 'review value,' considers…
-
AI agent BioCheck Agent generates detailed biomedical fact-checking reports
Researchers have developed a new AI agent called BioCheck Agent, designed to generate detailed fact-checking reports for biomedical information. This agent utilizes an enhanced search strategy, exclusively querying PubM…
-
Automated fact-checking systems show domain-dependent performance, retrieval remains key
A new paper evaluates the robustness of automated fact-checking (AFC) systems across different domains and metrics, finding that system rankings are highly dependent on the specific dataset and evaluation criteria. The …
-
AI struggles with rare-disease diagnosis, new paper reveals
A new research paper explores the challenges of selective prediction in rare-disease diagnosis using AI. The study found that even advanced open-weight LLMs struggle with ultra-rare diseases, achieving low recall rates.…
-
New method maps embedding model similarity spaces for better RAG
A new research paper introduces Synthetic Query Probing, a method to analyze and map similarity score spaces across different embedding models. This technique addresses the challenge that scores are not directly compara…
-
Biomedical claim verification: LLMs show promise in evidence generation
A new study published on arXiv explores the effectiveness of evidence-generating Large Language Models (LLMs) for biomedical claim verification. The research, conducted on the CARE-XAI benchmark, compares various LLM ap…
-
Hyperbolic geometry retrieval system enables RAG on edge devices
Researchers have developed a novel hybrid retrieval system that leverages hyperbolic geometry for retrieval-augmented generation (RAG) on edge devices. This system projects word embeddings into hyperbolic space, allowin…
-
New BANDMAS framework slashes multi-agent communication costs
Researchers have developed BANDMAS, a novel framework for multi-agent collaboration that optimizes communication by intelligently scheduling data packets. This system analyzes semantic features of messages to determine …
-
New AI workflow enhances claim-evidence traceability in writing
Researchers have developed a new workflow called evidence-ledger adjudication to improve the traceability of claims made by AI agents in relation to supporting evidence. This system pairs each claim with an evidence pac…
-
New Protocol Enhances Citation Faithfulness in Agentic LLM Synthesis
A new research paper introduces a protocol and a guard system designed to improve the reliability of citation faithfulness checks in agentic large language model (LLM) systems. These systems, like OpenScholar and PaperQ…
-
New SIFT method improves LLM fact-checking accuracy
Researchers have developed a new method called SIFT (claim-conditioned re-scoring) to improve the accuracy of fact-checking systems that use large language models (LLMs). These systems often incorrectly label claims as …
-
Small LLMs match GPT-4o/GPT-5 on biomedical claim verification
A new study demonstrates that fine-tuning smaller language models like Mistral-7B using QLoRA can achieve performance comparable to or exceeding larger models such as GPT-4o and GPT-5 on biomedical claim verification ta…