PulseAugur
EN
LIVE 17:37:33
ENTITY SciFact

SciFact

PulseAugur coverage of SciFact — every cluster mentioning SciFact across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
2
10 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
2
10 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 12 TOTAL
  1. TOOL · CL_244827 ·

    New framework prioritizes LLM answer review under budget constraints

    Researchers have developed a new framework for prioritizing answers from large language models (LLMs) for human review, particularly when review budgets are limited. The proposed method, termed 'review value,' considers…

  2. TOOL · CL_218879 ·

    AI agent BioCheck Agent generates detailed biomedical fact-checking reports

    Researchers have developed a new AI agent called BioCheck Agent, designed to generate detailed fact-checking reports for biomedical information. This agent utilizes an enhanced search strategy, exclusively querying PubM…

  3. RESEARCH · CL_218043 ·

    Automated fact-checking systems show domain-dependent performance, retrieval remains key

    A new paper evaluates the robustness of automated fact-checking (AFC) systems across different domains and metrics, finding that system rankings are highly dependent on the specific dataset and evaluation criteria. The …

  4. TOOL · CL_206353 ·

    AI struggles with rare-disease diagnosis, new paper reveals

    A new research paper explores the challenges of selective prediction in rare-disease diagnosis using AI. The study found that even advanced open-weight LLMs struggle with ultra-rare diseases, achieving low recall rates.…

  5. RESEARCH · CL_187156 ·

    New method maps embedding model similarity spaces for better RAG

    A new research paper introduces Synthetic Query Probing, a method to analyze and map similarity score spaces across different embedding models. This technique addresses the challenge that scores are not directly compara…

  6. TOOL · CL_180507 ·

    Biomedical claim verification: LLMs show promise in evidence generation

    A new study published on arXiv explores the effectiveness of evidence-generating Large Language Models (LLMs) for biomedical claim verification. The research, conducted on the CARE-XAI benchmark, compares various LLM ap…

  7. TOOL · CL_181169 ·

    Hyperbolic geometry retrieval system enables RAG on edge devices

    Researchers have developed a novel hybrid retrieval system that leverages hyperbolic geometry for retrieval-augmented generation (RAG) on edge devices. This system projects word embeddings into hyperbolic space, allowin…

  8. TOOL · CL_181225 ·

    New BANDMAS framework slashes multi-agent communication costs

    Researchers have developed BANDMAS, a novel framework for multi-agent collaboration that optimizes communication by intelligently scheduling data packets. This system analyzes semantic features of messages to determine …

  9. TOOL · CL_174011 ·

    New AI workflow enhances claim-evidence traceability in writing

    Researchers have developed a new workflow called evidence-ledger adjudication to improve the traceability of claims made by AI agents in relation to supporting evidence. This system pairs each claim with an evidence pac…

  10. TOOL · CL_160666 ·

    New Protocol Enhances Citation Faithfulness in Agentic LLM Synthesis

    A new research paper introduces a protocol and a guard system designed to improve the reliability of citation faithfulness checks in agentic large language model (LLM) systems. These systems, like OpenScholar and PaperQ…

  11. RESEARCH · CL_107791 ·

    New SIFT method improves LLM fact-checking accuracy

    Researchers have developed a new method called SIFT (claim-conditioned re-scoring) to improve the accuracy of fact-checking systems that use large language models (LLMs). These systems often incorrectly label claims as …

  12. RESEARCH · CL_86680 ·

    Small LLMs match GPT-4o/GPT-5 on biomedical claim verification

    A new study demonstrates that fine-tuning smaller language models like Mistral-7B using QLoRA can achieve performance comparable to or exceeding larger models such as GPT-4o and GPT-5 on biomedical claim verification ta…