PulseAugur
实时 16:46:35
English(EN) Search, Inspect, Fetch: Exploiting Boolean Retrieval for Deep-Research Agents

新研究通过结构化检索和评分标准排名增强AI代理 · 跟踪4个来源

两篇新研究论文介绍了增强AI代理信息处理能力的创新方法。第一篇论文《搜索、检查、获取》提出了SIEVE接口,该接口利用字段化布尔检索(BQL)使代理能够将搜索限制在特定文档字段,从而提高准确性并减少令牌使用。第二篇论文《使用搜索评分标准训练文档重排序器》介绍了RubricRanker,一种使用LLM合成的搜索评分标准进行训练的文档重排序器,以确保检索到的文档集满足复杂的信息需求,在深度研究和RAG基准测试中表现优于基线。 AI

影响 这些进展可能显著提高AI代理在复杂研究任务中的效率和准确性。

排序理由 两篇在arXiv上发表的学术论文,详细介绍了AI研究代理的新方法。

在 Hugging Face Daily Papers 阅读 →

AI 生成摘要 · Google Gemini · 来自 7 个来源。 我们如何撰写摘要 →

新研究通过结构化检索和评分标准排名增强AI代理 · 跟踪4个来源

报道来源 [7]

  1. arXiv cs.CL TIER_1 English(EN) · Ming Zhang, Jiabao Zhuang, Wenqing Jing, Kexin Tan, Ziyu Kong, Jingyi Deng, Yujiong Shen, Yuhui Wang, Zhenghao Xiang, Qiyuan Peng, Yuhang Zhao, Ning Luo, Renzhe Zheng, Jiahui Lin, Mingqi Wu, Long Ma, Shihan Dou, Maxm Pan, Tao Gui, Qi Zhang, Xuanjing Huang ·

    深度研究代理能否检索和组织?使用专家分类法评估综合差距

    arXiv:2601.12369v5 Announce Type: replace Abstract: Deep Research Agents increasingly automate survey writing, yet existing benchmarks do not jointly test whether they retrieve the papers experts consider essential and organize those papers into paper-grounded taxonomies. We intr…

  2. arXiv cs.AI TIER_1 English(EN) · Wenhan Liu, Yu Lu, Qiaolin Xia, Hui Xu, Tong Zhao, Jian Xi, Yutao Zhu, Haijin Liang, Haibo Shi, Hao Wang, Zhicheng Dou ·

    使用搜索标准对训练文档进行重排序以用于深度研究代理

    arXiv:2608.03527v1 Announce Type: cross Abstract: Retrieval systems help deep research agents generate high-quality answers by providing relevant documents. However, existing retrievers typically select documents through relevance matching, while individually well-matched top-$k$…

  3. arXiv cs.AI TIER_1 English(EN) · Shuai Wang, Haodong Chen, Yu Yin, Shengyao Zhuang, Bevan Koopman, Guido Zuccon ·

    搜索、检查、获取:利用布尔检索构建深度研究代理

    arXiv:2608.02751v1 Announce Type: cross Abstract: Existing deep-research agents use a search-visit workflow that retrieves and reads whole pages, without considering the addressable structure that web sources expose through titles, headings, sections, and metadata. This prevents …

  4. arXiv cs.IR (Information Retrieval) TIER_1 English(EN) · Zhicheng Dou ·

    使用搜索标准对训练文档进行重排以用于深度研究代理

    Retrieval systems help deep research agents generate high-quality answers by providing relevant documents. However, existing retrievers typically select documents through relevance matching, while individually well-matched top-$k$ documents may not form a \textit{set} that satisf…

  5. arXiv cs.IR (Information Retrieval) TIER_1 English(EN) · Guido Zuccon ·

    搜索、检查、获取:利用布尔检索构建深度研究代理

    Existing deep-research agents use a search-visit workflow that retrieves and reads whole pages, without considering the addressable structure that web sources expose through titles, headings, sections, and metadata. This prevents agents from directly constraining retrieval to doc…

  6. arXiv cs.IR (Information Retrieval) TIER_1 English(EN) · Guido Zuccon ·

    搜索、检查、获取:利用结构感知布尔检索赋能深度研究代理

    Existing deep-research agents use a Search--Visit workflow that retrieves whole webpages without considering the structure they expose through titles, headings, sections, and metadata. This prevents agents from directly constraining retrieval to parts of a webpage and often carri…

  7. Hugging Face Daily Papers TIER_1 English(EN) ·

    搜索、检查、获取:利用布尔检索构建深度研究代理

    Existing deep-research agents use a search-visit workflow that retrieves and reads whole pages, without considering the addressable structure that web sources expose through titles, headings, sections, and metadata. This prevents agents from directly constraining retrieval to doc…