PulseAugur
中
实时 00:47:51
English(EN) The Fellowship of the Query: Learning Retrieval Actions

新研究探讨LLM长上下文检索和RAG策略 · 追踪9个来源

近期研究探索了大型语言模型(LLMs)如何处理长上下文,研究调查了性能提升背后的机制。一篇论文检查了思维链(CoT)推理,发现它能够实现定向检索和比广泛检索更紧凑的表示。另一项研究分析了不同的位置编码选择,如RoPE和滑动窗口注意力,如何将模型从位置检索转移到语义检索,从而影响问答等任务的性能。此外,研究比较了各种检索增强生成(RAG)策略,强调了重排序和后期交互对于科学问答的重要性,并引入了学习检索动作和自我评估探索的新框架,以增强知识检索。 AI

影响 检索和上下文处理的进步对于提高LLM在复杂、长周期任务和领域特定知识上的性能至关重要。

排序理由 多篇arXiv论文详细介绍了关于LLM检索机制和RAG策略的新研究。

在 arXiv cs.IR (Information Retrieval) 阅读 →

AI 生成摘要 · Google Gemini · 来自 12 个来源。 我们如何撰写摘要 →

新研究探讨LLM长上下文检索和RAG策略 · 追踪9个来源

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
多篇arXiv论文详细介绍了关于LLM检索机制和RAG策略的新研究。
Source corroboration
12 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
paper, model release
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
13 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.
Coverage growth since scoring
+1 source(s) since last score
New sources have picked up this story since our last re-score. Score will update on the next scoring pass.

完整方法见我们的编辑标准。

报道来源 [12]

  1. arXiv cs.AI TIER_1 English(EN) · Liang Twist Shan, Tianyu Hu, Hao Yan, Yiqiao Zhong ·

    定向检索、紧凑表示:CoT推理如何改进长上下文计数

    arXiv:2609.38958v1 Announce Type: new Abstract: Large language models (LLMs) have been rapidly improving in long-context tasks, powered by Chain-of-Thought (CoT) reasoning. However, the internal mechanisms underlying this improvement remain unclear. We investigate these mechanism…

  2. arXiv cs.CL TIER_1 English(EN) · Eric Enouen, Sainyam Galhotra ·

    机制转变:位置编码选择如何影响上下文检索

    arXiv:2609.38530v1 Announce Type: new Abstract: Language models increasingly use architectures that vary attention span and positional encoding across layers, such as applying RoPE with sliding-window attention and NoPE with global attention (SWA NoPE). However, how these choices…

  3. arXiv cs.AI TIER_1 English(EN) · Bhagyesh Rathi, Eshan Chawla, William B. Andreopoulos ·

    重排与后期交互提升检索质量:RAG策略在科学问答中的受控比较

    arXiv:2609.38473v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) is now the standard way to ground Large Language Models (LLMs) in external knowledge, yet the design space of retrieval pipelines is large and the trade-offs between variants are not well under…

  4. arXiv cs.CL TIER_1 English(EN) · Jingyuan Ma, Lynx Aster, He Zhang, Siyao Song, Weijie Yuan, Zhe Zhang, Kai Jia, Zhifang Sui ·

    Traverse:学习何时记忆、重置和重定向以进行长时程网络搜索

    arXiv:2609.37082v1 Announce Type: new Abstract: Long-horizon information-seeking agents often accumulate noisy or misleading context, causing early mistakes to persist and making recovery increasingly difficult. We introduce an autonomous search harness in which the agent manages…

  5. arXiv cs.IR (Information Retrieval) TIER_1 English(EN) · William B. Andreopoulos ·

    重排和后期交互提升检索质量:RAG策略在科学问答中的受控比较

    Retrieval-Augmented Generation (RAG) is now the standard way to ground Large Language Models (LLMs) in external knowledge, yet the design space of retrieval pipelines is large and the trade-offs between variants are not well understood, especially on domain-specific corpora at re…

  6. Perplexity blog TIER_1 English(EN) ·

    Photon:从头开始构建检索和排名引擎

    How specialized data formats, batched reads, and the separation of index building from query serving improve AI-native search.

  7. arXiv cs.AI TIER_1 English(EN) · Mohammed Al-Maamari, Saber Zerhoudi, Michael Granitzer, Jelena Mitrovi\'c ·

    查询的伙伴:学习检索动作

    arXiv:2609.28653v1 Announce Type: cross Abstract: Retrieval-augmented question answering requires control decisions about when to decompose a question, search, reformulate, extract evidence, synthesize facts, verify progress, and stop. We study whether trajectory fine-tuning can …

  8. arXiv cs.IR (Information Retrieval) TIER_1 English(EN) · Ebrahim Bagheri ·

    Seek:用于知识检索的自我评估探索

    LLM-based retrievers and rerankers have advanced passage ranking, yet both paradigms interact with the corpus in a single pass and commit to the resulting candidate set, leaving relevant documents permanently unrecoverable once missed. We introduce Seek, Self-Evaluative Explorati…

  9. arXiv cs.IR (Information Retrieval) TIER_1 English(EN) · Jelena Mitrović ·

    查询的伙伴:学习检索动作

    Retrieval-augmented question answering requires control decisions about when to decompose a question, search, reformulate, extract evidence, synthesize facts, verify progress, and stop. We study whether trajectory fine-tuning can improve small language models (SLMs) as next-actio…

  10. Towards AI TIER_1 English(EN) · M. Haseeb Hassan ·

    RAG 与长上下文在 2026 年:检索仍然占优

    <p>Anthropic’s own benchmark tells an uncomfortable story about dumping everything into the prompt: even with a generous context window, standard retrieval still missed the right chunk 5.7% of the time, and fixing that took a second retrieval technique, not a bigger window [1]. T…

  11. Towards AI TIER_1 English(EN) · Jaival Suthar ·

    构建知识库:从 RAG 实现到检索工程

    <h4>What a controlled retrieval experiments revealed about ranking, evidence, latency, and failure modes.</h4><figure><img alt="" src="https://cdn-images-1.medium.com/max/1024/1*LIOvevb1bkfEFw5HWs6gGg.png" /></figure><p>Most RAG demos are deceptively simple.</p><p>Ingest a docume…

  12. dev.to — LLM tag TIER_1 English(EN) · Ruchita Nimkar ·

    超越首次回答:当检索成为调查

    <h2> What happens when a question looks simple, but answering it correctly requires more than retrieving a few documents? </h2> <p>For my <strong>TigerGraph Hackathon</strong> project, I explored this question by implementing and comparing <strong>RAG, GraphRAG, and Agentic Graph…