PulseAugur
实时 07:04:45
English(EN) Environment Evolution for Terminal Agents

新的环境演化方法提升终端代理性能 · 追踪 4 个来源

研究人员开发了一种名为“环境演化”的新方法来改进终端代理的训练。该技术在策略外(off-policy)逐步增加训练环境的难度,随着模型的进步提供持续的学习信号。使用该方法进行的实验在 Terminal-Bench 2.1 基准测试中显示出显著的性能提升,Qwen3.6-27B 和 Qwen3.6-35B-A3B 模型分别提高了 14.4 和 18.0 个百分点。该方法已与 Hy4 preview、Claude Opus 5 和 GPT-5.6 Sol 等前沿模型进行了测试,证明了其在生成更具挑战性环境方面的有效性。 AI

影响 该方法可以加速能够与复杂环境交互的 AI 代理的开发和性能提升。

排序理由 该集群报道了一篇详细介绍 AI 代理训练新方法的学术论文。

在 Hugging Face Daily Papers 阅读 →

AI 生成摘要 · Google Gemini · 来自 4 个来源。 我们如何撰写摘要 →

新的环境演化方法提升终端代理性能 · 追踪 4 个来源

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
该集群报道了一篇详细介绍 AI 代理训练新方法的学术论文。
Source corroboration
4 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
paper, model release
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
5 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准

报道来源 [4]

  1. arXiv cs.AI TIER_1 English(EN) · Zhiyuan Fan, Tinghao Yu, Yuanjun Cai, Jiang Zhou, Jiangtao Guan, Jincheng Liu, Yun Yang, Dingxin Hu, Zhuo Han, Xing Wu, Feng Zhang, Lilin Wang ·

    终端代理的环境演化

    arXiv:2609.04128v1 Announce Type: new Abstract: Scaling interactive and verifiable environments is critical for training terminal agents. As frontier models become more capable, environments synthesized from scratch become less challenging and thus provide limited learning signal…

  2. Hugging Face Daily Papers TIER_1 English(EN) ·

    终端代理的环境演化

    Scaling interactive and verifiable environments is critical for training terminal agents. As frontier models become more capable, environments synthesized from scratch become less challenging and thus provide limited learning signals. Recent co-evolution methods iteratively synth…

  3. Hugging Face Daily Papers TIER_1 English(EN) ·

    终端代理的环境演化

    Environment evolution incrementally raises task difficulty off-policy to sustain continuous learning signals for terminal agents, improving benchmark performance through multi-agent harnesses.

  4. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    终端代理环境演进:18分提升 新arXiv预印本:终端AI代理通过逐步训练获得14至18个基准点提升

    Environment evolution for terminal agents: 18-point gain New arXiv preprint: terminal AI agents gain 14 to 18 benchmark points by training against progressively harder environments. Results are unverified. https://www. notatechguy.com/environment-ev olution-for-terminal-agents-18…