PulseAugur
实时 14:10:16
English(EN) Infinity-Parser2 Technical Report

Infinity-Parser2 模型通过合成数据和多任务强化学习推进文档解析

研究人员推出了 Infinity-Parser2,这是一个用于端到端文档解析的大型多模态模型。该模型利用可控数据合成管道和多任务强化学习来克服标注解析数据的稀缺性。主要贡献包括创建了 500 万样本的双语语料库 Infinity-Doc2-5M,以及一个用于跨八个联合训练目标的联合强化学习的新颖奖励系统。Infinity-Parser2-FlashInfinity-Parser2-Pro 两个变体已发布,其中后者在 olmOCR-BenchParseBench 等基准测试中取得了最先进的成果。 AI

影响 推进文档解析能力,可能提高处理各种文档类型的效率和准确性。

排序理由 该集群描述了一份技术报告,详细介绍了在 arXiv 上发布的新多模态模型和数据集。[lever_c_demoted from research: ic=1 ai=1.0]

在 arXiv cs.AI 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Infinity-Parser2 模型通过合成数据和多任务强化学习推进文档解析

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
该集群描述了一份技术报告,详细介绍了在 arXiv 上发布的新多模态模型和数据集。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
paper, model release, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
48 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准

报道来源 [1]

  1. arXiv cs.AI TIER_1 English(EN) · Zuming Huang, Jun Huang, Kexuan Ren, Baode Wang, Weizhen Li, Jianming Feng, Yu Wang, Yichen Yao, Shijun Lin, Yige Tang, Cheng Peng, Weidi Xu, Wei Chu, Yinghui Xu, Yuan Qi ·

    Infinity-Parser2 技术报告

    arXiv:2607.07836v1 Announce Type: new Abstract: We present Infinity-Parser2, a large multimodal model that couples a controllable data-synthesis pipeline with multi-task reinforcement learning for end-to-end document parsing, addressing the persistent scarcity of faithfully annot…