OmniDocBench v1.6
PulseAugur coverage of OmniDocBench v1.6 — every cluster mentioning OmniDocBench v1.6 across labs, papers, and developer communities, ranked by signal.
3 day(s) with sentiment data
-
PaddlePaddle releases HPD-Parsing model with record throughput · 2 sources tracked
PaddlePaddle has released HPD-Parsing, a new lightweight document parsing model that utilizes a Hierarchical Parallel Decoding paradigm. This model achieves a new state-of-the-art score of 94.91% on the OmniDocBench v1.…
-
OvisOCR2 document parsing model achieves state-of-the-art on benchmarks · 3 sources tracked
Researchers have introduced OvisOCR2, a new 0.8 billion parameter document parsing model capable of generating Markdown representations of documents, including text, formulas, and tables, in natural reading order. The m…
-
OvisOCR2: A new 0.8B OCR model converts documents to structured Markdown
OvisOCR2 is a new 0.8B parameter end-to-end OCR model that converts full document pages into structured Markdown. Based on Qwen3.5-0.8B, it reportedly achieves high scores on document understanding benchmarks like OmniD…
-
HunyuanOCR-1.5 enhances lightweight OCR VLMs with faster inference and improved capabilities · 3 sources tracked
Researchers have introduced HunyuanOCR-1.5, an enhanced lightweight vision-language model specifically designed for Optical Character Recognition (OCR). This model unifies various document processing tasks, including pa…
-
US power sector M&A hits record $200B amid AI data center boom; Baidu's OCR model goes open-source
The US power and utility sector is experiencing a surge in M&A activity, with deal values exceeding $200 billion in the first five months of the year, driven by the demand for energy infrastructure to support data cente…
-
Baidu's Unlimited OCR model tops leaderboards and sees rapid adoption
Baidu has released and open-sourced its end-to-end OCR model, Unlimited OCR. The model quickly achieved top rankings on GitHub and Hugging Face, demonstrating strong performance in long document parsing. Unlimited OCR a…
-
Baidu's PaddleOCR-VL-1.6 sets new SOTA in document parsing
Baidu's Wenxin has released PaddleOCR-VL-1.6, a new version of its open-source OCR tool. This update achieves over 96.33% accuracy on the OmniDocBench v1.6 benchmark, surpassing major models like Gemini-3-Pro and GPT-5.…
-
New benchmarks and models advance document parsing and table extraction
Researchers have introduced new benchmarks and improved models for document parsing and table extraction. Dr. DocBench focuses on expert-level document parsing, including complex structures like chemical formulas and mu…
-
PaddleOCR-VL-1.6 sets new SOTA in document parsing
PaddlePaddle has released PaddleOCR-VL-1.6, an advanced document parsing model that achieves state-of-the-art accuracy on several benchmarks, including OmniDocBench v1.6 with a score of 96.33%. This new version incorpor…