PureDocBench
PulseAugur coverage of PureDocBench — every cluster mentioning PureDocBench across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
NaviDC-OCR framework enhances document parsing with deformation-aware learning
Researchers have introduced NaviDC-OCR, a novel framework designed to enhance document parsing across both digital and camera-captured documents. This system addresses limitations in existing methods by incorporating de…
-
OvisOCR2 document parsing model achieves state-of-the-art on benchmarks · 3 sources tracked
Researchers have introduced OvisOCR2, a new 0.8 billion parameter document parsing model capable of generating Markdown representations of documents, including text, formulas, and tables, in natural reading order. The m…
-
OvisOCR2: A new 0.8B OCR model converts documents to structured Markdown
OvisOCR2 is a new 0.8B parameter end-to-end OCR model that converts full document pages into structured Markdown. Based on Qwen3.5-0.8B, it reportedly achieves high scores on document understanding benchmarks like OmniD…
-
New PureDocBench benchmark reveals document parsing is far from solved
Researchers have introduced PureDocBench, a new benchmark for document parsing that addresses issues with the existing OmniDocBench dataset, which suffers from annotation errors and potential contamination. PureDocBench…