FUNSD
PulseAugur coverage of FUNSD — every cluster mentioning FUNSD across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
LLMs struggle with noisy documents, new benchmark reveals
A new research paper benchmarks several open-source large language models (LLMs) for key-value pair extraction from documents, specifically examining their performance under Optical Character Recognition (OCR) noise. Th…
-
New VLM extracts document data without OCR, outperforming larger models
Researchers have developed a new method for extracting key-value pairs from document images without relying on traditional OCR preprocessing. They fine-tuned a compact 256M-parameter vision-language model called SmolDoc…
-
FRAGMENT framework uses factorized graphs for document generation and editing
Researchers have introduced FRAGMENT, a novel generative framework for document creation and editing that utilizes factorized graph representations. This approach separates the generation of document structure from its …