TrOCR
PulseAugur coverage of TrOCR — every cluster mentioning TrOCR across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
New synthetic dataset boosts Persian OCR capabilities
Researchers have introduced Persian Pixel, a large-scale synthetic dataset designed to improve Optical Character Recognition (OCR) for the Persian language. The dataset contains over 343,000 image-text pairs, generated …
-
TrOCR adapted for medieval manuscript recognition, study finds
Researchers have explored adapting the TrOCR model for handwritten text recognition (HTR) on medieval manuscripts, a task complicated by the model's pre-training on modern text. Through controlled experiments on a 13th-…
-
Mamba models offer faster OCR but lag Transformer accuracy on historical texts
Researchers have benchmarked State-Space Models (SSMs), specifically Mamba, against Transformers and BiLSTMs for Optical Character Recognition (OCR) on historical newspapers. The studies indicate that while Mamba-based …
-
AI agent built to safely summarize patient discharge data
This article details the creation of an AI agent designed to summarize patient discharge information from PDF documents. The agent focuses on extracting structured data like diagnoses, medications, and allergies, priori…
-
New dataset BullingerDB targets historical text recognition
Researchers have introduced BullingerDB, a new large-scale dataset designed for analyzing historical handwritten documents. The dataset, derived from the correspondence of Heinrich Bullinger, contains over 20,000 pages …
-
AI pipeline transcribes medieval legal texts with 88% accuracy
Researchers have developed an open-source pipeline to transcribe medieval English legal manuscripts, which are written in a highly abbreviated form of medieval Latin. The system uses neural networks for segmentation and…