PulseAugur
EN
LIVE 07:42:22
ENTITY TrOCR

TrOCR

PulseAugur coverage of TrOCR — every cluster mentioning TrOCR across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
3
6 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
3
6 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

2 day(s) with sentiment data

RECENT · PAGE 1/1 · 6 TOTAL
  1. TOOL · CL_158618 ·

    New synthetic dataset boosts Persian OCR capabilities

    Researchers have introduced Persian Pixel, a large-scale synthetic dataset designed to improve Optical Character Recognition (OCR) for the Persian language. The dataset contains over 343,000 image-text pairs, generated …

  2. RESEARCH · CL_107933 ·

    TrOCR adapted for medieval manuscript recognition, study finds

    Researchers have explored adapting the TrOCR model for handwritten text recognition (HTR) on medieval manuscripts, a task complicated by the model's pre-training on modern text. Through controlled experiments on a 13th-…

  3. RESEARCH · CL_105258 ·

    Mamba models offer faster OCR but lag Transformer accuracy on historical texts

    Researchers have benchmarked State-Space Models (SSMs), specifically Mamba, against Transformers and BiLSTMs for Optical Character Recognition (OCR) on historical newspapers. The studies indicate that while Mamba-based …

  4. TOOL · CL_74626 ·

    AI agent built to safely summarize patient discharge data

    This article details the creation of an AI agent designed to summarize patient discharge information from PDF documents. The agent focuses on extracting structured data like diagnoses, medications, and allergies, priori…

  5. RESEARCH · CL_59084 ·

    New dataset BullingerDB targets historical text recognition

    Researchers have introduced BullingerDB, a new large-scale dataset designed for analyzing historical handwritten documents. The dataset, derived from the correspondence of Heinrich Bullinger, contains over 20,000 pages …

  6. TOOL · CL_15592 ·

    AI pipeline transcribes medieval legal texts with 88% accuracy

    Researchers have developed an open-source pipeline to transcribe medieval English legal manuscripts, which are written in a highly abbreviated form of medieval Latin. The system uses neural networks for segmentation and…