PulseAugur
实时 09:18:52
English(EN) Institutional Books - Visual Elements: An open-source pipeline for extracting, classifying, deduplicating, and captioning visual elements from digital book collections

开源管道从历史书籍中提取视觉元素

研究人员开发了一个名为 Institutional Books - Visual Elements 的开源管道,旨在从数字化的历史图书收藏中提取、分类、去重和标注视觉组件。该管道以及一个包含 2260 万个视觉元素的初始数据集已经发布,以促进对数字化图书馆材料的新应用。该项目旨在使插图和照片等视觉元素更容易用于计算用途,包括 AI 模型训练和数字人文研究。 AI

影响 通过计算访问实现数字化图书馆收藏的新用例,包括 AI 模型训练。

排序理由 该项目描述了一个用于处理数字化图书中视觉元素的开源管道和数据集发布,属于文化遗产计算访问方面的研发。 [lever_c_demoted from research: ic=1 ai=0.7]

在 arXiv cs.CV 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

开源管道从历史书籍中提取视觉元素

报道来源 [1]

  1. arXiv cs.CV TIER_1 English(EN) · Jimmy Mendez, Matteo Cargnelutti, David Lowry-Duda, Catherine Brobston, Salwa Ismail, Greg Leppert, Amanda Watson, Jonathan Zittrain ·

    机构藏书 - 视觉元素:一个用于从数字图书收藏中提取、分类、去重和标注视觉元素的开源管道

    arXiv:2608.18957v1 Announce Type: new Abstract: Historical book collections contain rich visual elements - such as illustrations, photographs, engravings, and decorative art - that are frequently under-explored in large-scale digitization projects. While Optical Character Recogni…