LCSHBench
PulseAugur coverage of LCSHBench — every cluster mentioning LCSHBench across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
New SHELF benchmark tests LLMs on library bibliographic tasks
A new benchmark system called SHELF has been developed to evaluate the performance of language models on bibliographic tasks relevant to libraries and archives. The system generates synthetic data based on Library of Co…
-
New SHELF benchmark tests LLMs on library bibliographic tasks · 2 sources tracked
Researchers have developed SHELF, a Synthetic Harness for Evaluating LLM Fitness, designed to benchmark bibliographic tasks for libraries and archives. This Python system generates controlled benchmark data from labeled…
-
New benchmark evaluates multilingual Library of Congress subject heading assignment
Researchers have introduced LCSHBench, a new benchmark dataset designed to evaluate automated subject cataloging for Library of Congress Subject Headings (LCSH). The dataset comprises 22,346 books in 15 languages, sourc…