Emily M. Bender is seeking a specific type of dataset for natural language processing research. The ideal dataset would represent data as Directed Acyclic Graphs (DAGs) or a similar structure, with the generated text containing only information present in the graph. Bender specifically mentioned a preference for datasets not sourced from Wikipedia, as the WebNLG dataset, which is derived from Wikipedia, has already been utilized. AI
IMPACT This search could lead to new datasets for training and evaluating data-to-text generation models.
RANK_REASON The item describes a specific data request for academic research in NLP. [lever_c_demoted from research: ic=1 ai=1.0]
Read on Bluesky Jetstream — AI desk →
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →