PulseAugur
EN
LIVE 07:12:10

New benchmark Ekphrasis measures visual creative ideation in text-only LLMs

Researchers have introduced Ekphrasis, a new benchmark designed to evaluate the visual creative ideation capabilities of text-only language models. This benchmark measures a model's ability to generate useful, expressive, and novel textual visual plans, separating these qualities from mere fluency. Ekphrasis comprises 400 tasks across abstraction, combination, transformation, and adaptation, using pairwise comparisons and Bradley-Terry models to aggregate preferences and identify population clichés. A cross-modal grounding study indicates that the ordering of visual ideation capabilities identified by Ekphrasis largely holds true even after faithful rendering and blind image-level preference judgments, suggesting its effectiveness in assessing visual ideation beyond prose quality. AI

IMPACT This benchmark could lead to more nuanced evaluations of LLMs' creative potential beyond text generation.

RANK_REASON The cluster contains an academic paper introducing a new benchmark and methodology for evaluating language models. [lever_c_demoted from research: ic=1 ai=1.0]

Read on arXiv cs.CL →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

New benchmark Ekphrasis measures visual creative ideation in text-only LLMs

COVERAGE [1]

  1. arXiv cs.CL TIER_1 English(EN) · Hongyu Luo, He Wang, Huihao Jing, Hong Ting Tsang, Yuxuan Liu, Wuganjing Song, Yauwai Yim, Chunyang Li, Yangqiu Song ·

    Can Language Models Imagine Without Seeing? Ekphrasis: Measuring Visual Creative Ideation in Text-Only LLMs

    arXiv:2608.06967v1 Announce Type: new Abstract: Current evaluations do not isolate whether text-only language models can originate visual concepts before image generation. Fluent visual prose can hide visual-plan failures: an answer may appear creative while repeating familiar vi…