Researchers have developed VisPilot, a system that uses multimodal prompts, combining text, sketches, and direct manipulation, to improve visualization authoring with large language models (LLMs). This approach addresses the limitations of text-only prompts, which can be imprecise and lead to misinterpretations. An empirical study found that multimodal prompts help users communicate spatial constraints and design preferences more effectively, maintaining similar task efficiency to text-only methods. The findings offer design implications for future human-AI authoring systems. AI
IMPACT Multimodal prompting could improve the precision and efficiency of AI-assisted design tools.
RANK_REASON The cluster contains an academic paper detailing a new system and empirical study. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →