A new research paper explores whether adding syntactic and rhetorical information to text can improve the detection of incoherence in large language models. The study found that current model architectures were incompatible with this enriched data, leading to lower accuracy compared to plain text. However, the research also demonstrated that assessing textual coherence could be a useful proxy for identifying misleading content, as shown in experiments with a Brazilian disinformation dataset. Code and models for the study are publicly available. AI
IMPACT Current LLM architectures struggle with enhanced linguistic data, suggesting a need for architectural improvements to better understand and generate coherent text.
RANK_REASON Academic paper on LLM capabilities and evaluation. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →