PulseAugur
EN
LIVE 14:27:40

AI models struggle with long-form fiction, exhibiting restart, freeze, and rewind failures · 1 source tracked

A recent experiment involving 360 rounds of AI novel continuation revealed three primary failure modes across nine different models. These failures include models restarting the narrative from the beginning, getting stuck in repetitive loops of their own generated text, or rewinding the plot to an earlier point. The study found that even advanced models like Gemini-3.1 Pro and GPT-5.6 Terra exhibited these issues, highlighting a disconnect between surface-level prose mimicry and long-range narrative coherence. Models like DeepSeek V4 Flash and V4 Pro showed stronger performance, with the former excelling in a fantasy epic and the latter in a palace-intrigue novel, suggesting that model performance is genre-dependent. AI

IMPACT Highlights limitations in current LLMs for long-form creative writing, suggesting a need for better context management and coherence in future models.

RANK_REASON The item details a research experiment and its findings on AI model capabilities. [lever_c_demoted from research: ic=1 ai=1.0]

Read on dev.to — LLM tag →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

AI models struggle with long-form fiction, exhibiting restart, freeze, and rewind failures · 1 source tracked

COVERAGE [1]

  1. dev.to — LLM tag TIER_1 English(EN) · Foreverse ·

    We Ran 360 Rounds of AI Novel Continuation. Models Fail Long Fiction in Exactly Three Ways

    <p>We have now run 360 rounds of novel continuation across nine models and two books — an 8.9M-character Chinese fantasy epic and the palace-intrigue classic <em>Empresses in the Palace</em>. Same protocol every time: continue from a fixed anchor, 20 consecutive rounds, each roun…