PulseAugur
EN
LIVE 14:26:59

EVOKE 14B: Interactive world model generates persistent, re-promptable video

EVOKE 14B is a new interactive world model designed for persistent, coherent video generation. It operates in three steps with no classifier-free guidance, allowing for rapid generation of 384x640 video at 24 frames per second. The model decouples world state from the generation process, enabling it to maintain coherence and respond to user instructions over extended sessions, reportedly for hours. AI

IMPACT Enables longer, more interactive and coherent video generation sessions.

RANK_REASON Release of a new model with weights and a paper, but not from a tier-1 frontier lab. [lever_c_demoted from research: ic=1 ai=1.0]

Read on r/StableDiffusion →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

EVOKE 14B: Interactive world model generates persistent, re-promptable video

COVERAGE [1]

  1. r/StableDiffusion TIER_2 English(EN) · /u/Crazy-Repeat-2006 ·

    EVOKE 14B - a 3-step, CFG-free interactive world model

    <table> <tr><td> <a href="https://www.reddit.com/r/StableDiffusion/comments/1vpw1z4/evoke_14b_a_3step_cfgfree_interactive_world_model/"> <img alt="EVOKE 14B - a 3-step, CFG-free interactive world model" src="https://external-preview.redd.it/ZnoxYnJpdGhmcWpoMZfEeKDjYYs6jl-UXwoq76A…