PulseAugur
EN
LIVE 18:42:43

Qwen 3.8 27B model powers GTA-style game with 128k context

A user showcased a playable Grand Theft Auto-style game developed using the Qwen 3.8 27B model, which utilized its full 128k context window. The performance metrics indicate a baseline generation speed of 36-40 tokens/s for short or unique prompts, with typical sustained speeds ranging from 50-72 tokens/s during longer runs. Peak speeds reached up to 197 tokens/s, attributed to the use of n-gram caching, which significantly boosts efficiency when generating repetitive content. AI

IMPACT Demonstrates the practical application of large context window models for game development and complex prompt execution.

RANK_REASON User-generated showcase of a model's capabilities in a specific application.

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Qwen 3.8 27B model powers GTA-style game with 128k context

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/dsdt ·

    Another qwen 3.8 27b showcase - gta style prompt - also a remainder to use ngram in your configs.

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vtpo6w/another_qwen_38_27b_showcase_gta_style_prompt/"> <img alt="Another qwen 3.8 27b showcase - gta style prompt - also a remainder to use ngram in your configs." src="https://preview.redd.it/ewxbqcpuakkh1.…