PulseAugur
中
实时 01:14:28
English(EN) Trained a ~20K LM (probably smallest) that can still write stories

拥有20K参数的微型语言模型可生成连贯的故事

一位研究人员开发了MacroStories,一个拥有约20,000个参数的语言模型,这比TinyStories等之前的模型要小得多。该模型计算资源需求极低,可在CPU上快速运行,能够生成100-300字的连贯叙事,包括目标、问题、行动和结局。该项目旨在探索语言模型中连贯叙事生成的最低限度。 AI

影响 展示了高效、专业化语言模型的潜力,这些模型可以在最少的硬件上运行。

排序理由 发布了一个新的、显著更小的语言模型,专注于特定能力(叙事生成)。[lever_c_demoted from research: ic=1 ai=1.0]

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

拥有20K参数的微型语言模型可生成连贯的故事

本文如何被排名

Signal score
1 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
发布了一个新的、显著更小的语言模型,专注于特定能力(叙事生成)。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, other
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Same-day
Cluster formed today. Ranking reflects the current source set at time of score.

完整方法见我们的编辑标准。

报道来源 [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/x_Raincandy_x ·

    训练了一个约20K的LM(可能是最小的),仍能写故事

    <!-- SC_OFF --><div class="md"><p>I’ve been pushing TinyStories-style models downward in size, and this is the smallest one so far:</p> <p>MacroStories — 19,969 parameters, 81 KB FP32</p> <p><a href="https://huggingface.co/raincandy-u/MacroStories">https://huggingface.co/raincand…