PulseAugur
实时 15:38:57
English(EN) "Talkie" is a 13B language model trained on 260B tokens of pre-1931 English text. Its instruction-tuned variant uses historical etiquette, letters, encyclopedia

新的 130 亿参数语言模型 "Talkie" 在 1931 年前的英文文本上进行了训练

一款名为 "Talkie" 的新语言模型已发布,拥有 130 亿参数,并在 1931 年前的 2600 亿个 token 的英文文本上进行了训练。Talkie 的指令微调版本融入了历史礼仪、信件、百科全书和诗歌,以提供对时间性训练数据的受控视角。 AI

影响 该模型独特的训练数据可能为深入了解历史语言模式及其对人工智能的影响提供新见解。

排序理由 该集群描述了一个具有特定训练数据特征的新语言模型的发布,符合研究类别。[lever_c_demoted from research: ic=1 ai=1.0]

在 Mastodon — mastodon.social 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

新的 130 亿参数语言模型 "Talkie" 在 1931 年前的英文文本上进行了训练

本文如何被排名

Signal score
11 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
该集群描述了一个具有特定训练数据特征的新语言模型的发布,符合研究类别。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

完整方法见我们的编辑标准

报道来源 [1]

  1. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    "Talkie" 是一个拥有 130 亿参数的语言模型,在 2600 亿个 1931 年前的英文文本上进行了训练。其指令微调版本使用了历史礼仪、信件和百科全书

    "Talkie" is a 13B language model trained on 260B tokens of pre-1931 English text. Its instruction-tuned variant uses historical etiquette, letters, encyclopedias and poetry, offering a controlled view of temporal training data. # LLM # AI # Talkie https:// talkie-lm.com/introduci…