PulseAugur
中
实时 02:46:22
English(EN) model : add support for talkie-1930-13b by niklassheth · Pull Request #22596 · ggml-org/llama.cpp

新的‘复古’LLM 使用 1931 年前的英文文本进行训练

一个名为 talkie-1930-13b-it 的新语言模型已被开发出来,它仅使用 1931 年前的英文文本进行训练。这个拥有 130 亿参数的模型是 talkie-1930-13b-base 的指令微调版本,而 talkie-1930-13b-base 最初是在 2600 亿个 token 上训练的。微调过程利用了源自历史参考作品的指令-响应对的独特数据集,随后进行了强化学习以增强其遵循指令的能力。 AI

影响 该模型提供了一种新颖的历史语言模拟方法,可能为研究过去的沟通风格和知识开辟新的途径。

排序理由 该集群描述了一个具有独特训练数据集和方法论的新语言模型发布,并附有报告和 GitHub 存储库。[lever_c_demoted from research: ic=1 ai=1.0]

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

新的‘复古’LLM 使用 1931 年前的英文文本进行训练

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
该集群描述了一个具有独特训练数据集和方法论的新语言模型发布,并附有报告和 GitHub 存储库。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, paper
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
126 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/pmttyji ·

    模型:添加对 niklassheth 的 talkie-1930-13b 的支持 · Pull Request #22596 · ggml-org/llama.cpp

    <!-- SC_OFF --><div class="md"><blockquote> <p><a href="https://huggingface.co/talkie-lm/talkie-1930-13b-it">https://huggingface.co/talkie-lm/talkie-1930-13b-it</a> </p> </blockquote> <p><strong>talkie-1930-13b-it</strong> </p> <p>talkie-1930-13b-it is a 13B vintage language mode…