PulseAugur
中
实时 23:01:09
English(EN) A self-taught AI never trained on law just topped a Swiss law-exam benchmark

人工智能模型在未接受法律培训的情况下,在瑞士法律考试基准测试中名列前茅

由韩国初创公司VIDRAFT开发的名为Darwin-180B-RSI的开放权重180B模型,在法律推理基准测试中取得了最高排名,尽管它从未接受过法律数据训练。该模型采用了一种名为模型级递归自我改进(RSI)的新颖方法,通过解决练习题来学习,然后仅在自己的正确答案上进行训练。这种方法似乎培养了一种通用的推理能力,可以迁移到新领域,正如其在瑞士法律考试中相对于GPT-5、Claude-4.5-Sonnet和Gemini 2.5 Pro等领先模型的优异表现所证明的那样。 AI

影响 证明了通用推理技能可以在没有领域特定数据的情况下进行训练,有可能加速人工智能在新领域的适应。

排序理由 该条目详细介绍了一种新颖的人工智能模型训练方法及其在特定基准测试上的表现,与研究发现一致。[lever_c_demoted from research: ic=1 ai=1.0]

在 dev.to — LLM tag 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

人工智能模型在未接受法律培训的情况下,在瑞士法律考试基准测试中名列前茅

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
该条目详细介绍了一种新颖的人工智能模型训练方法及其在特定基准测试上的表现,与研究发现一致。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, paper
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
2 days old
Coverage has settled into its steady-state source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. dev.to — LLM tag TIER_1 English(EN) · ai maya ·

    一个从未接受过法律训练的自学AI刚刚在瑞士法律考试基准测试中名列前茅

    <p>Darwin-180B-RSI, an open-weight 180B model from the Korean startup VIDRAFT, is now <strong>#1 on LEXam and LEXam-hard</strong>, the two legal-reasoning leaderboards that Hugging Face lists as official benchmarks. The model was never trained on legal data.</p> <p>With these two…