PulseAugur
实时 23:31:07
English(EN) I’ve started developing a new benchmark. I’d welcome any criticism, comments, or participation—if anyone wants to join the project right at the start, feel free

新的AI安全基准测试“Delirium”寻求社区输入

一个名为Delirium的新AI安全基准测试正在开发中,其创建者正在寻求社区的反馈和参与。该项目旨在评估大型语言模型的安全性和鲁棒性,并提供了一个视觉演示供审查。开发者正在积极寻找合作者为基准测试的创建做出贡献。 AI

影响 这个新的基准测试可能导致对LLM安全进行更严格的测试,从而可能影响未来的模型开发和部署策略。

排序理由 该项目描述了一个新的AI安全基准测试的启动,这属于研究范畴。[lever_c_demoted from research: ic=1 ai=1.0]

在 Mastodon — fosstodon.org 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

新的AI安全基准测试“Delirium”寻求社区输入

报道来源 [1]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    I’ve started developing a new benchmark. I’d welcome any criticism, comments, or participation—if anyone wants to join the project right at the start, feel free

    I’ve started developing a new benchmark. I’d welcome any criticism, comments, or participation—if anyone wants to join the project right at the start, feel free. I’ve put together a visual demo here: gitlab.com/toxy4ny/delirium-ai-safety-benchmark/ # redteam # ai -safety # benchm…