PulseAugur
实时 15:32:20
English(EN) A gate that rejects four out of four looks like a gate doing hard work. Mine was recording two opposite facts under one name: the safeguard firing, and the mode

AI模型安全门记录冲突结果

一位Mastodon用户分享了他们使用AI模型安全功能的经历,指出一个旨在拒绝有问题输入的门似乎运行有效。然而,他们观察到一个差异,即系统记录了两个冲突的结果:安全门被触发,以及模型根本没有提供任何响应。这表明模型性能和安全指标的记录方式可能存在问题。 AI

排序理由 社交媒体平台上的用户生成内容,讨论了关于AI模型行为的技术观察,缺乏更广泛的行业意义。

在 Mastodon — fosstodon.org 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

AI模型安全门记录冲突结果

报道来源 [1]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    A gate that rejects four out of four looks like a gate doing hard work. Mine was recording two opposite facts under one name: the safeguard firing, and the mode

    A gate that rejects four out of four looks like a gate doing hard work. Mine was recording two opposite facts under one name: the safeguard firing, and the model never answering at all. https:// devaland.com/blog/metric-said- it-was-working # AI # LLM # observability