PulseAugur
实时 10:28:19
English(EN) DeFiFlowBench: Benchmarking and Improving Safe Executability in Natural-Language DeFi Workflow Synthesis

新基准揭示AI生成的DeFi工作流中的安全缺陷

研究人员开发了DeFiFlowBench,这是一个旨在评估自然语言生成的去中心化金融(DeFi)工作流安全性的新基准。该基准包含207个提示,揭示了现有的提示方法即使在声明了安全谓词和价格影响上限的情况下,也经常导致不安全执行。为了解决这个问题,该团队提出了Koan-Safe,一个结合了意图解析、可替换生成器和具有默认安全参数的结构修复的系统,显著减少了在未见过提示上的不安全执行。 AI

影响 凸显了AI驱动的金融自动化中的关键安全问题,需要强大的评估框架。

排序理由 该项目是一篇研究论文,详细介绍了一个新的基准和用于评估AI生成工作流的拟议系统。[lever_c_demoted from research: ic=1 ai=1.0]

在 arXiv cs.LG 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

新基准揭示AI生成的DeFi工作流中的安全缺陷

本文如何被排名

Signal score
11 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
该项目是一篇研究论文,详细介绍了一个新的基准和用于评估AI生成工作流的拟议系统。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
paper, safety, product
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Same-day
Cluster formed today. Ranking reflects the current source set at time of score.

完整方法见我们的编辑标准

报道来源 [1]

  1. arXiv cs.LG TIER_1 English(EN) · Abhinav Rajeev Kumar, Harshit Arora, Varun Singh, Manikandan Nanjappan ·

    DeFiFlowBench:在自然语言 DeFi 工作流合成中对安全可执行性进行基准测试和改进

    arXiv:2609.11504v1 Announce Type: new Abstract: A structurally valid DeFi workflow can still authorize a costly trade. We introduce DeFiFlowBench, a benchmark of 207 team-authored prompts for natural-language DeFi workflow synthesis. It measures graph coverage, configuration comp…