跳到正文
原文
arXiv cs.AI· Stefan G. Creadore·· 10 小时前AI 评分42

Praxa:一种证据约束的 AI 智能体执行框架

From Proposal to Verified Effect: Praxa, an Evidence-Bound Harness for Governed AI Agent Execution

AI 摘要

Praxa 提出一种通过确定性准入、代理执行和外部回读显式化智能体状态转换的证据约束架构。在 Terminal-Bench Core 0.1.1 测试中,其可靠性层比基线多消耗 37.49% 输入 token 和 50.73% 输出 token;另一对比实验显示候选方案虽减少 37.11% token 和 33.84% 成本,但未证明质量或延迟提升。当前证据未确立对抗安全、生产安全性或通用优势。

来源:arXiv cs.AI · arxiv.org