TechCrunch · AI· Aditya Mehta·· 5 小时前AI 评分62
Goodfire 推出内部激活监控,拦截失控 AI 智能体
Goodfire says its new ‘inside-out’ monitors catch rogue AI agents at a fraction of the cost
AI 导读
Goodfire 发布一种监控 AI 智能体的新方案,通过探针读取模型内部激活信号,而非让第二个模型逐字审查输出,已在 Baseten 平台向客户开放。在 Kimi K3 测试中,监控约 1500 个会话花费约 51 美元,而用较便宜的 AI 模型逐步检查需 233 美元、顶级模型约 10000 美元;探针捕获了 94% 的恶意黑客会话,并将 8.7% 的无害会话送交二次检查。
来源:TechCrunch · AI · techcrunch.com