AI在线 AI在线

Patronus AI Launches Percival: One-Minute Diagnosis of Hidden Faults in Hundred-Step Agent Chains

As enterprises increasingly deploy autonomous AI agent systems, the demand for monitoring and debugging these complex systems is rapidly growing. Today, AI security company Patronus AI, headquartered in San Francisco, released its latest product, Percival, a monitoring platform capable of automatically identifying fault patterns in AI agent systems and providing repair recommendations."Percival is the industry's first intelligent agent that can automatically track agent trajectories, identify complex faults, and systematically output repair suggestions," said Anand Kannappan, CEO and co-founder of Patronus AI, in an exclusive interview with VentureBeat.Solving the Real-World Challenges of "Uncontrollable" AI AgentsDifferent from traditional machine learning, AI agents can autonomously execute large-scale operation processes involving multiple stages.

As enterprises increasingly deploy autonomous AI agent systems, the demand for monitoring and debugging these complex systems is rapidly growing. Today, AI security company Patronus AI, headquartered in San Francisco, released its latest product, Percival, a monitoring platform capable of automatically identifying fault patterns in AI agent systems and providing repair recommendations.

"Percival is the industry's first intelligent agent that can automatically track agent trajectories, identify complex faults, and systematically output repair suggestions," said Anand Kannappan, CEO and co-founder of Patronus AI, in an exclusive interview with VentureBeat.

Solving the Real-World Challenges of "Uncontrollable" AI Agents

Different from traditional machine learning, AI agents can autonomously execute large-scale operation processes involving multiple stages. However, it is precisely this "multi-step autonomy" that makes fault debugging extremely challenging: a small error in the early stage may evolve into a serious deviation in subsequent processes, and multi-agent collaborative scenarios further exacerbate this complexity.

Percival is designed to address this pain point, capable of identifying over 20 common faults across four major categories, including reasoning errors, execution errors, planning misalignments, and domain-specific errors. More importantly, it is not a "post-hoc" solution but actively monitors the entire agent trajectory, possessing "contextual memory" to understand the ins and outs of errors in specific contexts.

"Percival itself is also an AI agent, so unlike traditional evaluators, it does not make static judgments but can track and learn fault evolution paths at the system level," said Darshan Deshpande, a researcher at Patronus.

Holographic Projection Robot Design (2)

Image source note: Image generated by AI, licensed by Midjourney

From One Hour to One Minute: Significant Improvement in Debugging Efficiency

In practical applications, Percival has significantly improved fault analysis efficiency. Patronus stated that its early customers have compressed the time required to debug complex agent processes from about one hour to 1 to 1.5 minutes, greatly alleviating the maintenance burden on engineering teams.

To standardize evaluation capabilities, Patronus also released the TRAIL Benchmark Test (Tracking Reasoning and Agent Issue Localization). The results showed that even the strongest models currently available scored only 11% on this test. This highlights the urgent need for professional AI regulatory tools.

Enterprise Deployment and Integration: High-Complexity Agent Safety Barriers

Percival has been adopted by several clients, including Emergence AI and Nova. Satya Nitta, CEO of Emergence AI, which focuses on developing systems for "agent creation agents," said that Percival provides critical assurance for achieving controllability in large-scale autonomous systems.

Nova, on the other hand, is using Percival to build an AI-driven platform to help businesses migrate SAP systems and integrate legacy code, with their agent system processes involving hundreds of steps, far exceeding human-controlled complexity.

Percival can seamlessly integrate with mainstream frameworks such as Hugging Face Smolagents, Langchain, Pydantic AI, and OpenAI Agent SDK, covering a wide range of agent development ecosystems.

Accelerating Growth in AI Security and Regulatory Tracks

With AI technology rapidly commercializing, enterprises generate billions of lines of AI code daily. Kannappan pointed out: "Systems are becoming more and more autonomous, while human supervision capabilities are far from keeping up."

相关资讯

Patronus AI 推出 Percival:一分钟诊断百步代理链中的隐藏故障

随着企业越来越多地部署自主运行的 AI 代理系统,对这些复杂系统的监控与调试需求也迅速增长。 总部位于旧金山的 AI 安全公司 Patronus AI 今日发布了其最新产品 Percival,一个能够自动识别 AI 代理系统中故障模式并提出修复建议的监控平台。 “Percival 是业界首个可以自动追踪代理轨迹、识别复杂故障,并系统化输出修复建议的智能代理。
5/15/2025 11:01:55 AM
AI在线

3Cap 王康曼:我为什么投资 Cerebras Systems?

访谈 | 陈彩娴撰文丨朱可轩、赖文昕编辑丨陈彩娴本月初,美国知名 AI 芯片创业公司 Cerebras Systems 官宣,其已经向美国证券交易委员会 (“SEC”) 提交了一份有关其普通股首次公开发行的表格 S-1 登记声明草案——这一声明,进一步证实了外界对其今年计划上市的猜想。 Cerebras Systems 成立于 2015 年,创始人是 Andrew Feldman,是一家以打破英伟达垄断为目标的美国 AI 芯片创业公司。 它们为业内熟知的标签有二:一是研发了世界上最大的芯片,从最初的 WSE-1到今年新发布的 WSE-3 均体量庞大;二是曾在 2018 年 D 轮获得 OpenAI CEO Sam Altman 的注资。
8/13/2024 7:33:00 PM
朱可轩

钻石冷却的GPU即将问世:温度能降20度,超频空间增加25%

现阶段这一方案的前景如何? 我们尚不得而知。 未来 GPU 的发展方向,居然和钻石有关系?
11/18/2024 1:27:00 PM
机器之心
  • 1