Durable approval is not the same as valid approval
A developer argues that durable approval systems for AI agents can block valid commands, leading to unnecessary restrictions.

- Durable approval systems can mistakenly block valid AI agent commands due to rigid deny-lists.
- The conflation of durability and validity in approval systems leads to over-restrictive behavior.
- Real-world deployment of AI agents requires more nuanced validation to avoid false positives.
- Balancing safety and functionality is critical as AI agents become more autonomous.
In a recent technical post, developer Jackson Xie highlights a critical flaw in durable approval systems for AI agents. The system, designed to prevent harmful commands, can mistakenly block valid actions due to its rigid deny-list approach. Xie explains that while durable approval aims to enhance safety, it often conflates durability with validity, leading to over-restrictive behavior.
The post delves into a real-world example where a general command tool for a local agent was deployed with a deny-list in front. Despite the system's intention to block invalid commands, it inadvertently rejected legitimate requests, demonstrating how durability does not always equate to correctness. Xie emphasizes the need for more nuanced validation mechanisms that distinguish between harmful and valid commands without imposing unnecessary restrictions.
This issue is particularly relevant as AI agents become more autonomous, requiring systems that balance safety with functionality. The discussion underscores the importance of refining approval mechanisms to avoid false positives that hinder productivity.
Highlights the need for refined approval mechanisms in AI agent systems to avoid unnecessary restrictions.
Raises awareness about the limitations of current safety systems in AI agents.
- Durable approval
- A system designed to persistently block invalid or harmful commands in AI agents.
- Deny-list
- A predefined list of commands or actions that are explicitly blocked by a system.
AI ToolsAI Writes 41% of Code. Only 29% of Devs Trust It. Review It Like a Senior Engineer
AI ToolsEurope's free satellite service just made it easier to track wildfires
AI ToolsAnthropic loosens Fable 5's biology restrictions but keeps the guardrails on for virology and toxicology
How Cohere Health digitizes clinical policies using Amazon Bedrock AgentCore - Amazon Web Services (AWS)
How TReNDS automates root-cause analysis with Amazon Bedrock - Amazon Web Services (AWS)
NeonatAIlogy: The GIRISH (Goal, Input, Role, Iterative Refinement, Safety Verification, and Human Accountability) Framework for Structured, Human-Accountable Artificial Intelligence in Neonatal Intensive Care - Cureus
Researchers propose the GIRISH framework for human-accountable AI in neonatal intensive care. The framework aims to ensure safe and effective AI integration in healthcare.
Oneida County summer youth program introduces students to AI - Rome Sentinel
Oneida County's summer youth program is introducing students to artificial intelligence. The program aims to educate students about AI and its applications.
OpenAI says it slowed Astra model development over security concerns
OpenAI has temporarily halted parts of its Astra AI model development because of security vulnerabilities. The move reflects growing scrutiny of AI safety in advanced models.
In the News: John Abraham Discusses AI Safety Concerns - Newsroom | University of St. Thomas
John Abraham, a prominent AI safety advocate, shares his thoughts on the pressing concerns surrounding AI development in an interview with the University of St. Thomas.
Who is liable when Artificial Intelligence goes rogue? - FOX 29 Philadelphia
The issue of liability when artificial intelligence systems fail or cause harm is becoming increasingly important. Experts are debating who should be held responsible in such cases.
BusinessAfter Rippling blew millions on AI in months, it built an employee ROI tool
Rippling introduces AI Spend Console to monitor AI tool spending by employees and teams, addressing cost inefficiencies after rapid AI adoption.