📣 Send us your press release
Site updates every 15 minutes
Technology

AI Agent Evaluation Trust Rises, Failure Rate Static in New Survey

A new survey reveals that enterprise trust in automated AI agent evaluations nearly tripled in July, yet the failure rate for agents passing these evaluations remained unchanged.

12 August 2026
AI Agent Evaluation Trust Rises, Failure Rate Static in New Survey

Enterprise trust in automated AI agent evaluation methods saw a significant increase in July, according to VentureBeat Pulse Research. The share of organizations fully trusting these automated evaluations rose from 5% to 13%. Concurrently, complaints about evaluations not aligning with real-world outcomes decreased.

Despite the heightened confidence, the actual failure rate of AI agents in production has not improved. Nearly half of organizations (49%) reported deploying an agent or LLM feature in the past year that passed internal evaluations but subsequently caused customer-facing failures. This figure remains statistically indistinguishable from the previous month's 50%.

The surge in trust appears to be concentrated among enterprises with no prior negative experiences. 24% of these organizations fully trust automated evaluations, compared to only 4% among those that have encountered failures. Paradoxically, previous failures do not slow the adoption of autonomy; instead, they appear to accelerate it.

On the vendor side, the market shows signs of consolidation. The use of dedicated evaluation tooling is increasing, with ease of integration becoming a primary selection criterion over cost. The research is based on responses from 108 enterprises with 100 or more employees.

Original source: venturebeat.com