Enterprises burned by bad AI evals more likely to remove humans from loop

Across 108 enterprises, trust in automated agent evaluation nearly tripled from 5% in June to 13% in July, yet the failure rate it predicts did not move. Organizations that got burned by a bad eval are the most likely to remove humans from the loop.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Suno user reports published song changed after a year
- Suno users debate AI music quality patterns
- Databricks uses AI to accelerate incident investigation
- AWS introduces Agentic Resource Discovery (ARD) spec for agent discovery
- Users frustrated by Enter-to-send in AI chat interfaces