85% of companies burned by an AI mistake are racing to cut the humans who might catch the next one

EDITOR BRIEF
VentureBeat Intelligence survey data shows rising enterprise trust in automated AI evaluations, even though 49% of respondents said an AI agent or LLM feature passed testing and later caused a customer-visible problem. Companies that experienced these failures were far less likely to fully trust automated checks, but many are still pushing toward reduced human involvement in deployment decisions.
INSIGHTS
The findings point to a widening gap between confidence in AI evals and their proven ability to prevent real-world failures. As enterprises scale agent deployments, demand may grow for continuous monitoring and mitigation tools that catch issues after release rather than relying on pre-launch tests alone.
COMMENTS
Discussion
> geekhaus:~$ next read?
Next read recommendations
virtualizationhowto.com
Broadcom removes public VMware VDDK downloads, potentially complicating migrations from vSphere to rival virtualization platforms

The Verge
Audi’s new A2 E-tron is its most affordable and efficient EV yet
TechCrunch