Treatmybrand


a Kainjoo SA Venture
Ch. du Vernay 14a
1196 Gland
+41.21.561.34.96
[email protected]

Support


Monday to Friday
8AM to 8PM
[email protected]
Back

85% of Companies Affected by AI Failures Accelerate Removal of Human Oversight in Deployments

Companies that have experienced AI features passing internal tests but failing in real-world use are now moving more quickly to reduce human involvement in deployment decisions, despite rising overall trust in automated evaluations. According to recent research by VentureBeat, 13% of surveyed enterprises trusted automated evaluations in July, up from 5% the previous month, while fewer respondents expressed concern about poor alignment between tests and actual outcomes. However, nearly half of the respondents reported customer-facing issues caused by AI features that had previously cleared testing, with 24% experiencing multiple such incidents. Intriguingly, organizations that had encountered these failures were less confident in automated checks but were more likely to allow AI-driven changes without human approval, suggesting a maturity in deployment processes rather than recklessness. Production monitoring for quality still lags behind, with many companies focusing more on whether AI systems function rather than whether their outputs are correct. Investment trends show increasing budgets for human review and automated tools aimed at catching errors missed by evaluations. The emerging market for independent evaluation tools highlights shifting priorities towards integration ease and consistent evaluation results. Despite improvements, the report underscores that internal testing alone is insufficient for ensuring AI reliability in production, pointing to a need for continuous monitoring and human oversight as a crucial safety net.

Venturebeat
Venturebeat