Only 1 in 4 Customers Trust AI With a Refund

The Trust Gap

New research from Five9, surveying 3,000 consumers and 600 CX and contact center decision-makers across the US, UK, and Germany, finds that half of consumers believe AI can handle most customer service tasks they need help with. But that confidence collapses sharply once money enters the conversation: under a quarter of consumers said they trust AI to accurately approve refunds, make financial adjustments, or determine when a case needs escalation, and only 18% believe AI can understand complex issues or handle sensitive situations appropriately.

Why This Matters for BFSI and Collections

This is not a general AI skepticism problem; it is specific to financial decision-making, which puts it squarely in collections’ territory. A payment plan approval, a settlement negotiation, or a hardship determination is functionally the same category of decision customers say they do not trust AI to make alone. The data offers a clear mitigation, though: 55% of consumers said they trust AI more when there is an obvious option to reach a human, and just over half said they are less likely to do business with a company that removes human support entirely. For collections operations, the message is not that AI should stay out of financial conversations. It is that visible, easy human access has to be part of the design, not a fallback discovered only after something goes wrong.

What the Numbers Do Not Say Out Loud

The most revealing finding in this research is not about AI accuracy at all; it is about a perception gap between leadership and customers. Ninety-six percent of CX leaders believe their organization preserves context when a conversation moves from AI to a human agent. Meanwhile, 83% of consumers say they often or sometimes have to repeat themselves during that same handoff. That is not a small measurement error; it is two groups describing what looks like a different system entirely. CX leaders report the technical symptoms clearly enough: human agents lacking visibility into the AI interaction, glitches in the transfer itself, and context or data getting lost mid-channel-switch. But if leadership genuinely believes context is being preserved 96% of the time while customers experience the opposite, that gap itself is the actual risk.

The Practical Read

For collections and BFSI operations, the handoff moment deserves as much design attention as the AI conversation itself, arguably more, given that 87% of consumers report frustration by the time they reach a human after an AI interaction. A customer escalating from a voicebot mid-payment-negotiation should not have to restate their account details, their hardship situation, or what has already been discussed. The human agent needs that context waiting for them, not assembled after the fact. This matters even more given that 52% of consumers said an AI-powered channel was their only support option available to them. Getting the AI-to-human transition right is not a secondary concern behind getting the AI conversation right. For anything touching money, it is the more important half of the design.

[Read the full report]

Related Post

Enterprises Are Underestimating Multi-Model AI Failure Rates by 2.25x

A new study evaluating 67 frontier models from 21 providers, including GPT-5.5, Claude Opus 4.8, and Gemini 3.1 Pro, found that combining multiple AI models does not create the safety net most enterprises assume. On the MATH-500 benchmark, standard correlation metrics predicted a 2.3% “co-failure rate,” the share of prompts where every model in a […]

Oriserve’s Generative Voice AI Platform is Driving Strategic Transformation in BFSI Revenue Operations

 Oriserve (ORI), a bootstrapped startup with a team of over 100 professionals based in Mumbai and Delhi, is revolutionising enterprise communications as a next-generation voice-based Generative AI platform tailored for Banking, Financial Services and Insurance (BFSI). With over 1.2 billion conversations orchestrated globally, ORI is establishing a formidable presence in India and the Middle East […]

When 66% of Enterprises Deploy AI Agents Without Human Review

A June 2026 VB Pulse survey of 157 enterprise respondents found that half of organizations have shipped an AI agent or LLM feature that passed internal evaluation and still caused a customer-facing failure. One in four experienced this more than once. Despite that, 66% already permit production deployment without human review or are building toward […]