AI takes the call—but customers still want a human on the line

The gist
Despite the AI takeover in customer service, most consumers still demand a human touch—and companies are scrambling to build trust before automation backfires.
What to know
- By early 2026, 85% of consumers still prefer human agents for local service, with only 14% willing to buy based on AI recommendations alone.
- Companies expect AI to handle half of all support interactions within 18 months—but only 31% have real measurement systems or oversight in place.
- Fortune 2000 firms are ditching fragmented chatbots for unified AI platforms, letting AI resolve up to 60% of tier two queries while freeing humans for complex cases.
Trust Hinges on Transparency
Consumers remain wary of AI-led service, demanding clear boundaries, accountability, and visible human oversight to build confidence and reduce frustration.
By early 2026, consumer trust in AI-driven customer experience remains notably low despite growing adoption, with only 14% willing to purchase based solely on AI recommendations and a mere 2% preferring exclusively AI-led interactions. This skepticism is fueled not only by doubts about AI’s reliability but also by perceptions that AI often hides behind automation rather than genuinely assisting customers, leading to frustration and a preference for human interaction, as evidenced by ServiceForge’s finding that 85% of consumers favor humans when contacting local service providers.
Consumers clearly distinguish the role AI should play in customer experience, favoring it as a discovery and tactical tool rather than a decision-maker or frontline service agent. For instance, 31% are persuaded to buy only when AI provides detailed, transparent product information, while 84.9% prefer human agents for complex or emotional interactions, underscoring the 'use AI where it earns the right' approach advocated by analysts who recommend reserving AI for low-emotion tasks like scheduling and routing to build trust gradually.
Transparency, accountability, and clear role definitions between AI and human agents emerge as critical levers to overcome consumer resistance and build trust. CX leaders are urged to define precisely what AI handles, what humans own, and what success looks like, especially during high-stress moments, while companies like Zoom exemplify this by scoring bots and humans with the same quality standards and embedding AI into workflows that reduce repeat contacts and customer effort, thereby shifting the focus from speed to reliable, complete resolution.
Trust erosion risks remain high when AI implementations lack reliability and seamless integration, with consumers attributing failures more to poor business execution than to AI technology itself. Research by Pegasystems and YouGov highlights consumer demands for 'background-blending AI agents' that consistently deliver successful outcomes, while monitoring trust signals such as complaints mentioning 'recording' or 'creepy' is essential to maintain confidence, especially as fast AI decisions can lead to costly mistakes if customers feel trapped or misled.
Governance Gaps Threaten Scale
Without robust measurement and real-time oversight, rapid AI adoption in support risks unchecked errors and brand-damaging automation failures.
By early 2026, organizations were rapidly scaling agentic AI in customer support, with 78% expecting these agents to handle at least half of interactions within 18 months. However, this rapid adoption outpaced the establishment of robust accountability frameworks, as only 31% had measurement systems for agentic AI and 44% lacked them even for generative AI. Experts emphasized the critical need to assign clear ownership for AI measurement, define quality metrics beyond mere containment, and institute rigorous review cadences—such as bi-weekly evaluations during the initial 90 days—to prevent systemic failures like misrouted disputes or hallucinated policies from proliferating unchecked across thousands of customers.
The operational governance of AI agents demands a multifaceted control environment that balances automation with human oversight. Leading companies like Wonderful have integrated automated evaluations, role-based access controls, audit logging, and privacy safeguards directly into their platforms, recognizing that AI agents connected to core systems of record amplify both value and risk. This necessitates strict approval workflows, traceability, and rollback mechanisms to mitigate costly errors. Moreover, real-time observability—tracking resolution rates, latency, user sentiment, and business tags—is indispensable for maintaining quality assurance and enabling swift intervention when AI agents encounter complex scenarios or edge cases.
The paradigm is shifting from simply building autonomous AI agents to mastering governed autonomy, where structured supervision and accountability become paramount. IBM exemplifies this evolution by deploying real-time operating models that treat AI agents as part of a workforce, monitored continuously by humans who can intervene on critical outputs. This approach underscores the necessity of dedicated ownership and budgeting for agent observability and oversight, moving beyond treating these as incidental features. Clear escalation paths, ownership of thresholds, and efficient recovery processes are essential to prevent customer frustration and costly brand damage, especially given that automation scales mistakes faster than human agents, as Hans van Dam warns.
Recent advances highlight the imperative of unifying human and AI workflows within a single management system to operationalize accountability effectively. Microsoft’s Dynamics 365 Contact Center exemplifies this blended workforce model by enabling joint planning, coaching, and measurement of service reps and AI agents in one platform. This unified view is critical to avoid merely shifting customer pain points through better dashboards without resolving them. Comprehensive audits of AI-assisted service flows—identifying stop points, handoff ownership, early harm metrics, and recovery paths—combined with empowering supervisors to detect issues in real time, form the backbone of resilient AI governance that safeguards customer experience.
Foundational to scaling accountable AI in contact centers is establishing rigorous quality management for human agents first, as emphasized by Dave Rennyson of SuccessKPI. This baseline enables scalable validation of AI agent performance through benchmarking AI on real conversation samples before deployment, ensuring competence and minimizing hallucinations. Furthermore, a unified governance framework that dynamically manages both human and AI agents based on real-time performance data optimizes task allocation and customer outcomes. Underpinning all these efforts is robust data governance and infrastructure, which organizations treating as prerequisites to effective hybrid contact center management, recognizing that siloed data environments render accountability nearly impossible.
Humans and AI: Clear Roles
Effective customer experience now depends on structured collaboration, with humans and AI assigned to their strengths and seamless escalation for sensitive moments.
By early 2026, industry leaders emphasized that successful AI integration in customer experience hinges on clearly defined roles between AI and human agents, especially during high-stress interactions. Platforms like Trace exemplify this approach by breaking workflows into discrete steps routed to the appropriate agent—human or AI—while embedding human-in-the-loop reviews and permissions to maintain control and accountability. This clarity prevents customers and agents from being 'left hanging,' fostering trust through transparent collaboration rather than mere automation.
AI voice agents, such as Zendesk's, demonstrate how automation can handle routine calls end-to-end while seamlessly escalating complex issues with full context, preserving the essential human touch. However, balancing automation with human judgment requires setting clear boundaries—avoiding sensitive topics like payments or angry customers—and focusing on the quality of handoffs rather than just call deflection metrics. This nuanced approach aligns with findings that 95% of customers expect explanations when AI makes decisions, underscoring that losing control equates to losing trust.
The evolving hybrid workforce model redefines human agents’ roles from safety nets to active collaborators in continuous AI improvement. As Declan Ivory notes, freeing humans from repetitive tasks allows them to concentrate on critical elements like tone, handoffs, and proactive outreach, which are pivotal for trust and loyalty. Companies like Burger King leverage AI assistants to support frontline staff in real time, while maintaining clear controls and escalation protocols to manage risks and uphold customer confidence.
Continuous training and workforce transformation are vital to balancing automation with human judgment, as highlighted by Cresta’s Training Simulator which uses AI-powered simulated customers to prepare agents for complex escalations. Despite 79% of service professionals investing in agentic AI, Zendesk reports that 55% of agents lack ongoing training, revealing a readiness gap that undermines handling high-stakes calls. Organizations that embrace continuous, data-driven upskilling see faster ramp times, reduced attrition, and improved business outcomes, reinforcing that effective AI-human collaboration demands sustained investment in human capital.
AI as Accountable Labor
New metrics like Agentic Work Units are transforming AI from experimental tool to measurable, governable workforce with clear operational impact.
By early 2026, the introduction of Agentic Work Units (AWUs) by Salesforce marked a pivotal shift in AI performance measurement, moving the focus from vague interaction counts to quantifiable, outcome-based metrics that treat AI as accountable labor. This approach enables organizations to manage automation operationally—tracking volume, quality, rework, and risk by intent—thereby reducing hidden costs like customer recontacts and cleanup. As AWUs become a standardized unit of work, AI transitions from experimental projects to integral, governable line items within customer service operations.
The rapid evolution of AI capabilities demands continuous adaptation of quality assurance frameworks, with monthly updates to QA rubrics and skills training to capture emerging failure modes effectively. Yet, as 78% of organizations anticipate agentic AI handling over half of customer support interactions within 18 months, only a minority have established measurement frameworks, underscoring a governance gap. Experts emphasize assigning clear ownership and defining resolution-focused quality metrics with bi-weekly reviews during the critical first 90 days post-deployment to ensure AI delivers true resolution rather than mere containment.
AI-powered quality assurance platforms, such as Solidroad and Observe.AI, have revolutionized performance management by enabling evaluation of 100% of customer interactions rather than traditional sampling. DoorDash’s partnership with Observe.AI and AWS exemplifies this transformation, achieving near-complete automated quality coverage across 19,000 agents and reducing issue detection time from weeks to near real-time. This shift from binary scoring to diagnostic insights uncovers behavioral drivers behind customer pain points, allowing quality teams to focus on higher-value analysis, consistent coaching, and stronger accountability across human and AI agents alike.
Effective AI performance management in hybrid human-AI contact centers hinges on robust data governance and unified frameworks that treat human and AI agents equivalently. Dave Rennyson highlights the necessity of establishing a 'ground truth' benchmark by testing AI on real conversation samples before going live to prevent hallucinations and model drift. This integrated approach enables dynamic, real-time workforce decisions optimized for customer experience, while strong data infrastructure ensures scalable validation and accountability—foundations without which managing AI at scale becomes nearly impossible.
Unified Platforms, Real Results
Fortune 2000 firms are ditching patchwork chatbots for all-in-one AI platforms that boost efficiency, consistency, and human agent value.
By early 2026, the customer service landscape is witnessing a decisive shift from fragmented chatbot solutions to unified AI platforms that consolidate all service functions into a single, seamless system. This integration eliminates the cumbersome need to patch together disparate third-party tools, ensuring consistent context between AI and human agents and streamlining service delivery. As one industry analysis put it, 'There is no concept of adding on... it's all just part of this one thing,' a model particularly favored by Fortune 2000 companies seeking to avoid the complexity of managing multiple AI vendors or custom builds.
Facing mounting pressure from boards to simultaneously cut costs and elevate customer satisfaction, service leaders are increasingly turning to scalable AI platforms capable of autonomously handling a growing share of tier two queries—aiming for efficiencies that push AI resolution rates from 50% toward 60%. This evolution not only reduces operational burdens but also enhances service quality by enabling human agents to focus on complex, high-value interactions, a dynamic highlighted by Craig Walker of Dialpad who emphasizes AI’s role in augmenting routine tasks like order status checks and password resets.
Beyond automation, integrated AI platforms are transforming customer service management through real-time agent coaching and actionable insights for supervisors, fostering a culture of continuous improvement that delivers measurable ROI. Craig Walker underscores that successful AI adoption hinges on methodical groundwork—cleaning knowledge bases, analyzing ticket patterns, piloting controlled deployments, and scaling AI agents thoughtfully across the organization—ensuring that technology enhancements translate into tangible business outcomes.
Data Quality Drives Outcomes
Firms leveraging integrated, high-quality data and real-time analytics are seeing faster resolutions, empowered agents, and measurable service gains.
By early 2026, financial services firms like Capita leveraged AI-powered analytics platforms such as CallSight to revolutionize quality assurance, shifting from periodic manual audits to continuous, organization-wide analysis of 100% of customer calls. This transformation not only reduced average handling time by 12–15% and improved first-call resolution by 15%, but also empowered human teams to focus on interpreting structured insights and delivering high-value coaching rather than routine compliance tasks, thereby enhancing both operational efficiency and customer satisfaction.
Multiquip’s journey underscores the critical importance of data quality and integration in AI-driven customer support. Initial reliance on traditional PDF-based content led to unreliable AI outputs and hallucinations, prompting a strategic pivot to Boomi’s agentic AI platform that leveraged existing system interfaces for scalable, trustworthy automation. This approach freed experienced technicians from mundane inquiries, enabling them to adopt advisory roles that enriched customer guidance and boosted both employee productivity and service quality.
Traeger’s adoption of Amazon Connect’s AI-powered cloud contact center exemplifies how real-time access to comprehensive customer data and knowledge bases can dramatically improve agent effectiveness, especially during peak demand. By reclaiming control over their technology stack from third-party vendors, Traeger uncovered true performance metrics—revealing a first-contact resolution rate of only 33% versus inflated vendor claims—and implemented targeted training and knowledge management improvements. Complementing AI virtual agents handling routine queries with skilled human agents tackling complex issues further enhanced operational efficiency and customer satisfaction.
DoorDash’s collaboration with Observe.AI and AWS highlights the power of iterative co-creation in scaling AI-driven customer experience across 19,000 agents. Their conversational intelligence platform automated nearly 100% of interaction evaluations, accelerating issue detection from weeks to near real-time and uncovering behavioral drivers behind customer pain points. Importantly, automation augmented rather than replaced human quality teams, enabling them to focus on nuanced analysis such as customer safety and fairness, thus reinforcing trust and accountability. This partnership exemplifies how continuous refinement aligned AI capabilities with operational realities and customer priorities.
Bespin Global’s deployment of generative AI customer service assistants within a secure private cloud for AIA Life demonstrates that innovation and regulatory compliance can coexist in highly regulated sectors. By meticulously verifying and optimizing AI components like document parsing and answer generation, they achieved improved operational efficiency and consistent customer responses while adhering to strict security standards. Beyond query handling, Bespin plans to extend AI support to broader call center digital transformation initiatives including agent training, quality assurance, and customer feedback analysis, showcasing the expanding role of generative AI in trusted, accountable customer experience.





