Quarterly Report · Q1 2026 Voice Agents · Customer Trust

Voice Agents · Q1 2026

Six cohorts, one quarter, measurable transformation.

48%
of customer calls in Q1 resolved without human intervention, across six production cohorts — up from 31% at end-of-year 2025.

This report summarizes the first full quarter of Voice Agents in production at commercial scale. Three cohorts grew, three new ones launched, and the operational signal across all six converged sooner than projected. What follows is what moved, why, and where we are heading into Q2.

Reporting period 2026-01-01 — 2026-03-31
Prepared by Customer Trust & Product Analytics
Audience Board · Leadership · Customer-facing teams
Executive Summary

What moved in Q1

Six production cohorts handled a combined 312,400 customer calls in Q1 2026. Voice Agents resolved 48% of those calls end-to-end — transactional steps included — without escalation. The remaining 52% followed the first-class human-handoff path, every escalation logged and reviewable. Operationally, this quarter is the first in which the Voice Agents product performs above the threshold required to redirect new customer engineering effort from onboarding onto cohort expansion1.

Three signals drove the quarter: resolution-rate stabilization across cohorts, a narrowing variance band between the best- and worst-performing deployments, and a sharper handoff profile — when the agent does escalate, it escalates faster and with better context attached to the ticket.

Figure 1
Resolution rate by cohort, Q1 2026
60% 45% 30% 15% 0% 46% 52% 41% 54% 43% 51% — — Q1 median · 48% Cohort A Cohort B Cohort C Cohort D Cohort E Cohort F

Resolution = call closed without human agent joining the line. Cohorts A–C entered Q1 already in production; D–F launched during the quarter.

Voice Agents · Q1 2026 — 2 — Ontopix · Customer Trust
Volume and Variance

Volume grew; variance shrank

Across the quarter, weekly audit volume climbed from 19,100 calls in the first full week to 27,800 in the last. That's a function of two things: cohorts D–F coming online in late January, and the early cohorts raising the cap on calls they route through the agent. The narrower story is about variance: the spread between the highest- and lowest-resolution cohort halved by quarter-end.

Figure 2
Weekly volume by cohort-age (stacked)
30k 20k 10k 0 W1 W2 W3 W5 W7 W9 W10 W11 W12 W13
Cohorts D–F (launched in Q1) Cohorts A–C (in production pre-Q1)

Volume is weekly call count routed through Voice Agents before any handoff decision. Excludes aborted lines and wrong-number filters.

The variance band between cohorts halved in 13 weeks. That is not a model improvement. That is an operational one — the rollout playbook is starting to converge. — Customer Trust internal review, 2026-04-09

The playbook convergence matters more than the headline resolution rate. A high-performing cohort without a reproducible rollout is a demo; a portfolio of cohorts landing within ten points of the median is a product. Q1 is the first quarter we've had the second.

Voice Agents · Q1 2026 — 3 — Ontopix · Customer Trust
Looking Ahead

What's next for Q2

Three priorities carry forward into Q2. Each is tracked against a numerical gate; none is marketed.

Q2 2026 priorities and their gating metrics.
PriorityGateCurrentTarget
Expand Cohort B & D envelopes Call types accepted per cohort 11 16
Median escalation latency Seconds from decision to human join 14.2 ≤ 10.0
Evidence-per-escalation Structured context fields attached 3.1 5.0
Rubric version drift Weeks since last rubric publish 4 ≤ 6

The cost of the 52%

It is easy to read the 48% resolution number as the headline. The more useful number is the 52% — the calls where the agent elected to escalate. In Q1, those escalations averaged 3.1 structured context fields attached to the handoff ticket; when an operator picked up the line, the customer was already identified, intent-tagged, and routed to the right specialist queue. That context is not incidental. It is the product argument.

In Q2 we push that number to five. Which sounds incremental until you watch a human agent pick up a Voice Agents handoff and read context instead of asking questions. The escalation stops being an interruption to the customer; it becomes a continuation of a conversation the AI already started well.

What this quarter is not

It is not a signal to expand headcount. It is not a reason to stop measuring. It is not evidence that the rollout playbook is complete — three cohorts (A–C) are an old sample; three (D–F) have been in production for less than 90 days. Statistical significance across the cohort set is in late-Q2 territory at the earliest.

Q2 is the first quarter where we ask the question Audibot is designed to answer: do the calls we didn't resolve look materially different from the calls we did?2 If they do, the rollout playbook extends. If they don't, we have a limit we need to explain.

  1. Threshold defined in the 2025-Q4 deployment plan: 45% median resolution across ≥ 6 cohorts for ≥ 6 weeks. Met in week 9 of Q1; held through week 13.
  2. Audibot's v2 API ships the criterion-level evidence needed to make that comparison tractable at population scale. See SPEC-AUDIBOT-v2 §3.
Voice Agents · Q1 2026 — 4 — Ontopix · Customer Trust