Skip to content
Zumu

About Zumu

Five things we believe about answering a phone

Not a mission statement. A list of mechanisms we built because we believe them. Each one below links to the product page that makes it true, so you can check our work instead of taking our word for it.

A supervisor leaning in beside an agent, listening to a live call together

The judgment calls still belong to a person.

What we believe

Five beliefs, five mechanisms

No belief on this page is decorative. Each one is the reason a specific part of the product exists.

  1. The recording should survive the handoff.

    Most voice AI treats a transfer as an exit: dial a second call, drop the tape, hand your person a stranger. Zumu bridges the human into the same room the call is already in, so the recording never stops and a supervisor can still listen live after your own person picks up.

    How the handoff works
  2. You should never pay AI rates for human minutes.

    Once a call transfers, the AI billing clock stops. Only the minutes the agent actually handled reach the invoice; everything your own team does after that is recorded and reviewed, never billed. For one operator, that excluded about 36% of the weekly bill.

    See how metering works
  3. Latency is respect.

    Every extra second before the agent speaks is a second your caller spends wondering if anyone is there. We treat the gap between a caller's last word and the agent's first sound as a craft problem, worked one optimization at a time, not a number to chase for a slide.

    Read the latency engineering
  4. An agent should not guess. It should look it up.

    Point Zumu at your documents, recordings, and website, and it builds an actual knowledge graph, not a pile of embeddings, then queries it live while the call is still going. When the answer is not in there, the agent says so and brings in a person.

    How the knowledge graph works
  5. The team keeps the judgment calls.

    AI is the front door, not the replacement. A supervisor can listen to any live call, whisper a correction the caller never hears, or take the call over outright. The queue-hold experience belongs to your operation, not to chance.

    See live operations

How we build

Engineering honesty, not a highlight reel

A claim on this page is only worth as much as the query behind it. Here is how we keep that true.

  • We publish what we measure.

    Eleven separate engineering decisions sit between a caller's last word and the agent's first sound: a pre-rendered greeting, a prompt-cache warm ping, worker prewarming, and more. We narrate them as craft, in the order they run, rather than compress them into one score.

  • If we cannot show the query, we do not publish the number.

    Every figure on this site traces back to a query we can run again against production, not a benchmark we ran once. Where we do not have a verified number, we say so instead of rounding one up.

  • The bill has to survive a finance review.

    Every call carries a full session report: cost by component, which model provider served it, how long it waited in queue, and the trace of every tool the agent used. Nothing is rolled into a line item you cannot open.

The latency work and the ledger behind this page are things we can walk you through on a call, not slogans.

Talk to us

Come see the judgment calls we did not automate

Bring a call you wish had gone better. We will point an agent at your own documents and let you hear what it does with the next one.

About Zumu | The AI Call Center Platform