Will it understand callers with strong accents or a noisy background?
Usually yes, and it is the first thing we test. Speech recognition handles regional accents well now. Heavy background noise is the harder problem.
The realistic failure is not the accent on its own, it is a bad line plus a building site plus somebody speaking quickly. We test against recordings of your actual callers before launch rather than against a clean demo, because that is where the difference shows.
Where it cannot follow somebody after a couple of attempts, the right behaviour is to stop trying and put them through. An agent that keeps asking a frustrated caller to repeat themselves does more damage than one that gives up early.
How many callers hang up when they realise it is an AI?
Some do, and anyone telling you none do is selling. The number that matters is not hang-ups, it is how many of those callers you were reaching before.
The comparison people make is against a human answering instantly. That is rarely the real alternative. The real one is usually a voicemail nobody returns, or a ring-out while you are on a job, and a caller who hangs up on an AI was already lost in that scenario.
It is still worth measuring rather than assuming. Call outcomes can be reported so you can see hang-up rate against booking rate and decide with numbers.
We tried an AI phone system before and it was awful. Why would this be different?
Usually because the first one was a phone menu with a nicer voice. The test is whether it can finish the job on the call rather than take a message.
The common version answers, collects a name and number, and emails it to you. That is an answering machine with extra steps, and callers work it out within a sentence.
What actually makes the difference is unglamorous: whether it reads your live availability, whether it writes the booking back into your system, whether it hands off cleanly when it is out of its depth, and whether anybody maintains it after launch.
Can a caller skip the AI and get straight to a person?
Yes, and it should be easy. Asking for a human should transfer the call, not start a negotiation.
We set an explicit escape route. A direct request, repeated confusion, or any sign of distress ends the automation and routes the call: to your team in working hours, and to a message or whoever is on call outside them.
A system that traps people is the fastest way to lose the customers you were trying to catch.
What stops it inventing a price or promising something you do not offer?
Constraints, not good intentions. It works from a fixed set of facts you approve, and is instructed to say it will check rather than guess.
Prices, lead times, service areas and policies are supplied as data rather than left to the model. Anything outside that set gets a straight "I will have somebody confirm that" and a task for your team, which is what a sensible new starter would say anyway.
We then try to break it before launch: talk it into quoting, discounting, and committing to work you do not do, and tighten it until it stops.
Can it get an email address or a postcode right over the phone?
Yes, with read-back. It repeats the value and has the caller confirm it before saving.
Getting this wrong quietly is worse than failing loudly. A mistyped email means the confirmation never arrives and nobody finds out until an appointment is missed, so read-back is standard on anything that has to be exact.
Where a value can be checked automatically it is: postcode format, and the obvious mishearings in common email domains.
Which calls should we never let it handle?
Anything turning on clinical, legal or financial judgement, and anything where somebody is distressed. Those should reach a person quickly.
The agent should recognise the situation and route rather than attempt it. Diagnosis, advice on a live matter, a complaint that has escalated, and anybody upset are all straightforward handoffs.
Agreeing that list is part of the build. Being specific about what the agent will not do is more useful than claiming it handles everything.
How do I test it before trusting it with real customers?
Ring it yourself, then have your most impatient colleague ring it. Keep it on overflow or after-hours only until you are satisfied.
A staged rollout is the honest way in. Conditional forwarding means it only picks up calls that were going to voicemail anyway, so the downside is capped at what you were already losing.
We also read real transcripts with you in the first weeks. Nothing exposes a weak agent faster than seeing what it actually said.
How would I know if it is losing bookings a human would have won?
By reading transcripts and comparing outcomes, not by watching a dashboard of answered calls.
Answer rate always looks good, because the agent picks up everything. The number worth watching is calls that ended with no booking and no callback task, and those transcripts tell you whether it was the caller or the agent that gave up.
Better to look at that honestly in the first month than to discover it in the sixth.
Should it answer every call, or only the ones we miss?
Start with only the ones you miss. Move to full-time later if the transcripts justify it.
Overflow-only is the lower-risk setup and captures most of the value, because the missed calls were the leak. Your team keeps the relationships they already handle well.
Some businesses do move to answering everything, usually where the front desk is genuinely overloaded and consistency matters more than familiarity. That is a decision to make on evidence rather than at the start.
How does it tell an emergency from something that can wait?
By rules you approve, not its own judgement. You define what counts as urgent and what happens when it does.
For a plumber that might be water that will not stop; for a clinic, specific symptoms. The agent screens against your list, and anything matching goes straight to your on-call route instead of into a booking flow.
It errs towards escalation. A false alarm costs a phone call. The opposite mistake costs considerably more.