Skip to content
Customer service Human oversight

Can AI safely reply to your customer emails?

Yes, to some of them, and the interesting question is which. A way of splitting an inbox that keeps the speed without risking the relationship.

· 7 min read

The fear behind this question is specific and reasonable: a machine sends something confident and wrong to a customer, and you find out when they escalate.

That failure is avoidable, but not by making the model better. It is avoided by being deliberate about which messages it is allowed to answer at all.

Split the inbox before you automate it

Most shared inboxes contain four kinds of message, and they have completely different risk profiles.

Questions with a documented answer. Opening hours, order status, how to reset something, what your returns policy says. There is a right answer, it exists in writing, and the cost of getting it wrong is low. These are safe to answer automatically, drawing on your own approved wording rather than the model’s general knowledge.

Requests that need an action. Change a delivery address, cancel a booking, add a user. The reply is easy; the action behind it is the risk. Automate the understanding and the drafting, and let the workflow perform the action only where it is reversible and logged.

Messages that need judgement. A complaint, an unusual request, a negotiation, anything with a legal or safety flavour. These should never be answered automatically, and the useful thing automation can do is recognise them quickly and get them to the right person with a summary — which is often a bigger improvement than an automatic reply would have been.

Messages the classifier is unsure about. This is a category in its own right and it needs a route. Uncertainty is information; a system that has no way to express it will simply guess.

What “safely” actually requires

  • Your wording, not the model’s. Replies drawn from your approved answers, not generated freehand. This is the difference between a system that sounds like you and one that invents a policy.
  • A confidence threshold with a real consequence. Below the line it does not send. Not “sends with a caveat” — does not send.
  • A visible record. What arrived, how it was classified, what was sent, and on what basis.
  • An obvious way for the customer to reach a person. Every automated reply should make that easy rather than burying it.
  • Sampling. Someone reads a sample of automated replies every week. This costs very little and catches drift long before a customer does.

Start with drafts, not sends

The lowest-risk way to begin is to have the workflow draft and a person send. You get most of the time saving immediately, and — more usefully — you get a few weeks of evidence about how good the drafts actually are on your real inbox before anything goes out unsupervised.

After a month you will know precisely which categories are safe to release. That is a much better basis for the decision than anybody’s assurance, including ours.

Bring us your process.

In a free 30-minute discovery call we will look at how it works today, where the manual effort sits, what is worth automating, what is not, and where people should stay in the loop. No technical brief, no budget, no obligation.