Lyrion has joined the Open Secure AI Alliance, alongside

NVIDIAMicrosoftPalantirIBMAdobe
Read the announcementRead more

Escalation done properly

Green Fern

Handing a thread from an agent to a person is a design problem, and it is the one that quietly decides whether automation saves time or just moves the pile. Done badly, escalation means a human opens a mystery and starts from scratch. Done well, it means they open a decision that is already framed.

The whole value is in what travels with the handover.

The bad handover

The bad version drops a raw thread into a person's lap with a tag that says “needs human”. Now the person has to read the whole history, work out what the agent already tried, guess why it stopped, and reconstruct the context the agent had and threw away.

This is worse than no automation, because you have added a layer and kept all the work. People notice, and they stop trusting the escalations, and then they start reading everything again just in case.

What has to travel

A good escalation carries a short summary of what the customer wants, what the agent already did, the reason it stopped, the relevant history, and a clear recommendation with the decision the person actually has to make. The person should be able to act in a minute, not reconstruct in ten.

The recommendation matters. “Here is what I would do and why, but this is above my limit” is a handover. “I do not know, you deal with it” is a dump. The first respects the person's time. The second wastes it.

Escalation is not failure

It is tempting to treat every escalation as a miss, and to push the automation rate up by escalating less. That is the wrong metric to chase. The right escalation rate is however many threads genuinely need judgement, no fewer.

An agent that escalates the hard seven out of forty-one and handles the rest cleanly is doing exactly its job. An agent that escalates two because it was tuned to look good is setting a person up to catch its mistakes later.

Close the loop back

When the person decides, that decision should flow back: logged against the account, and where it is a repeatable pattern, turned into a rule the agent can follow next time. Escalation is not just a safety valve. It is how the system learns where the real line is.

Get this right and your escalations get rarer over time, not because the agent hides them, but because yesterday's judgement calls become today's known cases.

The part people underestimate

Most of the value is not the reply itself. It is that the context is gathered, the history is attached, and the decision a person still makes arrives already framed. The queue stops being a pile of cold threads and becomes a short list of real decisions.