Insights · Honest · By Muzamil Hasan · 7 min read

What automation cannot do yet: the tasks to keep human

Automation is good at anything with a fixed rule: the same input, the same steps, the same output, every time. It is bad at anything that needs a judgment call, a read of someone's tone, or a decision nobody wrote down in advance. That is the line, and it holds up across almost every small business task we have looked at.

In practice, that means an angry customer, a real negotiation, an emergency, or any moment where the "right" answer depends on context a machine cannot see. Those stay with a person. Everything routine and repeatable around them is fair game. The rest of this article walks through why, with the specific situations where handing off to automation backfires.

What "automation" and "judgment call" mean here

Automation is software that follows a fixed set of steps without a person doing them by hand: send the reminder, log the payment, update the record. It includes simple rule-based scripts and the smarter systems often called AI agents, which can read a message and decide which of several known steps to take.

A judgment call is a decision with no single correct answer written down anywhere. It requires weighing things no database holds: how upset this specific customer sounds, whether this is the moment to bend a policy, whether a delay is a minor annoyance or a real emergency. A rule can cover the common cases. It cannot cover the case nobody anticipated, which is exactly what judgment calls are for.

Angry customers: automation makes it worse, not better

This is the clearest failure mode, and it shows up in the data. Review roundups of customer service software report that a large majority of people, commonly cited around 86%, still say they want a human once a complaint or a complicated issue is involved. Not for the easy stuff. For the moment something has already gone wrong.

The reason is not that the software is unfriendly. It is that an upset customer needs to feel heard before they need a solution, and a scripted response reads as dismissive even when the words are correct. A bot that says "I understand your frustration" after being told a wedding cake never showed up does not land as understanding. It lands as a form letter.

Our rule: automation can acknowledge the complaint fast, log it, and pull up the order history so a person has full context in ten seconds instead of five minutes. The apology, the judgment call on what to offer, and the actual conversation stay human.

Illustration: A human clay hand and the mascot gently shaking hands over a tiny table

Negotiation and anything with real stakes

Negotiation depends on reading a person in real time. What will they accept. What do they actually care about versus what they are just saying. When to hold firm and when a small concession saves the relationship. None of that lives in a transcript an AI agent can pattern-match against, because it changes with the specific person on the other end.

The same goes for anything where a wrong call is expensive or hard to undo: a refund above a certain size, a contract change, a payment plan for a customer who is behind. Automation can gather the facts and lay out the options. The decision about which option to pick, when real money or a real relationship is on the line, belongs to someone who can be held accountable for it.

Emergencies

An emergency needs judgment about severity, and it needs it immediately. "My water heater is making a weird noise" and "water is pouring through my ceiling" can arrive in nearly identical words. Only a person reliably tells them apart on the first read. A system tuned to route the first one into a normal queue will sometimes route the second one there too. That mistake costs a customer real damage, and it costs you the relationship.

For businesses where "emergency" is a real category, plumbers, HVAC shops, anyone with after-hours calls, that matters more than most. The honest setup keeps a live escalation path for anything that sounds urgent, checked by a person, not just logged by a bot for the morning.

Why this line does not move as fast as the marketing suggests

A 2024 MIT study estimated that automation was economically viable, meaning it was both technically possible and cheaper than a person, for only about 23% of the tasks it looked at. That number keeps moving as models improve, but it has moved slower than most vendor pitches imply, and the tasks it covers first are the well-defined ones, not the judgment-heavy ones this article is about.

That gap is not a flaw to route around. It is the actual shape of what works right now. A vendor who tells you AI handles everything is selling you the 2030 version of the product today, at today's price, with today's error rate.

What happens when a system hits one of these limits

The honest answer is that it should stop and ask, not guess. That means someone approves anything customer-facing or financial before it goes out, and there is a record of what the system did and did not do on its own. We cover exactly how that gate works, and what should never run without a person's tap, in should an AI ever talk to your customers without approval.

There is also a real liability question buried in all of this: a business was held responsible in court for what its own chatbot told a customer, wrong information and all. That case, and what it means for you, is in when the AI gets it wrong: who pays.

Where we come in

We build the automation for the routine half of this equation: the reminder that goes out, the record that updates, the summary that lands in your inbox instead of an hour of spreadsheet work. We do not build a system that argues with an angry customer or negotiates a contract. That system does not exist yet, whatever the pitch deck says.

If you are trying to figure out which of your own tasks belong on which side of that line, the automation profile tool walks through your actual week and sorts it for you. Some of what it finds will point back to you. That is the honest answer more often than any vendor wants to admit, us included when it is true.

See also what to automate first for the other half of this question: the tasks that are safe to hand off today.