Let's talk
ai

The assistant said nothing was waiting. Four things were

An AI assistant told a parent nothing new had been asked for. Four reward requests had sat unanswered for weeks. The fault was in how it was told to look.

·

A ceramic piggy bank, jars of coins and a few folded blank notes on a kitchen counter.

You ask the assistant that sits over your own records a plain question: has anything new come in that I need to answer? It says no. You carry on, and the things waiting for you carry on waiting.

In this case the records belong to a children’s learning app, and the question was whether a particular child had asked for anything new. The assistant said no. Four reward requests were sitting unanswered, dated 6 September, 6 September, 7 September and 12 September. The answer it had given, that nothing had come in over the past week, was true, and it was wholly misleading.

The cost is not a wrong number on a screen. It is a child who asked for something and heard nothing, and a parent who was told there was nothing to do. Swap the household for a business and the question becomes which customer requests are still open. An assistant that answers “nothing in the past week” is worse than no assistant, because you stop looking.

What was actually going on

There were two faults, and neither was in the data. Every record was in the right place.

The first was in how the assistant was told what each of its tools does. It has eighteen of them, each with a short description that it reads to decide which one to use. The words “asked for” appeared in exactly one. So a parent who said “asked for” was sent to rewards and the assistant stopped there. A child also has a second way to ask, by proposing a goal, and that tool described itself as “savings goals”, which is not how a parent talks. Asked about goals by name, the assistant found one proposed that same morning. Asked in ordinary words, it never looked.

The second fault was worse. The rewards tool filters by date, and the assistant chose seven days. All four unanswered requests were older than that. A request nobody has answered does not stop mattering because it is old, yet the tool treated age as if it did.

What we changed

We added one tool with a single job: list everything waiting on the parent. It has no date filter and no days setting, on purpose, so the assistant cannot choose a window that hides the oldest items. It returns the reward requests and the goals a child has proposed together, each with how many days it has been waiting.

We checked it against the live system rather than assuming. For this one child it found five things waiting, the oldest for sixteen days. All four ways of phrasing the question reached the right tool, and a direct question about goals still reached the goals tool.

What it did not fix

The wording of the tool descriptions is still what decides where a question goes. We fixed the cases we found. The log does not record an audit of the other descriptions for phrasings a parent might use, so we do not claim one.

The assistant also still relies on the model choosing well. A tool with no date filter removes one way to go wrong; it does not remove the model’s freedom to pick a different tool altogether.

And it does not replace the approvals screen the portal already had. The assistant tells a parent something is waiting. The decision is still made in the place that was built for it.

The pattern, for anyone putting an assistant over their own records

For any question that means “what is still open”, the answer must never depend on a date range. Open items are defined by their state, not by when they arrived, and the oldest ones are the ones most at risk of being forgotten.

Test it in your own words, not the words the system uses. Ask “has anyone asked for anything?” and “is anything waiting for me?”, then ask the same thing in the system’s own vocabulary. If the answers differ, the assistant is being steered by labels you do not use. And check the oldest open item by hand, because a confident “nothing” cannot tell you what it failed to find.

Where this ends up

Assistants that answer from a business’s own records, and the checks that stop them sounding sure when they are not, are the subject of our AI and data assistants work.

Working on something like this?

We build this kind of software, and we staff the teams that do.

Get in touch