When we started Trail, the obvious demo was the impressive one: give the agent a goal, watch it run, come back to a finished result. We built that version first. It lasted three weeks.
The teams testing it did not want a finished result they had not seen being made. They wanted to know what the agent was about to do, why, and how to stop it. One product lead put it plainly: “I don’t need it to be clever. I need to not be surprised.”
So we removed most of the autonomy. Trail now works in a loop that is almost dull to describe. It reads the goal the team wrote down. It proposes the next step, with the sentence in the goal that justifies it. It waits. When someone confirms, it carries that step through, shows the result, and proposes the next one. Nothing happens silently. If it cannot tell who owns something, it says so instead of guessing.
Three things followed from that decision.
First, trust arrived faster than capability. Teams let Trail take on more once they had seen it decline to act. The permission to be ambitious was earned by being predictable.
Second, the most valued message turned out to be “nothing to do.” A tool that can say there is no next step is a tool people believe when there is one.
Third, the routine parts are where the value is. Drafting the follow-up, gathering the three documents, checking the numbers match: these are the steps that disappear between “we decided” and “it happened.” Trail does not need to be brilliant to be useful there. It needs to be reliable.
We still believe the ambitious version is where this goes. But the path runs through boring, and we would rather walk it with twelve teams who trust us than sprint it alone.
