2026-08-16
Leadership Is a Manual Control Loop
Working with an agent can feel less like coding and more like leading a capable new hire. That is a useful observation. It is also a warning that the operating system still lives in somebody's head.
The leadership metaphor fits
You rarely get the best result by throwing a vague request at an agent and waiting. You explain the destination, share the context that changes the decision, define what good looks like, inspect what comes back, and correct the route. Anyone who has led people recognizes the rhythm.
The comparison explains why coding skill alone does not guarantee good agent-run work. The hard part moves away from producing each line and toward making intent legible. Judgment shows up in the brief, in the boundary around the work, and in the decision to accept or reject the result.
But metaphors can hide costs. If every run depends on one person remembering how to brief, monitor, review, and redirect the agent, the company has not built a new production system. It has assigned that person a very fast report with no durable management structure.
Good management does not scale by conversation
A skilled operator can recover a weak request through follow-up. They notice the missing assumption, add context at the right moment, challenge a plausible answer, and ask for another pass. The finished result may be excellent. The route that produced it is still trapped inside a conversation.
That becomes a liability as the work multiplies. Ten agents do not give one leader ten times the capacity when each one needs personal clarification, live supervision, and a custom review. The bottleneck simply moves from typing code to holding every production decision in working memory.
Human availability is a poor control surface for repeatable work. People get interrupted. Their standards drift. They explain the same rule differently on Tuesday than they did on Monday. An agent can execute around the clock, but that advantage disappears when the machinery must wait for a person to remember what happens next.
Turn the leadership loop into machinery
A software factory takes the useful parts of leadership and makes them executable. The destination becomes a specification with named outcomes and constraints. Delegation becomes routing based on the kind of work, the authority required, and the evidence expected. Review becomes a gate that can actually stop the line.
Context stops being a thoughtful preamble someone writes from scratch. The factory retrieves the relevant rules, repository state, prior decisions, and production evidence for the job. It gives the agent enough authority to act and no more. When a run fails, the failure returns to the system as a condition the next attempt must survive.
That is the difference between prompting and operating. A prompt asks a worker to behave well. A factory makes the route carry the standards, preserve the record, and reject work that cannot prove itself outside the worker's own explanation.
Judgment moves up, and jobs disappear below it
Mechanizing the loop does not eliminate human judgment. It changes where judgment earns its place. People choose which outcomes matter, which consequences are acceptable, and which standards deserve to govern every future run. They inspect exceptions that reveal a rule is weak, then improve the rule instead of manually rescuing the same class of work forever.
The repeatable coordination underneath those decisions will not remain a human job. Translating a clear request into tasks, checking routine conformance, moving status, requesting the predictable revision, and assembling the record are exactly the kinds of work factories retain and execute more consistently than a rotating set of people can.
Our prediction is that organizations will need fewer people whose value is being the live control loop. The people who move their judgment into specifications, gates, routing, and recovery will compound it across every run. The people who keep performing the loop by hand will compete with machinery that never forgets the lesson.
Do not manage the same run twice
Treat the leadership feeling as a diagnostic. Every time you clarify an instruction, ask why the needed context was missing. Every time you reject an answer, identify the claim the gate should have tested. Every time you intervene, decide whether the authority boundary, route, or recovery policy was incomplete.
Then put the answer back into the factory. The next agent should inherit the stronger specification. The next review should receive the evidence automatically. The next failure should stop at the boundary you just discovered instead of requiring another person to recognize it in time.
Leading an agent well can produce one good result. Building the leadership into the system produces an asset. The conversation is where you discover the loop. The factory is where you make sure nobody has to carry it by hand again.
In response to Working with AI Feels More Like Leadership Than Coding by Allen.