compare

Every tool promises discipline.

Spec kits, skill packs, virtual teams, task graphs — each hands your agent a process to follow. One question separates them: when the model skips a step at hour six, what stops it? For most tools the answer is nothing — the process is a recommendation. team.management makes it a gate. Every number below carries a date, and every source is linked.

That is not a criticism of any of them. A well-written instruction is genuinely useful, and most of the time the model does follow it. The question is what happens the rest of the time.

Anything the model reads is advice. It sits in the context window alongside everything else, and it competes with everything else. Late in a long session, with the window full and the goal three steps away, advice is the first thing to go. Nothing announces that it happened.

A gate is different because it does not live in the context window at all. It runs in the harness. Some gates answer a tool call before it happens — the edit tools stay locked until the work has been agreed. Others refuse to let the job reach its next step until a check has actually run. Either way the model does not get a say.

That distinction is the axis every page below is measured on, and it cuts both ways — it is also the limit. A gate binds the agent, not the person running it. You can always change your own config. What it removes is the silent kind of failure, where a step was skipped and the summary said otherwise.

The pages come in three families. Workflow tools are the ones you would run instead of this, or alongside it — most compose better than they compete. Models and subscriptions are about what you pay and what you get for it, with a date on every number, because those numbers move month to month. Process and verification compares methods rather than products: what actually holds when an agent goes off-script, and what checks the work it claims to have done.