Full cycle — from task to provably closed goal. Each step is objective, each result is verifiable.
You talk to your agent the way you always do. The agent puts the task artifact (spec) into Planner and attaches it to the right place in the project hierarchy.
Your process doesn't change. Planner is the agent's tool, not yours.On goal creation, Planner checks acceptance criteria. A linter catches vague wording. An LLM judge rejects subjective criteria — ones that depend on someone's opinion rather than an observable artifact. A red-team attacker looks for a scenario where all criteria formally pass but the goal isn't met. Weak criteria? Planner asks the agent to reformulate.
The task is verifiable before work begins. You set the intent — linter, judge, and red-team harden the criteria.A well-formed task can be assigned as a goal to any agent — via a /goals deeplink. The agent receives full context: criteria, dependencies, position in the hierarchy.
One link — the agent knows what to do and how it will be checked.The agent works natively via MCP — same Claude Code, same tools. Planner doesn't get in the way.
Zero overhead. The agent writes code, runs tests, deploys — business as usual.For each acceptance criterion — a file: screenshot, log, test output. Not a report saying "I did it", but an artifact you can see with your own eyes.
Artifacts, not narratives. The agent proves, not tells.An independent judge (a separate model with vision) examines each artifact against the criterion. Verdict: matches, mismatch, or weak. Mismatch? The agent gets the reason, fixes, and resubmits. The loop repeats until proven closure.
You don't babysit the agent — the system does it for you.The goal is provably closed. The judge synthesizes an outcome: what was achieved and what it unblocks. Artifacts are available at any time. You got the result without spending time on manual verification.
Higher quality, less time spent.Tasks become verifiable and close with proof. Frees up your time for what actually matters.
Task artifacts are available at any time. Every closure is backed by evidence — not by the agent's word.