
An AI-ready project brief: the template that stops endless prompting
An AI project brief connects outcomes, scope, sources, constraints, and validation. Use this template before assigning work to an agent.
Your AI agent says the job is done. Learn how to check the actual result, find supporting evidence and catch missing work before you sign it off.
Last reviewed on August 30, 2026

“Done.” The new field appears on the form. But leave it empty and submission fails. The client asked for an optional field: the task is not finished.
To verify an AI agent’s work, start with the observable result, not its closing message.
What was requested?
Return to the original criteria.
What actually works?
Try the relevant user journey.
What is still unchecked?
Name the limits, then decide.
Match the evidence to the promise. A screenshot may be enough to check alignment. It cannot prove that a form works.
The agent says
“The form is ready”
The new field appears in a screenshot.
Useful evidence
Both submissions succeed
With and without a value in the optional field. The data reaches the right place.
The agent says
“The page is updated”
A file was saved on its machine.
Useful evidence
The right version is accessible
Open the relevant URL. Distinguish the preview from the public site.
Anthropic draws this distinction between an execution trace and the final outcome. A tool call proves an attempt; its effect still needs checking.
Return to our fictional client email: replace a PDF and add an optional Company field, keeping the price and logo unchanged.
Missing access? Record the limit. Without access to the system receiving registrations, you can check the confirmation screen but not the final stored record.
Ask for three short sections. This is a fictional example, not a test result produced for this article.
For a real assignment, include the preview URL and test references. “Verified” without accessible evidence leaves you to reconstruct the investigation.
Ask for a verifiable handover
Compare your result with the original request, including anything that should remain unchanged.
Verified: list the checks actually performed and links to the evidence.
Not checked: state what still needs testing and which access is missing.
Pending: list the decisions or permissions needed.
Do not invent evidence. An attempt is not a success; a preview is not a published change. Take no external action without the required authorisation.
The criteria are met
Accept the work
Evidence is accessible and any remaining limits do not prevent acceptance of the agreed scope.
A result is missing
Request a correction
Name the defect: “Submission fails without a company name.” That is more useful than “check again.”
If an essential check is impossible, leave the task pending. A second AI can help review it, but its agreement does not replace a successful test.
In Stellary, keep acceptance criteria in card checklists and evidence with the task. Define permissions using the AI agents guide.
For recurring assignments, keep a few cases to rerun after changing models: an ordinary request, a missing attachment and contradictory instructions. One successful demonstration is not enough to judge reliability.
Match the review to the risk: read through a rewrite, test a functional journey, and explicitly approve messages or publications before execution.
Sometimes, for a visual state. For delivery, stored records or forms, check the behaviour and the result at the receiving end.
It should state what is verified and what remains unknown. A missing essential check prevents treating the task as fully approved.

An AI project brief connects outcomes, scope, sources, constraints, and validation. Use this template before assigning work to an agent.

Task boundaries, context, worktrees, contracts, and integration: a method for running multiple AI coding agents in parallel without multiplying conflicts.

A simple guide to choosing local, cloud, or hybrid AI based on your documents, budget, internet connection, and everyday needs.

Fable 5.1 or Opus 5? Choose by mission difficulty and cost per successful task — not by putting Anthropic’s strongest model everywhere.
Stellary brings together your board, docs, and AI agents in one command center.