Result
What must be different when the task is done?
Goal to Proof is a lightweight closure contract for AI agents: turn authorized, non-trivial work into an observed outcome—not a plausible completion claim.
npx skills add aiopshwang/goal-to-proof
Agent Skills
Goal to Proof is an AI agent completion-verification skill. It prevents an agent from substituting a plan, placeholder, partial artifact, isolated check, or unobserved external action for the requested result.
A file, change, or record was created.
Syntax, unit test, citation, or schema is valid.
The connected workflow runs.
The user, device, render, account, or remote sees it.
Stop at the first layer that directly proves the request.
The fields stay internal when obvious and become visible only when ambiguity or authority could change the work.
What must be different when the task is done?
Who or what must be able to use or observe it?
What direct observation separates success from plausible failure?
What is authorized, excluded, or requires the user?
The user and agent have different jobs. Making the split explicit prevents both approval theater and unauthorized action.
The repository root is the plugin root. Platform packages point to the same canonical skills/goal-to-proof/SKILL.md.
Portable installer for supported skill hosts.
npx skills add aiopshwang/goal-to-proof
Add the repository catalog, then install the plugin.
codex plugin marketplace add aiopshwang/goal-to-proof
codex plugin add goal-to-proof@goal-to-proof
Install as a namespaced managed plugin.
claude plugin marketplace add aiopshwang/goal-to-proof
claude plugin install goal-to-proof@goal-to-proof
$goal-to-proof in Codex. In a Claude Code managed plugin install, use /goal-to-proof:goal-to-proof; a standalone skill may be exposed as /goal-to-proof.Package presence is not a blanket compatibility claim. See validation evidence and claim levels for what each check proves.
The initial contract came from aggregate analysis of prior real working sessions. Repeated behaviors—not private content or corpus metadata—became the public principles.
Intent alignment, scope control, authorized autonomy, end-to-end proof, durable checkpoints, and evidence-first reporting.
Raw conversations, personal names, secrets, one-off preferences, private content, and session transcripts.
The corpus counts explain the source of the design. They are not a benchmark or proof of performance.
No. Goal to Proof makes the current authorization boundary more explicit. Publication, spending, disclosure, irreversible action, and material scope expansion still require user authority.
No. Tests can be excellent proof for a software claim, but the requested boundary might be a rendered document, a public release, an account state, or a research conclusion. Proof must match the claim.
No. Tools, permissions, environments, and the task itself can block completion. The skill requires the agent to report the verified layer and name what remains unverified.
No. It is an AI-agent completion skill. This site uses conventional search fundamentals and clear, source-linked answers, but the project makes no ranking claim. See the SEO/GEO answer.