/g The Contract Codex ← 返回动态 Catalog

Voice E2E Goal Contract

Run a long-horizon Codex goal against a TypeScript voice system using a reading list, working rules, concrete done-when criteria, and anti-pattern fences.

难度:intermediate 分类:testing 来源:Tecton & Tide Codex goal run

完整 Prompt(可直接复制)

/goal
GOAL:
Complete Voice E2E Goal Contract for a project with failing or missing verification gates: Run a long-horizon Codex goal against a TypeScript voice system using a reading list, working rules, concrete done-when criteria, and anti-pattern fences.

CONTEXT:
- Before editing, read the nearest AGENTS.md/CLAUDE.md, current issue or PLAN.md, and any failing logs already in the repo.
- Inspect test suites, lint config, CI logs, coverage reports, and failing output.
- Establish a baseline by running or locating evidence for: `four target end-to-end voice scenarios pass, transcript review shows no prompt loops, and unavailable metrics are documented honestly`.

CONSTRAINTS:
- Keep the scope limited to this goal; do not expand into unrelated cleanup.
- Do not weaken tests, delete assertions, or mask errors to make verification pass.
- Respect the repository's AGENTS.md/CLAUDE.md instructions and existing patterns.
- Do not weaken lint, typecheck, or test rules to create a green result.
- Fix production or fixture causes before changing expectations.

DONE WHEN:
- The implementation or documentation directly satisfies: Run a long-horizon Codex goal against a TypeScript voice system using a reading list, working rules, concrete done-when criteria, and anti-pattern fences.
- The verification command or evidence path succeeds: `four target end-to-end voice scenarios pass, transcript review shows no prompt loops, and unavailable metrics are documented honestly`.
- The final diff is scoped to the relevant files and has no unrelated formatting churn.

VERIFY:
- Run `four target end-to-end voice scenarios pass, transcript review shows no prompt loops, and unavailable metrics are documented honestly` or the closest repo-local equivalent if the exact command is not available.
- Capture before/after evidence for the behavior, metric, report, or artifact involved.
- If verification cannot run locally, stop and report the missing dependency instead of guessing success.

OUTPUT:
- Summarize changed files, key decisions, verification output, and remaining risks.
- Include any follow-up that is required for production rollout or human review.

STOP RULES:
- Pause if secrets, production access, stakeholder decisions, or destructive data operations are required.
- Pause after three failed fix attempts on the same symptom and challenge the root-cause hypothesis.
- Do not mark the goal complete until the current repository state has been audited against DONE WHEN.

来源与证据

原始来源: Tecton & Tide Codex goal run

证据摘要: All four target end-to-end voice scenarios passed verification; source: Tecton & Tide Codex goal run; type: third-party-review; verification: four target end-to-end voice scenarios pass, transcript review shows no prompt loops, and unavailable metrics are documented honestly