Repository navigation
docs: fix stale and broken agent instructions from prompt audit - #842
Conversation
A prompt audit of the Claude Code configuration found broken references, stale release facts, and instruction files that contradict each other. - AGENTS.md: agent prompts live in .claude/agents/, publishing uses Trusted Publishing, the module map points to ARCHITECTURE.md, and the staleness check reads the root ARCHITECTURE.md. - sdk-dev.md: fix the GOLDEN_PRINCIPLES.md link and the SDK symbol names; add subagent frontmatter (also to pm.md). - pm.md, marketing.md: defer work order to #729 instead of the old phase plan. - dashboard-dev.md: drop non-public plan prices. - next-ticket skill: drop the AG-06 hold (#735 and #767 are closed). - CLAUDE.md, AGENTS.md: remove shouted capitals; rules are unchanged. Proof and a re-check script are in proof/prompt-audit-cleanup/. Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
Codex Review SummaryThis comment shows the latest Codex review activity on this pull request.
ℹ️ About Codex in GitHubYour team has set up Codex to review pull requests in this repo. Reviews are triggered when you
Codex reacts with 👀 while any review is running, comments if it has suggestions, and reacts with 👍 once all reviews finish with no findings. |
Bugbot couldn't run - usage limit reachedBugbot is counted against Cursor usage for this user or team, and this run hit a usage or spend limit. A user or team admin can review and increase usage limits in the Cursor dashboard. (requestId: 609dec97-418a-43a3-b0e0-ef840265e399) |
🤖 Claude reviewReviewThis is a docs/config-only cleanup with no SDK code changes. The proof chain is complete and all 24 checks passed. A few items worth noting: Non-blocking inconsistencies
The agent-file paths were correctly updated to
Minor concern in
|
|
Replies to the Claude review above. No code change; reasons per point:
🤖 Addressed by Claude Code |
What
Fixes stale and broken agent instructions found by a prompt audit of the Claude Code configuration (target: Claude Opus 5.5). Docs and agent config only; no SDK code changes.
Broken references, now corrected
AGENTS.md: agent prompt paths pointed to.Codex/agents/, which does not exist. They now point to.claude/agents/.AGENTS.md: the release comment namedPYPI_TOKEN.publish.ymluses Trusted Publishing (OIDC).AGENTS.md: replaced the old module graph and module table with a pointer toARCHITECTURE.md. The table listed 3 CLI subcommands; there are 13..claude/agents/sdk-dev.md: fixed theGOLDEN_PRINCIPLES.mdlink,RateLimitExceeded→RetryLimitExceeded, andMODEL_PRICES→DEFAULT_PRICE_TABLEinprice_table.py..claude/agents/pm.md: removedv1.2.6as the latest release (state lives inmemory/state.md)..agents/skills/next-ticket/SKILL.md: removed the AG-06 hold (AG-06: Add a supported OpenAI Responses and Agents SDK integration #735 and AG-06 remains held: #735 was closed by a release-prep reference #767 are closed).Conflicts between instruction files, older text rewritten to match newer
pm.mdandmarketing.md: the 3-phase plan and Phase 1 launch table now defer to AgentGuard weekly growth plan: verified limits, portable evidence, and adoption #729 (perSKILL.md:20andpm.md:49).marketing.md: dropped "expand to full observability" (CLAUDE.md:74avoids it). Wedge and audience now matchmemory/distribution.mdanddocs/enforcement-boundary.md.dashboard-dev.md: removed non-public plan prices (marketing.md:51,site/compare.html:146).AGENTS.md: the staleness check reads the rootARCHITECTURE.md, asCLAUDE.mddoes.Roster
name/descriptionfrontmatter tosdk-dev.mdandpm.md, so Claude Code registers them as subagents (CLAUDE.mdlists them as Claude-side assets).Wording
CLAUDE.md,AGENTS.md) and the contract heading. The rules are unchanged.Follow-ups added to
ops/FOLLOWUP.md: the root report files, and whetherops-cadence.ymland the PR template should track the rootARCHITECTURE.md.Proof
Saved under
proof/prompt-audit-cleanup/.python scripts/sdk_preflight.pypython scripts/ci_tools_requirements_guard.pypython scripts/review_readiness_guard.pyruff check(Makefilelintpaths)pytest sdk/tests/test_architecture.pybandit -r sdk/agentguard/ -s B101,B110,B112,B311 -qpython scripts/sdk_release_guard.pypython scripts/generate_pypi_readme.py --checkpytest sdk/tests/ --cov=agentguard --cov-fail-under=80python proof/prompt-audit-cleanup/verify.pymakeis not installed on this Windows host, so each target ran directly.make mcpwas not run: nomcp-server/file changed. CI runs it.showwork:
prompt-audit-cleanupclosed checks-only (24/24 green).prompt-audit-acceptancedeclaresverify.pyas its acceptance check and closes on the outcome.The release-guard markers in
AGENTS.md,CLAUDE.mdandsdk-dev.mdare unchanged.Not in this PR
~/.claudeconflicts (auto-merge vs the comment loop,pnpm test:e2e). They are outside the repo and reported to the owner.QA_REPORT.md,WORK_PLAN.md,RESEARCH.mdout of the root.🤖 Generated with Claude Code
Note
Low Risk
Documentation and agent-config only; changes how agents are instructed, not product behavior, auth, or data paths.
Overview
Aligns Claude/agent instruction files with current repo facts so automated agents stop following stale paths, wrong SDK names, and superseded launch plans. No runtime SDK or dashboard code changes.
Reference and roster fixes:
AGENTS.mdand related docs now point at.claude/agents/, Trusted Publishing instead ofPYPI_TOKEN, and rootARCHITECTURE.mdfor staleness.sdk-dev.mdcorrects guard/price guidance (RetryLimitExceeded,DEFAULT_PRICE_TABLE) and theGOLDEN_PRINCIPLES.mdlink.pm.mdandmarketing.mddrop the old 3-phase launch tables and defer ordering to GitHub #729; marketing positioning tightens to runtime guardrails (not “full observability”).dashboard-dev.mdand the next-ticket skill remove non-public prices and the closed AG-06 hold.sdk-devandpmsubagents getname/descriptionfrontmatter.Proof and follow-ups: Adds
proof/prompt-audit-cleanup/(saved check outputs plusverify.pyto re-assert fixes), showwork session/claim receipts, a smallCLAUDE.mdreview-loop wording tweak, and prompt-audit leftovers logged inops/FOLLOWUP.md.Reviewed by Cursor Bugbot for commit 1bb18b4. Bugbot is set up for automated code reviews on this repo. Configure here.