Repository navigation
feat: apply the Opus 5.5 usage guide across repo instructions and plugins - #4352
Conversation
Adds reference/model-adaptation/opus-5-5.md sourced from the Opus 5.5 usage guide and the official prompting and model-config docs, routes it from fable-5 meta-rule 3, and keeps the Opus 5 and Opus 4.8 chapters as fallback targets for flagged requests. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_015zXN1hq49epZqs3APTdCz8
…gins Design and artifact skills now name specific styles to leave out and extend the list when the user dislikes a choice; adhd:shape takes the next no-input step in the same message; playwright reads the screenshot itself for visual questions; dometrain grounding marks what no lesson confirmed; ai-briefing's report leads with what waits on the user. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_015zXN1hq49epZqs3APTdCz8
…lugins Review findings carry file and line, why it is wrong, and how to show it fails; PR mode leads with merge-blocking findings; fan-out audits check each subagent's evidence before accepting it; doc-drift-detector checks for self-contradicting numbers, dates, and names; long-run reports lead with what waits on the user; knowledge queues the Opus 5.5 docs. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_015zXN1hq49epZqs3APTdCz8
…topping point Audit and research reports in bugs, codebase-health, mutation-testing, discovery, architecture, and machine-health open with what waits on the user. Dispatch briefs in codebase-health, batch-simplify, coupling, mutation-testing, review:fanout, course-digest, and architecture:improve say when the subagent is done and when to return early. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_015zXN1hq49epZqs3APTdCz8
…tion audits Root AGENTS.md gains a stop and keep-going rule. audit-instructions detects think-carefully steers and reasoning-reproduction requests on Opus 5.5 targets, vague design steers, and misplaced settled-answer instructions; audit-prompting-postures checks finish lines, named stops, task files, fan-out evidence checks, and report shape; write-for-agents teaches the same when authoring. Live docs now resolve `opus` to Opus 5.5. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_015zXN1hq49epZqs3APTdCz8
…ugins Autonomous dispatch, lane stop gates, and session-flow no longer treat a status summary or an offer to continue as a stop; spawn specs name done and when to stop; fan-outs check worker evidence and end in one table; run reports lead with what waits on the user; PR prep lists only merge-blocking findings with file, line, and repro; known-issues covers the flagged-message model switch. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_015zXN1hq49epZqs3APTdCz8
Adds the per-item provenance record, and closes the last two gaps: the worker and merge lane launch prompts forbid ending a turn on a status summary and lead the cycle report with escalations, and the lane-stop alert names waiting on the operator. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_015zXN1hq49epZqs3APTdCz8
PR prep leads with merge-blocking findings instead of dropping the rest; implement-dispatch's autonomous keep-going defers to the resident-vs-clear conditions; the AGENTS.md report headings yield to a skill's own report shape; the loop-lane known gap now covers Opus 5.5's classifiers; the ledger attributes G2.1 and G3.1 surfaces correctly and maps the guide's summary items. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_015zXN1hq49epZqs3APTdCz8
Its description is now 1007 codepoints, under the 1024 maximum. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_015zXN1hq49epZqs3APTdCz8
630f2b9 to
f43d49e
Compare
Codex Review SummaryThis comment shows the latest Codex review activity on this pull request.
ℹ️ About Codex in GitHubYour team has set up Codex to review pull requests in this repo. Reviews are triggered when you
Codex reacts with 👀 while any review is running, comments if it has suggestions, and reacts with 👍 once all reviews finish with no findings. |
|
Claude finished @kyle-sexton's task in 3m 12s —— View job
|
|
Claude finished @kyle-sexton's task in 1m 17s —— View job Security review — no issues found
Reviewed the full diff against Skip-gate assessment: this PR has no security-relevant surface. It is almost entirely documentation and skill/agent instruction text (
I checked the full diff for injection, secret-handling, and command-execution patterns ( No GitHub Actions workflow files are touched by this PR, so there is nothing in zizmor's advisory lane to defer. Conclusion: no CRITICAL, IMPORTANT, or SUGGESTION security findings. Per this lane's skip gate, no inline comments were posted. |
|
Last security-reviewed head: |
|
Claude has reviewed this PR 1 time. The lane skips further automatic reviews after 5; deleting this comment resets the count. |
There was a problem hiding this comment.
💡 Codex Review
Here are some automated review suggestions for this pull request.
Reviewed commit: f43d49efc2
ℹ️ About Codex in GitHub
Your team has set up Codex to review pull requests in this repo. Reviews are triggered when you
- Open a pull request for review
- Mark a draft as ready
- Comment "@codex review".
If Codex has suggestions, it will comment; otherwise it will react with 👍.
Codex can also answer questions or update the PR. Try commenting "@codex address that feedback".
…arget model write-for-agents removed think-carefully lines for every target, while audit-instructions I8-f scopes that advice to Opus 5.5 because the model-agnostic guide still recommends thinking steers. The skill now defers to I8-f for the model the text will run on. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_015zXN1hq49epZqs3APTdCz8
No related issue: initiated directly from a session request to apply the Opus 5.5 usage guide across the repository; follow-ups are filed as #4346, #4347, #4348, #4349, and #4351.
Summary
Applies "Getting the most out of Opus 5.5 in Claude and Claude Code" (claude.dev, 2026-09-22) to this repository's own instructions and to every plugin's skills, agents, hook text, and prompts. The branch also adds an Opus 5.5 adaptation chapter to
playbooks.docs/upstream/opus-5-5-usage-guide.mdrecords each guide item with its verdict and the files that carry it.Where the guide's advice is Opus-5.5-specific, it lives only in
plugins/playbooks/reference/model-adaptation/opus-5-5.md. Shared skills carry only the model-agnostic items, because Fable 5.1, Sonnet, and Haiku also read them.Fix
AGENTS.mdnow says when to keep going and when to stop and ask. It asks the agent to stop before destructive actions or anything outside this checkout, keep permission prompts on, keep a task file on long runs, and end with "Blocked on me / Changed / Found" unless a skill defines its own report shape. The loop-lane launch prompts forbid ending a turn on a status summary.claude-config:audit-instructionsdetects:audit-prompting-posturesP1, P6, P9, and the new P11 check for finish lines, named stops, task files, fan-out evidence checks, and report order.docs-hygiene:write-for-agentsteaches the same items at authoring time.opus-5-5.mdchapter, and fable-5 meta-rule 3 now routes it and re-resolves the chapter after a flagged-message fallback.opus-5.mdandopus-4-8.mdstay, because flagged Fable and Opus 5.5 requests fall back to Opus 5 (biology) and Opus 4.8 (cybersecurity). Each of those chapters now says so at its top, and playbooks: retire the Opus 5 and Opus 4.8 chapters once they stop being fallback targets #4349 tracks retiring them.doc-drift-detectorchecks documents for contradictions within themselves.opusresolves to Opus 5.5, and known-issues covers the flagged-message model switch.Every changed plugin is version-bumped and has a new CHANGELOG entry.
Decisions made by the owner:
review:code-reviewkeeps its "block or flag" bar.Verification
scripts/validate-plugins.sh: passes, including the strict catalog manifest.scripts/check-changelog-parity.sh --check-bump origin/mainand--check-preserved origin/mainboth pass: 37 changelogs were changed and 2612 headings were compared.instruction-scan.test.shpasses 97/97 andemit-findings.test.shpasses 119/119.lane-stop-gate.test.shpasses 103 of 104. The one failure is case 52, CR handling, which fails the same way at base on Windows.check-skill.shpasses on the changed skills except for failures that already exist onmain:mutation-testing:audit,architecture:improve, andcode-tidying:audit-dead-codeare over the length limit (Three skill descriptions exceed the 1024-codepoint limit and fail check-skill #4351). The branch changes none of them.claude-configaudit-automation-gaps (findings-state.test.sh,inventory.test.sh),claude-memoryaudit (instruction-load-stats.test.sh,nested-agents-check.test.sh),claude-opschangelog (changelog-status.test.sh), andrepo-hygieneclean (git-branch-audit.test.sh). None of these tests read the files this branch edits.findings-state.test.shandnested-agents-check.test.shwere confirmed to exit 1 onmaintoo.A fresh-context verifier checked the ledger against the diff. It found five problems:
All five are fixed in
4a0c6e7dc. The verifier found no weakened gate, no new request to show reasoning, no broken ledger link, and no unbumped plugin.markdownlint was not run locally because there is no
node_moduleshere. CI runs it.Related
🤖 Generated with Claude Code
https://claude.ai/code/session_015zXN1hq49epZqs3APTdCz8