Skip to content

feat(skills): add self-audit — four-dimension reasoning quality gate before delivery - #1361

Closed
YuhaoLin2005 wants to merge 7 commits into
anthropics:mainfrom
YuhaoLin2005:main
Closed

YuhaoLin2005 wants to merge 7 commits into
anthropics:mainfrom
YuhaoLin2005:main

Conversation

@YuhaoLin2005

@YuhaoLin2005 YuhaoLin2005 commented Jun 27, 2026 •

Copy link
Copy Markdown

Why this exists

anthropics/skills has skills for almost every task in the development lifecycle — but nothing that checks whether the AI's reasoning itself was sound before shipping.

code-reviewer checks other people's code. security-review checks for vulnerabilities. This skill checks a different dimension: did the AI actually deliver what was asked, without contradicting itself, rationalizing gaps, or claiming verification it didn't do?

What it does

Four questions before shipping:

  1. Completeness — Did I answer every request?
  2. Consistency — Did I contradict myself or the rules?
  3. Groundedness — Did I show evidence, or just claim?
  4. Honesty — Am I being honest about what wasn't done?

Relationship to existing skills

Existing skill Covers This skill covers
code-reviewer Code correctness Reasoning correctness
security-review Vulnerability detection Assumption detection
shipping-and-launch Deployment safety Delivery completeness

Extracted from practice

Derived from a five-library learning-capture workflow used across 200+ sessions. The pattern of "audit before shipping" reduced repeated mistakes measurably — bugs that slipped through one session were caught in the next because the audit habit forced them to surface.

@YuhaoLin2005 YuhaoLin2005 changed the title feat(skills): add self-audit — four-dimension output quality gate feat(skills): add self-audit — four-dimension reasoning quality gate before delivery Jun 27, 2026
…ead See-Also, clarify Hard Constraints, expand Verification
@YuhaoLin2005

Copy link
Copy Markdown
Author

This PR was replaced by #1367 (clean feature branch, same contribution). All discussion and review continues there.

👉 #1367

@YuhaoLin2005

Copy link
Copy Markdown
Author

Apologies — this PR was short-lived because I opened it from my fork's main branch during an urgent rebuild (recovering from a git rebase disaster on #1360). I've since learned to always use feature branches — it's now a hard rule in my config.

The contribution is now on #1367 with a clean self-audit feature branch. xg-gh-25's review from SwarmAI is there, and based on their suggestions I've added: (1) mechanical Step 0 file existence check before reasoning audit, (2) requirement traceability in the Completeness dimension. Full pytest results (22/22) are in the latest commit.

👉 #1367

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant