What claude-sdlc-wizard v2.0.0 just shipped
Two lanes with npm dist-tags
npm install agentic-sdlc-wizard # Reliable (@latest) — tried and true
npm install agentic-sdlc-wizard@frontier # Frontier (@frontier) — newest models
Plugin install (same result, no npm):
claude plugin install sdlc-wizard@sdlc-wizard-marketplace
The portable pattern (builder/reviewer/advisor + 95% escalation ladder)
| Role |
What it does |
When |
| Builder |
Writes code, runs tests |
Always (it's you) |
| Reviewer (brain 1) |
Cross-model adversarial check |
When builder <95% confident, or before every merge |
| Advisor (brain 2) |
Deepest thinker, full context |
Only when builder + reviewer can't reach 95% together |
Escalation ladder:
- Builder tries first. 95%+ confident → ship. No brain call.
- <95% → ask brain 1 with what you learned. Different model family = different blind spots.
- Still <95% or disagreement → ask brain 2 with what BOTH learned. Reconcile all three.
Each rung passes forward what it learned. No model redoes work.
claude-sdlc-wizard model table
| Role |
Reliable (@latest) |
Frontier (@frontier) |
| Builder |
Opus 4.6[1m] max (Anthropic) |
Opus 5.5 xhigh (Anthropic) |
| Reviewer (brain 1) |
GPT-5.6 Sol xhigh (OpenAI) |
GPT-5.6 Sol xhigh (OpenAI) |
| Advisor (brain 2) |
Fable 5.1 high (Anthropic) |
Fable 5.1 high (Anthropic) |
codex-sdlc-wizard — THE INVERSE
Builder is OpenAI. Reviewer is Anthropic (cross-family diversity). Families flip:
| Role |
Reliable (@latest) |
Frontier (@frontier) |
| Builder |
GPT-5.6 Sol xhigh (OpenAI) |
??? (OpenAI, newest) |
| Reviewer (brain 1) |
Opus 4.6 or Fable 5.1 via claude -p |
??? (Anthropic) |
| Advisor (brain 2) |
??? (strongest available, full context) |
??? |
What needs porting
-
Two-lane dist-tag infrastructure — release.yml with scripts/derive-dist-tag.sh (already proven in claude-sdlc-wizard #713). Prerelease versions get --tag frontier, stable gets --tag latest.
-
Model table in all shipped docs — exact model IDs, exact effort levels, no aliases. "GPT-5.6 Sol xhigh" not "Sol" or "Codex high".
-
Brain escalation ladder — documented in wizard doc, skills, tested in tests/test-escalation-ladder.sh (live E2E, isolated sessions). 95% confidence threshold documented in 3+ shipped files.
-
Plugin parity (#725 on claude-sdlc-wizard) — codex --plugin should give the same experience as npm install.
-
Model pin tests — tests/test-model-pins.sh (6 assertions: frontier model present, reviewer present and defaulted, advisor present, no stale refs).
Reliable = what you actually use
Same principle: Reliable is what the maintainer dogfoods. Frontier tracks newest models for experimenting. Owner will figure out the Codex-side model pins interactively.
Reference
- claude-sdlc-wizard v2.0.0 (just shipped)
- claude-sdlc-wizard #725 (plugin parity)
- claude-sdlc-wizard #720 (CI → local tests)
- claude-sdlc-wizard tests/test-escalation-ladder.sh, tests/test-model-pins.sh
What claude-sdlc-wizard v2.0.0 just shipped
Two lanes with npm dist-tags
Plugin install (same result, no npm):
The portable pattern (builder/reviewer/advisor + 95% escalation ladder)
Escalation ladder:
Each rung passes forward what it learned. No model redoes work.
claude-sdlc-wizard model table
@latest)@frontier)max(Anthropic)xhigh(Anthropic)xhigh(OpenAI)xhigh(OpenAI)high(Anthropic)high(Anthropic)codex-sdlc-wizard — THE INVERSE
Builder is OpenAI. Reviewer is Anthropic (cross-family diversity). Families flip:
@latest)@frontier)xhigh(OpenAI)claude -pWhat needs porting
Two-lane dist-tag infrastructure —
release.ymlwithscripts/derive-dist-tag.sh(already proven in claude-sdlc-wizard #713). Prerelease versions get--tag frontier, stable gets--tag latest.Model table in all shipped docs — exact model IDs, exact effort levels, no aliases. "GPT-5.6 Sol xhigh" not "Sol" or "Codex high".
Brain escalation ladder — documented in wizard doc, skills, tested in
tests/test-escalation-ladder.sh(live E2E, isolated sessions). 95% confidence threshold documented in 3+ shipped files.Plugin parity (#725 on claude-sdlc-wizard) —
codex --pluginshould give the same experience as npm install.Model pin tests —
tests/test-model-pins.sh(6 assertions: frontier model present, reviewer present and defaulted, advisor present, no stale refs).Reliable = what you actually use
Same principle: Reliable is what the maintainer dogfoods. Frontier tracks newest models for experimenting. Owner will figure out the Codex-side model pins interactively.
Reference