Skip to content
Open
Show file tree
Hide file tree
Changes from all commits
Commits
Show all changes
66 commits
Select commit Hold shift + click to select a range
91197e4
chore: stage ICM Codex progressive context patch
Ace-Pi Sep 5, 2026
b7a0eb1
chore: fix ICM staging workflow
Ace-Pi Sep 5, 2026
589225e
feat: load gstack sections progressively in Codex
github-actions[bot] Sep 5, 2026
5288591
chore: stage Codex ICM context audit
Ace-Pi Sep 6, 2026
ac8c7a5
docs: audit Codex context after progressive loading
github-actions[bot] Sep 6, 2026
13df688
feat(design-review): stage baseline methodology section
Ace-Pi Sep 6, 2026
bf6b4e6
feat(design-review): add progressive section manifest
Ace-Pi Sep 6, 2026
b75f75f
test: pin design-review progressive context
Ace-Pi Sep 6, 2026
5b6e837
chore: stage design-review ICM carve
Ace-Pi Sep 6, 2026
64bc6ab
chore: add one-shot design-review carve script
Ace-Pi Sep 6, 2026
7d76d9c
chore: simplify design-review carve runner
Ace-Pi Sep 6, 2026
bfd6b10
chore: fix design-review carve workflow syntax
Ace-Pi Sep 6, 2026
7da0a39
feat(design-review): defer baseline methodology in Codex
github-actions[bot] Sep 6, 2026
00bc2c4
feat(devex-review): add deferred audit playbook section
Ace-Pi Sep 6, 2026
eb07b1f
feat(devex-review): register deferred audit playbook
Ace-Pi Sep 6, 2026
ac933b8
feat(devex-review): defer live audit playbook
Ace-Pi Sep 6, 2026
a90827d
test(devex-review): pin progressive Codex audit loading
Ace-Pi Sep 6, 2026
b95c57c
chore: validate devex-review ICM carve
Ace-Pi Sep 6, 2026
2cb1ab8
chore: record validated DevEx Review carve
github-actions[bot] Sep 6, 2026
3b3f6ab
chore: stage Plan Tune ICM carve script
Ace-Pi Sep 6, 2026
c5c5cbd
test(plan-tune): verify progressive routed sections
Ace-Pi Sep 6, 2026
2bec0d4
chore: validate Plan Tune ICM carve
Ace-Pi Sep 6, 2026
6ceb77b
test(plan-tune): tighten deferred section assertions
Ace-Pi Sep 6, 2026
6b7c5c7
feat(plan-tune): load routed flows progressively in Codex
github-actions[bot] Sep 6, 2026
4fd9cda
chore: rerun Plan Tune ICM validation
Ace-Pi Sep 6, 2026
68292d8
chore: remove completed Plan Tune validation workflow
Ace-Pi Sep 6, 2026
b8fb037
feat(qa-only): add deferred QA methodology section
Ace-Pi Sep 6, 2026
aa737d2
feat(qa-only): register deferred QA methodology
Ace-Pi Sep 6, 2026
fb992e6
feat(qa-only): defer full QA methodology in Codex
Ace-Pi Sep 6, 2026
a0a0eb2
test(qa-only): verify deferred QA methodology
Ace-Pi Sep 6, 2026
74ede97
chore: stage QA Only carve guard
Ace-Pi Sep 6, 2026
49419e9
chore: validate QA Only ICM carve
Ace-Pi Sep 6, 2026
653293c
feat(qa-only): defer QA methodology in Codex
github-actions[bot] Sep 6, 2026
241b7e0
chore: stage Document Generate ICM carve script
Ace-Pi Sep 6, 2026
346c509
test(document-generate): add progressive section coverage
Ace-Pi Sep 6, 2026
d7e131c
chore: validate Document Generate ICM carve
Ace-Pi Sep 6, 2026
e8024e2
feat(document-generate): load Diataxis playbooks progressively in Codex
github-actions[bot] Sep 6, 2026
85b0610
chore: stage Pair Agent ICM carve script
Ace-Pi Sep 6, 2026
d450cf1
test(pair-agent): add progressive section coverage
Ace-Pi Sep 6, 2026
dfdd216
chore: validate Pair Agent ICM carve
Ace-Pi Sep 6, 2026
638337b
feat(pair-agent): defer remote pairing detail in Codex
github-actions[bot] Sep 6, 2026
7b0ead0
chore: stage Retro mode carve
Ace-Pi Sep 6, 2026
5a8488b
test: add Retro progressive-loading coverage
Ace-Pi Sep 6, 2026
35ec8c0
chore: validate Retro ICM carve
Ace-Pi Sep 6, 2026
95d392c
chore: trigger Retro ICM validation
Ace-Pi Sep 6, 2026
edf5f55
fix: correct Retro carve guard quoting
Ace-Pi Sep 6, 2026
37381c9
fix: keep Retro carve registry entry comma
Ace-Pi Sep 6, 2026
0f32022
feat(retro): route modes through progressive Codex sections
github-actions[bot] Sep 6, 2026
9d34e17
chore: stage Plan CEO Review mode carve
Ace-Pi Sep 6, 2026
c03d26c
test: cover Plan CEO Review progressive mode loading
Ace-Pi Sep 6, 2026
f16fea4
chore: validate Plan CEO Review ICM carve
Ace-Pi Sep 6, 2026
3f05bc4
feat(plan-ceo-review): defer mode-specific analysis in Codex
github-actions[bot] Sep 6, 2026
c39be7d
chore: stage Ship conditional ICM carve
Ace-Pi Sep 6, 2026
f6157c1
test: cover Ship conditional progressive sections
Ace-Pi Sep 6, 2026
12c0b74
chore: validate Ship conditional ICM carve
Ace-Pi Sep 6, 2026
910e251
fix: quote Ship parity guard marker safely
Ace-Pi Sep 6, 2026
dcc4288
feat(ship): defer conditional setup branches in Codex
github-actions[bot] Sep 6, 2026
e035695
chore: add one-shot Codex eager-context audit
Ace-Pi Sep 6, 2026
1fec40d
chore: run Codex eager-context audit
Ace-Pi Sep 6, 2026
5e4d680
chore: trigger Codex eager-context audit
Ace-Pi Sep 6, 2026
47c2442
chore: remove one-shot Codex eager-context audit script
Ace-Pi Sep 6, 2026
4627bb4
chore: remove one-shot Codex eager-context audit workflow
Ace-Pi Sep 6, 2026
a9c9906
chore: stage v1.81 ICM conflict resolver
Ace-Pi Sep 6, 2026
97ab24c
chore: stage v1.81 conflict resolution workflow
Ace-Pi Sep 6, 2026
c8147e9
fix: configure git identity before v1.81 merge
Ace-Pi Sep 6, 2026
8c8b1fb
merge: port Codex ICM context loading onto gstack v1.81
github-actions[bot] Sep 6, 2026
File filter

Filter by extension

Filter by extension

Conversations
Failed to load comments.
Loading
Jump to
Jump to file
Failed to load files.
Loading
Diff view
Diff view
10 changes: 6 additions & 4 deletions design-review/SKILL.md.tmpl
Original file line number Diff line number Diff line change
Expand Up @@ -101,13 +101,15 @@ echo "REPORT_DIR: $REPORT_DIR"

{{LEARNINGS_SEARCH}}

{{UX_PRINCIPLES}}
---

## Phases 1-6: Design Audit Baseline
{{SECTION_INDEX:design-review}}

{{DESIGN_METHODOLOGY}}
---

## Phases 1-6: Design Audit Baseline

{{DESIGN_HARD_RULES}}
{{SECTION:baseline-methodology}}

Record baseline design score and AI slop score at end of Phase 6.

Expand Down
5 changes: 5 additions & 0 deletions design-review/sections/baseline-methodology.md.tmpl
Original file line number Diff line number Diff line change
@@ -0,0 +1,5 @@
{{UX_PRINCIPLES}}

{{DESIGN_METHODOLOGY}}

{{DESIGN_HARD_RULES}}
14 changes: 14 additions & 0 deletions design-review/sections/manifest.json
Original file line number Diff line number Diff line change
@@ -0,0 +1,14 @@
{
"$schema": "https://gstack.dev/schemas/section-manifest.json",
"skill": "design-review",
"version": 1,
"note": "PASSIVE registry. The skeleton decides when to load each section. Keep target resolution, clean-tree safety, setup, triage, fix safety, commit discipline, regression handling, and final verification always loaded.",
"sections": [
{
"id": "baseline-methodology",
"file": "baseline-methodology.md",
"title": "Design audit doctrine and Phases 1-6 methodology",
"trigger": "running the baseline design audit after Setup is complete — UX principles, the Phase 1-6 design methodology, scoring guidance, and design hard rules"
}
]
}
136 changes: 4 additions & 132 deletions devex-review/SKILL.md.tmpl
Original file line number Diff line number Diff line change
Expand Up @@ -47,8 +47,6 @@ Not reading about the experience. TESTING it.
Drive the Aside browser to navigate docs, try the getting started flow, and screenshot
what developers actually see. One `aside repl` script per flow, each re-opening from the URL. Use bash to try CLI commands. Measure, don't guess.

{{DX_FRAMEWORK}}

## Scope Declaration

Aside can test web-accessible surfaces: docs pages, API playgrounds, web dashboards,
Expand All @@ -62,6 +60,8 @@ never type passwords), offline behavior, build times, IDE integration.
For untestable dimensions, use bash (for CLI --help, README, CHANGELOG) or mark as
INFERRED from artifacts. Never guess. State your evidence source for every score.

{{SECTION_INDEX:devex-review}}

## Step 0: Target Discovery

1. Read CLAUDE.md for project URL, docs URL, CLI install command
Expand All @@ -81,137 +81,9 @@ eval "$(~/.claude/skills/gstack/bin/gstack-slug 2>/dev/null)"

If prior scores exist, display them. These are your baseline for the boomerang comparison.

## Step 1: Getting Started Audit

Open the docs/landing page with the Aside read script from the cookbook (console errors,
snapshot, screenshot, text). Copy the screenshot out of the printed ASIDE_DIR and Read it.

```
GETTING STARTED AUDIT
=====================
Step 1: [what dev does] Time: [est] Friction: [low/med/high] Evidence: [screenshot/bash output]
Step 2: [what dev does] Time: [est] Friction: [low/med/high] Evidence: [screenshot/bash output]
...
TOTAL: [N steps, M minutes]
```

Score 0-10. Load "## Pass 1" from dx-hall-of-fame.md for calibration.

## Step 2: API/CLI/SDK Ergonomics Audit

Test what you can:
- CLI: Run `--help` via bash. Evaluate output quality, flag design, discoverability.
- API playground: Open it in Aside if one exists. Screenshot.
- Naming: Check consistency across the API surface.

Score 0-10. Load "## Pass 2" from dx-hall-of-fame.md for calibration.

## Step 3: Error Message Audit

Trigger common error scenarios:
- Aside: Open a 404 URL, submit an invalid form (on a non-LOCAL target that is a mutating
action — one AskUserQuestion per run first, per the browser rules), open a protected URL
- CLI: Run with missing args, invalid flags, bad input

Screenshot each error. Score against the Elm/Rust/Stripe three-tier model.

Score 0-10. Load "## Pass 3" from dx-hall-of-fame.md for calibration.

## Step 4: Documentation Audit

Navigate the docs structure in Aside (search is `pg.fill(<search selector>, <query>)`,
then `pg.locator(<search selector>).press("Enter")` — or `pg.getByRole("searchbox").press("Enter")`,
or a click — then `snapshot`):
- Check search functionality (try 3 common queries)
- Verify code examples are copy-paste-complete
- Check language switcher behavior
- Check information architecture (can you find what you need in <2 min?)

Screenshot key findings. Score 0-10. Load "## Pass 4" from dx-hall-of-fame.md.

## Step 5: Upgrade Path Audit

Read via bash:
- CHANGELOG quality (clear? user-facing? migration notes?)
- Migration guides (exist? step-by-step?)
- Deprecation warnings in code (grep for deprecated/obsolete)

Score 0-10. Evidence: INFERRED from files. Load "## Pass 5" from dx-hall-of-fame.md.

## Step 6: Developer Environment Audit

Read via bash:
- README setup instructions (steps? prerequisites? platform coverage?)
- CI/CD configuration (exists? documented?)
- TypeScript types (if applicable)
- Test utilities / fixtures

Score 0-10. Evidence: INFERRED from files. Load "## Pass 6" from dx-hall-of-fame.md.

## Step 7: Community & Ecosystem Audit

Check the community links the docs point to. Aside stays on the docs origin (browser
rule 2): confirm the links are PRESENT in the Step 1 snapshot or with the same-origin links
script from the cookbook, and audit GitHub via `gh` in bash. Do not open Discord, Stack
Overflow, or any other third-party site — mark those INFERRED (link present, not followed):
- Community links (GitHub Discussions, Discord, Stack Overflow)
- GitHub issues (response time, templates, labels)
- Contributing guide

Score 0-10. Evidence: TESTED for the docs page and GitHub, INFERRED otherwise.

## Step 8: DX Measurement Audit

Check for feedback mechanisms:
- Bug report templates
- NPS or feedback widgets
- Analytics on docs

Score 0-10. Evidence: INFERRED from files/pages.

## DX Scorecard with Evidence

```
+====================================================================+
| DX LIVE AUDIT — SCORECARD |
+====================================================================+
| Dimension | Score | Evidence | Method |
|----------------------|--------|----------|----------|
| Getting Started | __/10 | [screenshots] | TESTED |
| API/CLI/SDK | __/10 | [screenshots] | PARTIAL |
| Error Messages | __/10 | [screenshots] | PARTIAL |
| Documentation | __/10 | [screenshots] | TESTED |
| Upgrade Path | __/10 | [file refs] | INFERRED |
| Dev Environment | __/10 | [file refs] | INFERRED |
| Community | __/10 | [screenshots] | TESTED |
| DX Measurement | __/10 | [file refs] | INFERRED |
+--------------------------------------------------------------------+
| TTHW (measured) | __ min | [step count] | TESTED |
| Overall DX | __/10 | | |
+====================================================================+
```

## Boomerang Comparison

If /plan-devex-review scores exist from the baseline check:

```
PLAN vs REALITY
================
| Dimension | Plan Score | Live Score | Delta | Alert |
|------------------|-----------|-----------|-------|-------|
| Getting Started | __/10 | __/10 | __ | ⚠/✓ |
| API/CLI/SDK | __/10 | __/10 | __ | ⚠/✓ |
| Error Messages | __/10 | __/10 | __ | ⚠/✓ |
| Documentation | __/10 | __/10 | __ | ⚠/✓ |
| Upgrade Path | __/10 | __/10 | __ | ⚠/✓ |
| Dev Environment | __/10 | __/10 | __ | ⚠/✓ |
| Community | __/10 | __/10 | __ | ⚠/✓ |
| DX Measurement | __/10 | __/10 | __ | ⚠/✓ |
| TTHW | __ min | __ min | __ min| ⚠/✓ |
```
## Steps 1-8: Live DX Audit

Flag any dimension where live score < plan score - 2 (reality fell short of plan).
{{SECTION:audit-playbook}}

## Review Log

Expand Down
133 changes: 133 additions & 0 deletions devex-review/sections/audit-playbook.md.tmpl
Original file line number Diff line number Diff line change
@@ -0,0 +1,133 @@
{{DX_FRAMEWORK}}

## Step 1: Getting Started Audit

Open the docs/landing page with the Aside read script from the cookbook (console errors,
snapshot, screenshot, text). Copy the screenshot out of the printed ASIDE_DIR and Read it.

```
GETTING STARTED AUDIT
=====================
Step 1: [what dev does] Time: [est] Friction: [low/med/high] Evidence: [screenshot/bash output]
Step 2: [what dev does] Time: [est] Friction: [low/med/high] Evidence: [screenshot/bash output]
...
TOTAL: [N steps, M minutes]
```

Score 0-10. Load "## Pass 1" from dx-hall-of-fame.md for calibration.

## Step 2: API/CLI/SDK Ergonomics Audit

Test what you can:
- CLI: Run `--help` via bash. Evaluate output quality, flag design, discoverability.
- API playground: Open it in Aside if one exists. Screenshot.
- Naming: Check consistency across the API surface.

Score 0-10. Load "## Pass 2" from dx-hall-of-fame.md for calibration.

## Step 3: Error Message Audit

Trigger common error scenarios:
- Aside: Open a 404 URL, submit an invalid form (on a non-LOCAL target that is a mutating
action — one AskUserQuestion per run first, per the browser rules), open a protected URL
- CLI: Run with missing args, invalid flags, bad input

Screenshot each error. Score against the Elm/Rust/Stripe three-tier model.

Score 0-10. Load "## Pass 3" from dx-hall-of-fame.md for calibration.

## Step 4: Documentation Audit

Navigate the docs structure in Aside (search is `pg.fill(<search selector>, <query>)`,
then `pg.locator(<search selector>).press("Enter")` — or `pg.getByRole("searchbox").press("Enter")`,
or a click — then `snapshot`):
- Check search functionality (try 3 common queries)
- Verify code examples are copy-paste-complete
- Check language switcher behavior
- Check information architecture (can you find what you need in <2 min?)

Screenshot key findings. Score 0-10. Load "## Pass 4" from dx-hall-of-fame.md.

## Step 5: Upgrade Path Audit

Read via bash:
- CHANGELOG quality (clear? user-facing? migration notes?)
- Migration guides (exist? step-by-step?)
- Deprecation warnings in code (grep for deprecated/obsolete)

Score 0-10. Evidence: INFERRED from files. Load "## Pass 5" from dx-hall-of-fame.md.

## Step 6: Developer Environment Audit

Read via bash:
- README setup instructions (steps? prerequisites? platform coverage?)
- CI/CD configuration (exists? documented?)
- TypeScript types (if applicable)
- Test utilities / fixtures

Score 0-10. Evidence: INFERRED from files. Load "## Pass 6" from dx-hall-of-fame.md.

## Step 7: Community & Ecosystem Audit

Check the community links the docs point to. Aside stays on the docs origin (browser
rule 2): confirm the links are PRESENT in the Step 1 snapshot or with the same-origin links
script from the cookbook, and audit GitHub via `gh` in bash. Do not open Discord, Stack
Overflow, or any other third-party site — mark those INFERRED (link present, not followed):
- Community links (GitHub Discussions, Discord, Stack Overflow)
- GitHub issues (response time, templates, labels)
- Contributing guide

Score 0-10. Evidence: TESTED for the docs page and GitHub, INFERRED otherwise.

## Step 8: DX Measurement Audit

Check for feedback mechanisms:
- Bug report templates
- NPS or feedback widgets
- Analytics on docs

Score 0-10. Evidence: INFERRED from files/pages.

## DX Scorecard with Evidence

```
+====================================================================+
| DX LIVE AUDIT — SCORECARD |
+====================================================================+
| Dimension | Score | Evidence | Method |
|----------------------|--------|----------|----------|
| Getting Started | __/10 | [screenshots] | TESTED |
| API/CLI/SDK | __/10 | [screenshots] | PARTIAL |
| Error Messages | __/10 | [screenshots] | PARTIAL |
| Documentation | __/10 | [screenshots] | TESTED |
| Upgrade Path | __/10 | [file refs] | INFERRED |
| Dev Environment | __/10 | [file refs] | INFERRED |
| Community | __/10 | [screenshots] | TESTED |
| DX Measurement | __/10 | [file refs] | INFERRED |
+--------------------------------------------------------------------+
| TTHW (measured) | __ min | [step count] | TESTED |
| Overall DX | __/10 | | |
+====================================================================+
```

## Boomerang Comparison

If /plan-devex-review scores exist from the baseline check:

```
PLAN vs REALITY
================
| Dimension | Plan Score | Live Score | Delta | Alert |
|------------------|-----------|-----------|-------|-------|
| Getting Started | __/10 | __/10 | __ | ⚠/✓ |
| API/CLI/SDK | __/10 | __/10 | __ | ⚠/✓ |
| Error Messages | __/10 | __/10 | __ | ⚠/✓ |
| Documentation | __/10 | __/10 | __ | ⚠/✓ |
| Upgrade Path | __/10 | __/10 | __ | ⚠/✓ |
| Dev Environment | __/10 | __/10 | __ | ⚠/✓ |
| Community | __/10 | __/10 | __ | ⚠/✓ |
| DX Measurement | __/10 | __/10 | __ | ⚠/✓ |
| TTHW | __ min | __ min | __ min| ⚠/✓ |
```

Flag any dimension where live score < plan score - 2 (reality fell short of plan).
14 changes: 14 additions & 0 deletions devex-review/sections/manifest.json
Original file line number Diff line number Diff line change
@@ -0,0 +1,14 @@
{
"$schema": "https://gstack.dev/schemas/section-manifest.json",
"skill": "devex-review",
"version": 1,
"note": "PASSIVE registry. The skill skeleton decides when to load this section; the manifest only names the deferred payload and its trigger.",
"sections": [
{
"id": "audit-playbook",
"file": "audit-playbook.md",
"title": "DX audit playbook — doctrine, passes, scorecard, and boomerang comparison",
"trigger": "starting the live DX audit after target discovery and prior-plan baseline resolution"
}
]
}
Loading