gstack/test
Garry Tan 5044c664c6
feat: add debug escalation tests (validation + LLM judge + E2E)
Skill validation: 11 new assertions covering Phase 8g trigger, structured
handoff fields, agent result handlers, debug escalation summary, Step 5.7
recommendation, ship reverted QA detection, and debug browse setup.

LLM judge: evaluates Phase 8g template quality — structured brief format,
result handling, working tree cleanup, sequential processing.

E2E: prompt-level deterministic test (verifies escalation prompt has all
required fields) + full flow stub (fixture TODO for planted regression).

Touchfile entries for diff-based test selection.

Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com>
2026-03-18 11:13:12 -07:00
..
fixtures feat: design review lite in /review and /ship + gstack-diff-scope (v0.6.3) (#142) 2026-03-17 20:12:55 -05:00
helpers feat: add debug escalation tests (validation + LLM judge + E2E) 2026-03-18 11:13:12 -07:00
gen-skill-docs.test.ts feat: interactive /plan-design-review + CEO invokes designer + 100% coverage (v0.6.4) (#149) 2026-03-17 22:48:48 -05:00
skill-e2e.test.ts feat: add debug escalation tests (validation + LLM judge + E2E) 2026-03-18 11:13:12 -07:00
skill-llm-eval.test.ts feat: add debug escalation tests (validation + LLM judge + E2E) 2026-03-18 11:13:12 -07:00
skill-parser.test.ts feat: SKILL.md template system, 3-tier testing, DX tools (v0.3.3) (#41) 2026-03-13 21:08:12 -07:00
skill-validation.test.ts feat: add debug escalation tests (validation + LLM judge + E2E) 2026-03-18 11:13:12 -07:00
touchfiles.test.ts feat: add debug escalation tests (validation + LLM judge + E2E) 2026-03-18 11:13:12 -07:00