gstack/test/helpers
Garry Tan 82cf085213
merge: integrate origin/main (v0.5.1-v0.6.4) into team-supabase-store
Resolves conflicts in package.json (keep unified cli-eval.ts + add
eval:select) and test/skill-llm-eval.test.ts (keep judgeCost/judgeCosts
helpers + add diff-based test selection).

Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com>
2026-03-17 23:01:35 -07:00
..
eval-store.test.ts merge: integrate origin/main (v0.4.0, v0.4.1) into team-supabase-store 2026-03-16 07:49:27 -05:00
eval-store.ts merge: integrate origin/main (v0.4.0, v0.4.1) into team-supabase-store 2026-03-16 07:49:27 -05:00
llm-judge.test.ts feat: wire eval-cache + eval-tier into LLM judge, pin E2E model 2026-03-15 16:47:35 -05:00
llm-judge.ts feat: wire eval-cache + eval-tier into LLM judge, pin E2E model 2026-03-15 16:47:35 -05:00
observability.test.ts fix: never clean up observability artifacts — partial file persists after finalize 2026-03-14 12:37:38 -05:00
session-runner.test.ts feat: wire costs[] from modelUsage into eval results 2026-03-15 16:47:27 -05:00
session-runner.ts feat: wire costs[] from modelUsage into eval results 2026-03-15 16:47:27 -05:00
skill-parser.ts feat: 3-tier eval suite with planted-bug outcome testing (EVALS=1) 2026-03-14 01:17:36 -05:00
touchfiles.ts feat: interactive /plan-design-review + CEO invokes designer + 100% coverage (v0.6.4) (#149) 2026-03-17 22:48:48 -05:00