gstack

History

Garry Tan ae0a9ad195 feat: GStack Learns — per-project self-learning infrastructure (v0.13.4.0) (#622 ) * feat: learnings + confidence resolvers — cross-skill memory infrastructure Three new resolvers for the self-learning system: - LEARNINGS_SEARCH: tells skills to load prior learnings before analysis - LEARNINGS_LOG: tells skills to capture discoveries after completing work - CONFIDENCE_CALIBRATION: adds 1-10 confidence scoring to all review findings Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> * feat: learnings bin scripts — append-only JSONL read/write gstack-learnings-log: validates JSON, auto-injects timestamp, appends to ~/.gstack/projects/$SLUG/learnings.jsonl. Append-only (no mutation). gstack-learnings-search: reads/filters/dedupes learnings with confidence decay (observed/inferred lose 1pt/30d), cross-project discovery, and "latest winner" resolution per key+type. Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> * feat: learnings count in preamble output Every skill now prints "LEARNINGS: N entries loaded" during preamble, making the compounding loop visible to the user. Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> * feat: integrate learnings + confidence into 9 skill templates Add {{LEARNINGS_SEARCH}}, {{LEARNINGS_LOG}}, and {{CONFIDENCE_CALIBRATION}} placeholders to review, ship, plan-eng-review, plan-ceo-review, office-hours, investigate, retro, and cso templates. Regenerated all SKILL.md files. Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> * feat: /learn skill — manage project learnings New skill for reviewing, searching, pruning, and exporting what gstack has learned across sessions. Commands: /learn, /learn search, /learn prune, /learn export, /learn stats, /learn add. Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> * docs: self-learning roadmap — 5-release design doc Covers: R1 GStack Learns (v0.14), R2 Review Army (v0.15), R3 Smart Ceremony (v0.16), R4 /autoship (v0.17), R5 Studio (v0.18). Inspired by Compound Engineering, adapted to GStack's architecture. Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> * test: learnings bin script unit tests — 13 tests, free Tests gstack-learnings-log (valid/invalid JSON, timestamp injection, append-only) and gstack-learnings-search (dedup, type/query/limit filters, confidence decay, user-stated no-decay, malformed JSONL skip). Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> * chore: bump version and changelog (v0.13.4.0) Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> * test: learnings resolver + bin script edge case tests — 21 new tests, free Adds gen-skill-docs coverage for LEARNINGS_SEARCH, LEARNINGS_LOG, and CONFIDENCE_CALIBRATION resolvers. Adds bin script edge cases: timestamp preservation, special characters, files array, sort order, type grouping, combined filtering, missing fields, confidence floor at 0. Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> * fix: sync package.json version with VERSION file (0.13.4.0) Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> * chore: gitignore .factory/ — generated output, not source Same pattern as .claude/skills/ and .agents/. These SKILL.md files are generated from .tmpl templates by gen:skill-docs --host factory. Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> * test: /learn E2E — seed 3 learnings, verify agent surfaces them Seeds N+1 query pattern, stale cache pitfall, and rubocop preference into learnings.jsonl, then runs /learn and checks that at least 2/3 appear in the agent's output. Gate tier, ~$0.25/run. Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> --------- Co-authored-by: Claude Opus 4.6 (1M context) <noreply@anthropic.com>		2026-03-29 17:02:01 -06:00
..
fixtures	feat: test coverage catalog — shared audit across plan/ship/review (v0.10.1.0) (#259 )	2026-03-22 11:28:16 -07:00
helpers	feat: GStack Learns — per-project self-learning infrastructure (v0.13.4.0) (#622 )	2026-03-29 17:02:01 -06:00
analytics.test.ts	feat: safety hook skills + skill usage telemetry (v0.7.1) (#189 )	2026-03-18 23:57:59 -05:00
audit-compliance.test.ts	fix: security audit compliance — credentials, telemetry, bun pin, untrusted warning (v0.12.12.0) (#574 )	2026-03-27 12:06:58 -06:00
codex-e2e.test.ts	feat: worktree isolation for E2E tests + infrastructure elegance (v0.11.12.0) (#425 )	2026-03-23 23:05:22 -07:00
gemini-e2e.test.ts	feat: worktree isolation for E2E tests + infrastructure elegance (v0.11.12.0) (#425 )	2026-03-23 23:05:22 -07:00
gen-skill-docs.test.ts	feat: GStack Learns — per-project self-learning infrastructure (v0.13.4.0) (#622 )	2026-03-29 17:02:01 -06:00
global-discover.test.ts	feat: /retro global — cross-project AI coding retrospective (v0.10.2.0) (#316 )	2026-03-22 13:52:47 -07:00
hook-scripts.test.ts	feat: safety hook skills + skill usage telemetry (v0.7.1) (#189 )	2026-03-18 23:57:59 -05:00
learnings.test.ts	feat: GStack Learns — per-project self-learning infrastructure (v0.13.4.0) (#622 )	2026-03-29 17:02:01 -06:00
review-log.test.ts	fix: community PRs + security hardening + E2E stability (v0.12.7.0) (#552 )	2026-03-26 23:21:27 -06:00
skill-e2e-bws.test.ts	fix: community PRs + security hardening + E2E stability (v0.12.7.0) (#552 )	2026-03-26 23:21:27 -06:00
skill-e2e-cso.test.ts	feat: /cso v2 — infrastructure-first security audit (v0.11.6.0) (#384 )	2026-03-23 06:57:22 -07:00
skill-e2e-deploy.test.ts	feat: /land-and-deploy first-run dry run + staging-first + trust ladder (v0.12.2.0) (#518 )	2026-03-26 11:08:31 -07:00
skill-e2e-design.test.ts	feat: CI evals on Ubicloud — 12 parallel runners + Docker image (v0.11.10.0) (#360 )	2026-03-23 10:17:33 -07:00
skill-e2e-learnings.test.ts	feat: GStack Learns — per-project self-learning infrastructure (v0.13.4.0) (#622 )	2026-03-29 17:02:01 -06:00
skill-e2e-plan.test.ts	test: E2E tests for plan review report and Codex offering (v0.11.15.0) (#449 )	2026-03-24 07:30:24 -07:00
skill-e2e-qa-bugs.test.ts	feat: CI evals on Ubicloud — 12 parallel runners + Docker image (v0.11.10.0) (#360 )	2026-03-23 10:17:33 -07:00
skill-e2e-qa-workflow.test.ts	feat: CI evals on Ubicloud — 12 parallel runners + Docker image (v0.11.10.0) (#360 )	2026-03-23 10:17:33 -07:00
skill-e2e-review.test.ts	fix: community PRs + security hardening + E2E stability (v0.12.7.0) (#552 )	2026-03-26 23:21:27 -06:00
skill-e2e-sidebar.test.ts	fix: sidebar agent uses real tab URL instead of stale Playwright URL (v0.12.6.0) (#544 )	2026-03-26 22:07:03 -06:00
skill-e2e-workflow.test.ts	feat: 2-tier E2E test system — granular touchfiles + gate/periodic split (v0.11.16.0) (#450 )	2026-03-24 15:24:00 -07:00
skill-e2e.test.ts	feat: test coverage catalog — shared audit across plan/ship/review (v0.10.1.0) (#259 )	2026-03-22 11:28:16 -07:00
skill-llm-eval.test.ts	feat: voice directive for all skills (v0.12.3.0) (#520 )	2026-03-26 17:31:53 -06:00
skill-parser.test.ts	feat: SKILL.md template system, 3-tier testing, DX tools (v0.3.3) (#41 )	2026-03-13 21:08:12 -07:00
skill-routing-e2e.test.ts	fix: community PRs + security hardening + E2E stability (v0.12.7.0) (#552 )	2026-03-26 23:21:27 -06:00
skill-validation.test.ts	fix: Codex hang fixes — plan visibility, stdout buffering, reasoning effort (v0.12.4.0) (#536 )	2026-03-26 18:19:26 -06:00
telemetry.test.ts	fix: security audit remediation — 12 fixes, 20 tests (v0.13.1.0) (#595 )	2026-03-28 08:35:24 -06:00
touchfiles.test.ts	feat: 2-tier E2E test system — granular touchfiles + gate/periodic split (v0.11.16.0) (#450 )	2026-03-24 15:24:00 -07:00
uninstall.test.ts	feat: community PRs — faster install, skill namespacing, uninstall, Codex fallback, Windows fix, Python patterns (v0.12.9.0) (#561 )	2026-03-27 00:44:37 -06:00
worktree.test.ts	feat: worktree isolation for E2E tests + infrastructure elegance (v0.11.12.0) (#425 )	2026-03-23 23:05:22 -07:00