* feat: add better params to working representation fetch in SDKs, return messages when added * fix: working representation routes now accepting all parameters properly, with tests * feat: add metadata/config fields to SDK objects where viable * fix: tests * feat: refactor SDKs to use representation config; [TEMP STAINLESS BUILD] update API * feat: add representation object to sdks * fix: use stainless sdk on branch * fix: update TypeScript SDK tsconfig to use node16 module resolution * fix: add isolatedModules = true to tsconfig * fix: lol * chore: coderabbit review * feat: make delete session real * feat: add observations routes with delete endpoints for documents. make session deletion real. * chore: type cleanup * fix: tests * chore: coderabbit review * fix: namespace by workspace * feat: add ability to customize messages_per_summary at both workspace and session level * chore: tests for summary config * chore: coderabbit cleanup * feat: make session and workspace config totally customizeable * feat: add search by peer knowledge (#250) * feat: search by peer perspective * fix: enforce workspace in filters, make messages distinct in join * fix: batch and merge migration steps * fix: add refresh, add config to workspace, add refresh function, make fields readonly * fix: search distinct * fix: merge migrations * fix: merge migrations * fix: batch deletions, improve comments, limit consolidate dream to 100 docs at a time, auth on observations routes * chore: review * chore: coderabbit * chore: review * chore: broken comment * feat: add set peer card route to API * feat: create advanced configuration parameters with message>session>workspace hierarchy * [wip] build unified testing harness * chore: lint * fix: cache invalidation, naming things, etc * feat: longmem tests * chore: peer config refactor * feat: consolidate dream working, refactor representation * fix: Various CR Comment Fixes * feat: Allow configurable Redis port for harness instances and update cleanup methods to be asynchronous. * feat: agentic ingestion task!!! * feat: agentic deriver * feat: dialectic agent and dreamer agent * chore: browbeat tests into passing * fix: nits * chore: remove old code, update config files * fix: simplify deriver * feat: dialectic agent prompt updates, re-introduce non_agent deriver, eval tweaks * feat: fast deriver, dreamer, then dialectic * fix: tweaks across the board * feat: add baseline tests * feat: truncation in tools and client, tweaks for evals * feat: add locomo, fix longmem judge!!! * fix: locomo f1 is trash, use llm judge * feat: trace creation * feat: add first draft of obex benchmark, fix embedding model, fix locomo methodology * fix: locomo session-optimized, better logging of cache usage and better cache usage * chore: use openrouter for baselines * fix: add test for merge migration * chore: opus-powered cleanup * fix: add config for vllm, better client * chore: clean up clients.py a bit * chore: move magic numbers to config, add tests for agent tools * fix: wrong mock in dialectic tests, make ToolContext a dataclass * feat: tweak prompts, make deriver explicit-only * feat: more prompt & tool tweaks * chore: more tweaks * feat: dream with subagents * fix: make dream trigger override scheduled, play around with dream agents * chore: cleanup deriver * chore: cleanup dialectic * chore: cleanup orchestrator * chore: comment out dream stuff, WIPing * fix: inc temp on retry, typechecking * feat: tweak dreaming * feat: contradiction obs * Add dream trees * chore: preserve reasoning_details from openrouter in client * fix: get_observation_context correct params * fix: use correct message id in tool * chore: cleanup longmem runner * chore: clean up tests, remove dream tests for now as rearchitecting around trees * chore: update stainless deps * Update threholding mechanism * chore: pre-commit hooks whitespace * chore: clean up types * feat: add explicit bench * fix: address additional basepyright issues * fix: adding logging as a fixture on honcho_llm_call and supporting dialectic loging. (#305) * fix: lock on db for tool calls * chore: clean up experimental derivers * chore: coderabbit review cleanup * feat: add streaming support to agentic dialectic * feat: prometheus token tracking for deriver and dialectic * fix: self-loops for isolated nodes * chore: PascalCase for prometheus parameter typing * feat: add reasoning levels to dialectic agent * chore: delete old file, add new fake env vars in unittest.yml * fix: all fields needed for dialectic reasoning level configs * feat: track dreaming usage in prometheus * chore: Create backwards compatabile conclusion and queue endpoints * fix: remove redundant try-catch, add trace label, move .limit to end of statement * fix: remove vignettes (for now), review fixes, remove merge migration, config cleanup * chore: code review / cleanup * chore: merge fixes * chore: clean up, remove reasoning_focus, reintroduce peer cards in dreamers * chore: code rabbit nitpicks * fix: add unique index for pending dreams in queue * fix: revert removal of surprisal in dreamer config --------- Co-authored-by: Vineeth Voruganti <13438633+VVoruganti@users.noreply.github.com> Co-authored-by: 3un01a <3un01a@plasticlabs.ai> Co-authored-by: 3un01a <3un01a.labs@gmail.com> Co-authored-by: ajspig <46900795+ajspig@users.noreply.github.com> |
||
|---|---|---|
| .. | ||
| revisions | ||
| README.md | ||
| __init__.py | ||
| conftest.py | ||
| registry.py | ||
| scaffold.py | ||
| test_pipeline.py | ||
| verifier.py | ||
README.md
These tests validate Alembic migrations end-to-end for structure, order, and correctness. They ensure reversibility, expected schema and data, and integration with the registry and pipeline. The key components are the verifier, the test pipeline, the registry, and the revisions under test.
Verifier
- The verifier runs checks when specific revisions are applied and reverted.
- It validates the schema after upgrade, verifies data migrations such as defaults, backfills, and transforms, and confirms reversibility after downgrade.
- Assertions are grouped per-revision or feature, and helpers use the SQLAlchemy inspector to introspect the database.
Test Pipeline
- The pipeline orchestrates the database lifecycle by creating a database, applying upgrades and downgrades, running verifications, and tearing down resources.
- It typically starts from base, upgrades to the revision immediately before a target revision, seeds the DB, runs the target migration, and then validates the schema + data
- It relies on shared fixtures such as
engine,connection, andalembic_config, and it ensures isolation per test
Registry
- The registry declares revisions and test metadata used to drive scenarios.
- It defines ordering and selection, attaches verifier callbacks to specific revisions or ranges, and centralizes per revision expectations.
Revisions
- Revisions are the migration scripts under
alembic/revisions. - Each revision should provide functions decorated with
register_before_upgrade()andregister_after_upgrade(). These are used to validate schemas and data before and after a migration is run
Running the Tests
- Tests can be run all together
pytest tests/alembicor individuallypytest tests/alembic -k "revision_number". For example, to run the test against a1b2c3d4e5f6_initial_schema.py, you would run the commandpytest tests/alembic -k "a1b2c3d4e5f6"