* feat: implement forced batching for representation tasks and adjust max tokens - Updated `REPRESENTATION_BATCH_MAX_TOKENS` to 1024 in `config.py`. - Enhanced `get_and_claim_work_units` in `QueueManager` to enforce batching based on token thresholds. - Added tests to ensure representation work units are only claimed when token counts meet or exceed the threshold. - Introduced a synthesis prompt for tool execution to improve final response generation. * chore: update queue-status docs to match new behavior * fix: make representation work query efficient * fix: Align alembic models and force dreams on for tests * fix: clean queue between tests --------- Co-authored-by: Vineeth Voruganti <13438633+VVoruganti@users.noreply.github.com> |
||
|---|---|---|
| .. | ||
| api-reference | ||
| contributing | ||
| documentation | ||
| guides | ||
| migrations | ||
| README.md | ||
| openapi.json | ||
README.md
This subdirectory contains the peer-paradigm documentation for Honcho (Honcho v2.0.0 onwards).