hermes-agent/website/docs/guides
Teknium 1c156736dc
docs: warn that mid-session model switches break prompt caching (#58747)
/model switches, primary-model fallback, and credential-pool key
rotation all change the prompt-cache key (model and/or account), so
the next turn re-reads the entire conversation at full input price.
Add cost warnings everywhere docs recommend or describe these paths:

- reference/slash-commands.md: cost note on both /model rows
- user-guide/features/fallback-providers.md: warning admonition
- user-guide/features/credential-pools.md: warning admonition
- user-guide/configuring-models.md: mid-session switch warning
- guides/tips.md: expand cache tip + /model tip
- reference/faq.md: warning on the switch-back-and-forth example
- user-guide/desktop.md: composer picker bullet
- developer-guide/context-compression-and-caching.md: new
  cache-aware design pattern (model identity is part of the key)
2026-07-05 04:34:05 -07:00
..
_category_.json
automate-with-cron.md
automation-blueprints.md
aws-bedrock.md
azure-foundry.md
build-a-hermes-plugin.md
cron-script-only.md
cron-troubleshooting.md
daily-briefing-bot.md
delegation-patterns.md
github-pr-review-agent.md
google-gemini.md
google-vertex.md
local-llm-on-mac.md
local-ollama-setup.md
microsoft-graph-app-registration.md
migrate-from-openclaw.md
minimax-oauth.md
oauth-over-ssh.md
operate-teams-meeting-pipeline.md
pipe-script-output.md
python-library.md
run-hermes-with-nous-portal.md
run-nemotron-3-ultra-free.md
team-telegram-assistant.md
tips.md
use-mcp-with-hermes.md
use-soul-with-hermes.md
use-voice-mode-with-hermes.md
webhook-github-pr-review.md
work-with-skills.md
xai-grok-oauth.md