mirror of https://github.com/NousResearch/hermes-agent.git synced 2026-04-28 06:51:16 +08:00

Go to file

Brian D. Evans e87a2100f6 fix(mcp): auto-reconnect + retry once when the transport session expires (#13383 )

Streamable HTTP MCP servers may garbage-collect their server-side
session state while the OAuth token remains valid — idle TTL, server
restart, pod rotation, etc.  Before this fix, the tool-call handler
treated the resulting "Invalid or expired session" error as a plain
tool failure with no recovery path, so **every subsequent call on
the affected server failed until the gateway was manually
restarted**.  Reporter: #13383.

The OAuth-based recovery path (``_handle_auth_error_and_retry``)
already exists for 401s, but it only fires on auth errors.  Session
expiry slipped through because the access token is still valid —
nothing 401'd, so the existing recovery branch was skipped.

Fix
---
Add a sibling function ``_handle_session_expired_and_retry`` that
detects MCP session-expiry via ``_is_session_expired_error`` (a
narrow allow-list of known-stable substrings: ``"invalid or expired
session"``, ``"session expired"``, ``"session not found"``,
``"unknown session"``, etc.) and then uses the existing transport
reconnect mechanism:

* Sets ``MCPServerTask._reconnect_event`` — the server task's
  lifecycle loop already interprets this as "tear down the current
  ``streamablehttp_client`` + ``ClientSession`` and rebuild them,
  reusing the existing OAuth provider instance".
* Waits up to 15 s for the new session to come back ready.
* Retries the original call once.  If the retry succeeds, returns
  its result and resets the circuit-breaker error count.  If the
  retry raises, or if the reconnect doesn't ready in time, falls
  through to the caller's generic error path.

Unlike the 401 path, this does **not** call ``handle_401`` — the
access token is already valid and running an OAuth refresh would be
a pointless round-trip.

All 5 MCP handlers (``call_tool``, ``list_resources``, ``read_resource``,
``list_prompts``, ``get_prompt``) now consult both recovery paths
before falling through:

    recovered = _handle_auth_error_and_retry(...)          # 401 path
    if recovered is not None: return recovered
    recovered = _handle_session_expired_and_retry(...)     # new
    if recovered is not None: return recovered
    # generic error response

Narrow scope — explicitly not changed
-------------------------------------
* **Detection is string-based on a 5-entry allow-list.**  The MCP
  SDK wraps JSON-RPC errors in ``McpError`` whose exception type +
  attributes vary across SDK versions, so matching on message
  substrings is the durable path.  Kept narrow to avoid false
  positives — a regular ``RuntimeError("Tool failed")`` will NOT
  trigger spurious reconnects (pinned by
  ``test_is_session_expired_rejects_unrelated_errors``).
* **No change to the existing 401 recovery flow.**  The new path is
  consulted only after the auth path declines (returns ``None``).
* **Retry count stays at 1.**  If the reconnect-then-retry also
  fails, we don't loop — the error surfaces normally so the model
  sees a failed tool call rather than a hang.
* **``InterruptedError`` is explicitly excluded** from session-expired
  detection so user-cancel signals always short-circuit the same
  way they did before (pinned by
  ``test_is_session_expired_rejects_interrupted_error``).

Regression coverage
-------------------
``tests/tools/test_mcp_tool_session_expired.py`` (new, 16 cases):

Unit tests for ``_is_session_expired_error``:
* ``test_is_session_expired_detects_invalid_or_expired_session`` —
  reporter's exact wpcom-mcp text.
* ``test_is_session_expired_detects_expired_session_variant`` —
  "Session expired" / "expired session" variants.
* ``test_is_session_expired_detects_session_not_found`` — server GC
  variant ("session not found", "unknown session").
* ``test_is_session_expired_is_case_insensitive``.
* ``test_is_session_expired_rejects_unrelated_errors`` — narrow-scope
  canary: random RuntimeError / ValueError / 401 don't trigger.
* ``test_is_session_expired_rejects_interrupted_error`` — user cancel
  must never route through reconnect.
* ``test_is_session_expired_rejects_empty_message``.

Handler integration tests:
* ``test_call_tool_handler_reconnects_on_session_expired`` — reporter's
  full repro: first call raises "Invalid or expired session", handler
  signals ``_reconnect_event``, retries once, returns the retry's
  success result with no ``error`` key.
* ``test_call_tool_handler_non_session_expired_error_falls_through``
  — preserved-behaviour canary: random tool failures do NOT trigger
  reconnect.
* ``test_session_expired_handler_returns_none_without_loop`` —
  defensive: cold-start / shutdown race.
* ``test_session_expired_handler_returns_none_without_server_record``
  — torn-down server falls through cleanly.
* ``test_session_expired_handler_returns_none_when_retry_also_fails``
  — no retry loop on repeated failure.

Parametrised across all 4 non-``tools/call`` handlers:
* ``test_non_tool_handlers_also_reconnect_on_session_expired``
  [list_resources / read_resource / list_prompts / get_prompt].

**15 of 16 fail on clean ``origin/main`` (``6fb69229``)** with
``ImportError: cannot import name '_is_session_expired_error'``
— the fix's surface symbols don't exist there yet.  The 1 passing
test is an ordering artefact of pytest-xdist worker collection.

Validation
----------
``source venv/bin/activate && python -m pytest
tests/tools/test_mcp_tool_session_expired.py -q`` → **16 passed**.

Broader MCP suite (5 files:
``test_mcp_tool.py``, ``test_mcp_tool_401_handling.py``,
``test_mcp_tool_session_expired.py``, ``test_mcp_reconnect_signal.py``,
``test_mcp_oauth.py``) → **230 passed, 0 regressions**.

Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>

2026-04-24 05:28:45 -07:00

.github

docs(website): dedicated page per bundled + optional skill (#14929 )

2026-04-23 22:22:11 -07:00

.plans

Merge PR #724 : feat: --yolo flag to bypass all approval prompts

2026-03-10 20:56:30 -07:00

acp_adapter

fix(acp): include MCP toolsets in ACP sessions

2026-04-24 03:04:42 -07:00

acp_registry

feat: restore ACP server implementation from PR #949 (#1254 )

2026-03-14 00:09:05 -07:00

agent

fix(credential_pool): add Nous OAuth cross-process auth-store sync

2026-04-24 05:20:05 -07:00

assets

Update banner image to new version

2026-02-25 11:53:44 -08:00

cron

feat(cron): per-job workdir for project-aware cron runs (#15110 )

2026-04-24 05:07:01 -07:00

datagen-config-examples

feat: add WebResearchEnv RL environment for multi-step web research

2026-03-05 14:34:36 +00:00

docker

fix(docker): fix HERMES_UID permission handling and add docker-compose.yml

2026-04-24 04:52:11 -07:00

environments

refactor: remove remaining redundant local imports (comprehensive sweep)

2026-04-21 00:50:58 -07:00

gateway

fix(model): repair Discord Copilot /model flow

2026-04-24 03:33:29 -07:00

hermes_cli

fix(model-normalize): pass DeepSeek V-series IDs through instead of folding to deepseek-chat

2026-04-24 05:24:54 -07:00

nix

fix(nix): make working directory writable

2026-04-23 02:06:16 -07:00

optional-skills

feat(optional-skills): add page-agent skill under new web-development category (#13976 )

2026-04-22 04:54:26 -07:00

packaging/homebrew

chore: prepare Hermes for Homebrew packaging (#4099 )

2026-03-30 17:34:43 -07:00

plans

fix(gemini): tighten native routing and streaming replay

2026-04-19 12:40:08 -07:00

plugins

feat(hindsight): optional bank_id_template for per-agent / per-user banks

2026-04-24 03:38:17 -07:00

scripts

chore(spotify): gate toolset off by default, add to hermes tools UI

2026-04-24 05:20:38 -07:00

skills

refactor(commands): drop /provider, /plan handler, and clean up slash registry (#15047 )

2026-04-24 03:10:52 -07:00

tests

fix(mcp): auto-reconnect + retry once when the transport session expires (#13383 )

2026-04-24 05:28:45 -07:00

tinker-atropos @ 65f084ee80

Add tinker-atropos submodule and update RL training tools

2026-02-04 10:36:01 -08:00

tools

fix(mcp): auto-reconnect + retry once when the transport session expires (#13383 )

2026-04-24 05:28:45 -07:00

tui_gateway

refactor(commands): drop /provider, /plan handler, and clean up slash registry (#15047 )

2026-04-24 03:10:52 -07:00

ui-tui

refactor(commands): drop /provider, /plan handler, and clean up slash registry (#15047 )

2026-04-24 03:10:52 -07:00

web

feat(dashboard): reskin extension points for themes and plugins (#14776 )

2026-04-23 15:31:01 -07:00

website

feat(copilot): add 401 auth recovery with automatic token refresh and client rebuild

2026-04-24 05:09:08 -07:00

.dockerignore

fix(docker): exclude runtime data/ from build context

2026-04-22 21:15:28 -07:00

.env.example

feat: add Ollama Cloud as built-in provider

2026-04-16 02:22:09 -07:00

.envrc

nix: add tui lockfile update script

2026-04-10 00:46:37 -04:00

.gitattributes

feat: web UI dashboard for managing Hermes Agent (#8756 )

2026-04-12 22:26:28 -07:00

.gitignore

Update .gitignore

2026-04-22 20:02:46 -07:00

.gitmodules

refactor: remove mini-swe-agent dependency — inline Docker/Modal backends (#2804 )

2026-03-24 07:30:25 -07:00

.mailmap

chore: add MestreY0d4-Uninter to AUTHOR_MAP and .mailmap

2026-04-15 15:03:28 -07:00

AGENTS.md

docs(agents): refresh AGENTS.md — fix stale facts, expand plugins/skills sections (#14763 )

2026-04-23 15:13:13 -07:00

batch_runner.py

refactor: remove remaining redundant local imports (comprehensive sweep)

2026-04-21 00:50:58 -07:00

cli-config.yaml.example

docs: document prompt_caching.cache_ttl in cli-config example

2026-04-24 03:21:29 -07:00

cli.py

fix(codex): route auth failures to fallback provider chain

2026-04-24 04:53:32 -07:00

constraints-termux.txt

feat: add tested Termux install path and EOF-aware gh auth

2026-04-09 16:24:53 -07:00

CONTRIBUTING.md

Update CONTRIBUTING.md

2026-04-23 15:08:41 -07:00

docker-compose.yml

fix(docker): safer docker-compose defaults for UID and dashboard bind

2026-04-24 04:52:11 -07:00

Dockerfile

fix(docker): reap orphaned subprocesses via tini as PID 1 (#15116 )

2026-04-24 05:22:34 -07:00

flake.lock

fix nix build

2026-04-11 15:30:37 -04:00

flake.nix

nix: add tui lockfile update script

2026-04-10 00:46:37 -04:00

hermes

fix: use argparse entrypoint in top-level launcher (#3874 )

2026-03-29 21:54:36 -07:00

hermes_constants.py

Merge branch 'main' of github.com:NousResearch/hermes-agent into feat/ink-refactor

2026-04-13 21:17:41 -05:00

hermes_logging.py

fix: detect and strip non-ASCII characters from API keys (#6843 )

2026-04-14 20:20:31 -07:00

hermes_state.py

fix(resume): redirect --resume to the descendant that actually holds the messages

2026-04-24 03:04:42 -07:00

hermes_time.py

refactor: extract shared helpers to deduplicate repeated code patterns (#7917 )

2026-04-11 13:59:52 -07:00

hermes-already-has-routines.md

docs: automation templates gallery + comparison post (#9821 )

2026-04-14 12:30:50 -07:00

LICENSE

fix: restore missing MIT license file

2026-03-07 13:43:08 -08:00

MANIFEST.in

chore: prepare Hermes for Homebrew packaging (#4099 )

2026-03-30 17:34:43 -07:00

mcp_serve.py

fix: point optional-dep install hints at the venv's python (#11938 )

2026-04-17 21:16:33 -07:00

mini_swe_runner.py

fix(kimi): omit temperature entirely for Kimi/Moonshot models (#13157 )

2026-04-20 12:23:05 -07:00

model_tools.py

fix: sanitize tool schemas for llama.cpp backends; restore MCP in TUI (#15032 )

2026-04-24 02:44:46 -07:00

package-lock.json

perf(browser): upgrade agent-browser 0.13 -> 0.26, wire daemon idle timeout

2026-04-22 16:33:36 -07:00

package.json

perf(browser): upgrade agent-browser 0.13 -> 0.26, wire daemon idle timeout

2026-04-22 16:33:36 -07:00

pyproject.toml

chore: release v0.11.0 (2026.4.23) (#14791 )

2026-04-23 15:31:59 -07:00

README.md

docs(readme): fix stale RL submodule instructions, skills table row, test runner (#14758 )

2026-04-23 15:12:04 -07:00

RELEASE_v0.2.0.md

chore: rebuild changelog with correct time window (Feb 25 12PM PST onwards)

2026-03-12 02:33:50 -07:00

RELEASE_v0.3.0.md

chore: release v0.3.0 (v2026.3.17)

2026-03-17 00:38:48 -07:00

RELEASE_v0.4.0.md

docs: revise v0.4.0 changelog — fix feature attribution, reorder sections

2026-03-23 22:42:22 -07:00

RELEASE_v0.5.0.md

chore: release v0.5.0 (v2026.3.28) (#3568 )

2026-03-28 13:11:39 -07:00

RELEASE_v0.6.0.md

chore: release v0.6.0 (2026.3.30) (#3985 )

2026-03-30 08:29:38 -07:00

RELEASE_v0.7.0.md

chore: release v0.7.0 (2026.4.3) (#4812 )

2026-04-03 11:14:55 -07:00

RELEASE_v0.8.0.md

docs: update v0.8.0 highlights — notify_on_complete, MiMo v2 Pro, reorder

2026-04-08 04:59:45 -07:00

RELEASE_v0.9.0.md

fix: add contributor audit script + fix missed contributors (#9264 )

2026-04-13 16:31:27 -07:00

RELEASE_v0.10.0.md

chore: release v0.10.0 (2026.4.16) (#11209 )

2026-04-16 12:53:06 -07:00

RELEASE_v0.11.0.md

chore: release v0.11.0 (2026.4.23) (#14791 )

2026-04-23 15:31:59 -07:00

rl_cli.py

refactor: consolidate get_hermes_home() and parse_reasoning_effort() (#3062 )

2026-03-25 15:54:28 -07:00

run_agent.py

fix(agent): fall back on rate limit when pool has no rotation room

2026-04-24 05:20:05 -07:00

SECURITY.md

docs: add terminal bypass test to Out of Scope section

2026-04-15 14:34:09 -07:00

setup-hermes.sh

fix(termux): make setup-hermes use android path

2026-04-09 16:24:53 -07:00

toolset_distributions.py

chore: fix 154 f-strings, simplify getattr/URL patterns, remove dead code (#3119 )

2026-03-25 19:47:58 -07:00

toolsets.py

chore(spotify): gate toolset off by default, add to hermes tools UI

2026-04-24 05:20:38 -07:00

trajectory_compressor.py

fix: sweep remaining provider-URL substring checks across codebase

2026-04-20 22:14:29 -07:00

utils.py

fix(agent): normalize socks:// env proxies for httpx/anthropic

2026-04-21 05:52:46 -07:00

uv.lock

chore(dev): add ruff linter to dev deps and configure in pyproject.toml (#14527 )

2026-04-23 17:20:18 +05:30

README.md

Hermes Agent ☤

The self-improving AI agent built by Nous Research. It's the only agent with a built-in learning loop — it creates skills from experience, improves them during use, nudges itself to persist knowledge, searches its own past conversations, and builds a deepening model of who you are across sessions. Run it on a $5 VPS, a GPU cluster, or serverless infrastructure that costs nearly nothing when idle. It's not tied to your laptop — talk to it from Telegram while it works on a cloud VM.

Use any model you want — Nous Portal, OpenRouter (200+ models), NVIDIA NIM (Nemotron), Xiaomi MiMo, z.ai/GLM, Kimi/Moonshot, MiniMax, Hugging Face, OpenAI, or your own endpoint. Switch with hermes model — no code changes, no lock-in.

A real terminal interface	Full TUI with multiline editing, slash-command autocomplete, conversation history, interrupt-and-redirect, and streaming tool output.
Lives where you do	Telegram, Discord, Slack, WhatsApp, Signal, and CLI — all from a single gateway process. Voice memo transcription, cross-platform conversation continuity.
A closed learning loop	Agent-curated memory with periodic nudges. Autonomous skill creation after complex tasks. Skills self-improve during use. FTS5 session search with LLM summarization for cross-session recall. Honcho dialectic user modeling. Compatible with the agentskills.io open standard.
Scheduled automations	Built-in cron scheduler with delivery to any platform. Daily reports, nightly backups, weekly audits — all in natural language, running unattended.
Delegates and parallelizes	Spawn isolated subagents for parallel workstreams. Write Python scripts that call tools via RPC, collapsing multi-step pipelines into zero-context-cost turns.
Runs anywhere, not just your laptop	Six terminal backends — local, Docker, SSH, Daytona, Singularity, and Modal. Daytona and Modal offer serverless persistence — your agent's environment hibernates when idle and wakes on demand, costing nearly nothing between sessions. Run it on a $5 VPS or a GPU cluster.
Research-ready	Batch trajectory generation, Atropos RL environments, trajectory compression for training the next generation of tool-calling models.

Quick Install

curl -fsSL https://raw.githubusercontent.com/NousResearch/hermes-agent/main/scripts/install.sh | bash

Works on Linux, macOS, WSL2, and Android via Termux. The installer handles the platform-specific setup for you.

Android / Termux: The tested manual path is documented in the Termux guide. On Termux, Hermes installs a curated .[termux] extra because the full .[all] extra currently pulls Android-incompatible voice dependencies.

Windows: Native Windows is not supported. Please install WSL2 and run the command above.

After installation:

source ~/.bashrc    # reload shell (or: source ~/.zshrc)
hermes              # start chatting!

Getting Started

hermes              # Interactive CLI — start a conversation
hermes model        # Choose your LLM provider and model
hermes tools        # Configure which tools are enabled
hermes config set   # Set individual config values
hermes gateway      # Start the messaging gateway (Telegram, Discord, etc.)
hermes setup        # Run the full setup wizard (configures everything at once)
hermes claw migrate # Migrate from OpenClaw (if coming from OpenClaw)
hermes update       # Update to the latest version
hermes doctor       # Diagnose any issues

📖 Full documentation →

CLI vs Messaging Quick Reference

Hermes has two entry points: start the terminal UI with hermes, or run the gateway and talk to it from Telegram, Discord, Slack, WhatsApp, Signal, or Email. Once you're in a conversation, many slash commands are shared across both interfaces.

Action	CLI	Messaging platforms
Start chatting	`hermes`	Run `hermes gateway setup` + `hermes gateway start`, then send the bot a message
Start fresh conversation	`/new` or `/reset`	`/new` or `/reset`
Change model	`/model [provider:model]`	`/model [provider:model]`
Set a personality	`/personality [name]`	`/personality [name]`
Retry or undo the last turn	`/retry`, `/undo`	`/retry`, `/undo`
Compress context / check usage	`/compress`, `/usage`, `/insights [--days N]`	`/compress`, `/usage`, `/insights [days]`
Browse skills	`/skills` or `/<skill-name>`	`/<skill-name>`
Interrupt current work	`Ctrl+C` or send a new message	`/stop` or send a new message
Platform-specific status	`/platforms`	`/status`, `/sethome`

For the full command lists, see the CLI guide and the Messaging Gateway guide.

Documentation

All documentation lives at hermes-agent.nousresearch.com/docs:

Section	What's Covered
Quickstart	Install → setup → first conversation in 2 minutes
CLI Usage	Commands, keybindings, personalities, sessions
Configuration	Config file, providers, models, all options
Messaging Gateway	Telegram, Discord, Slack, WhatsApp, Signal, Home Assistant
Security	Command approval, DM pairing, container isolation
Tools & Toolsets	40+ tools, toolset system, terminal backends
Skills System	Procedural memory, Skills Hub, creating skills
Memory	Persistent memory, user profiles, best practices
MCP Integration	Connect any MCP server for extended capabilities
Cron Scheduling	Scheduled tasks with platform delivery
Context Files	Project context that shapes every conversation
Architecture	Project structure, agent loop, key classes
Contributing	Development setup, PR process, code style
CLI Reference	All commands and flags
Environment Variables	Complete env var reference

Migrating from OpenClaw

If you're coming from OpenClaw, Hermes can automatically import your settings, memories, skills, and API keys.

During first-time setup: The setup wizard (hermes setup) automatically detects ~/.openclaw and offers to migrate before configuration begins.

Anytime after install:

hermes claw migrate              # Interactive migration (full preset)
hermes claw migrate --dry-run    # Preview what would be migrated
hermes claw migrate --preset user-data   # Migrate without secrets
hermes claw migrate --overwrite  # Overwrite existing conflicts

What gets imported:

SOUL.md — persona file
Memories — MEMORY.md and USER.md entries
Skills — user-created skills → ~/.hermes/skills/openclaw-imports/
Command allowlist — approval patterns
Messaging settings — platform configs, allowed users, working directory
API keys — allowlisted secrets (Telegram, OpenRouter, OpenAI, Anthropic, ElevenLabs)
TTS assets — workspace audio files
Workspace instructions — AGENTS.md (with --workspace-target)

See hermes claw migrate --help for all options, or use the openclaw-migration skill for an interactive agent-guided migration with dry-run previews.

Contributing

We welcome contributions! See the Contributing Guide for development setup, code style, and PR process.

Quick start for contributors — clone and go with setup-hermes.sh:

git clone https://github.com/NousResearch/hermes-agent.git
cd hermes-agent
./setup-hermes.sh     # installs uv, creates venv, installs .[all], symlinks ~/.local/bin/hermes
./hermes              # auto-detects the venv, no need to `source` first

Manual path (equivalent to the above):

curl -LsSf https://astral.sh/uv/install.sh | sh
uv venv venv --python 3.11
source venv/bin/activate
uv pip install -e ".[all,dev]"
scripts/run_tests.sh

RL Training (optional): The RL/Atropos integration (environments/) ships via the atroposlib and tinker dependencies pulled in by .[all,dev] — no submodule setup required.

Community

💬 Discord
📚 Skills Hub
🐛 Issues
🔌 HermesClaw — Community WeChat bridge: Run Hermes Agent and OpenClaw on the same WeChat account.

License

MIT — see LICENSE.

Built by Nous Research.

Languages

Python 88.1%

TypeScript 8.9%

TeX 1.7%

Shell 0.5%

Nix 0.3%

Other 0.5%