Go to file

Teknium 77c0bc6b13 fix(curator): defer first run and add --dry-run preview (#18373 ) (#18389 )

* fix(curator): defer first run and add --dry-run preview (#18373)

Curator was meant to run 7 days after install, not on the very first
gateway tick. On a fresh install (no .curator_state), should_run_now()
returned True immediately because last_run_at was None — so the gateway
cron ticker fired Curator against a fresh skill library moments after
'hermes update'. Combined with the binary 'agent-created' provenance
model (anything not bundled and not hub-installed), this consolidated
hand-authored user workflow skills without consent.

Changes:
- should_run_now(): first observation seeds last_run_at='now' and returns
  False. The next real pass fires one full interval_hours later (7 days
  by default), matching the original design intent.
- hermes curator run --dry-run: produces the same review report without
  applying automatic transitions OR permitting the LLM to call
  skill_manage / terminal mv. A DRY-RUN banner is prepended to the
  prompt and the caller skips apply_automatic_transitions. State is
  NOT advanced so a preview doesn't defer the next scheduled real pass.
- hermes update: prints a one-liner on fresh installs pointing at
  --dry-run, pause, and the docs. Silent on steady state.
- Docs: curator.md and cli-commands.md explain the deferred first-run
  behavior and warn that hand-written SKILL.md files share the
  'agent-created' bucket, with guidance to pin or preview before the
  first pass.

Tests:
- test_first_run_defers replaces the old 'first run always eligible'
  assertion — same fixture, inverted expectation.
- test_maybe_run_curator_defers_on_fresh_install covers the gateway tick
  path end-to-end.
- Three new dry-run tests cover state-advance suppression, prompt
  banner injection, and apply_automatic_transitions skipping.

Fixes #18373.

* feat(curator): pre-run backup + rollback (#18373)

Every real curator pass now snapshots ~/.hermes/skills/ into
~/.hermes/skills/.curator_backups/<utc-iso>/skills.tar.gz before calling
apply_automatic_transitions or the LLM review. If a run consolidates or
archives something the user didn't want touched, 'hermes curator
rollback' restores the tree in one command. Dry-run is skipped — no
mutation means no snapshot needed.

Changes:
- agent/curator_backup.py (new): tar.gz snapshot + safe rollback. The
  snapshot excludes .curator_backups/ (would recurse) and .hub/ (managed
  by the skills hub). Extract refuses absolute paths and .. components,
  and uses tarfile's filter='data' on Python 3.12+. Rollback takes a
  pre-rollback safety snapshot FIRST, stages the current tree into
  .rollback-staging-<ts>/ so the extract lands in an empty dir, and
  cleans the staging dir on success. A failed extract restores the
  staged contents.
- agent/curator.py: run_curator_review() calls curator_backup.
  snapshot_skills(reason='pre-curator-run') before apply_automatic_
  transitions. Best-effort — a failed snapshot logs at debug and the
  run continues (a transient disk issue shouldn't silently disable
  curator forever).
- hermes_cli/curator.py: new 'hermes curator backup' and 'hermes curator
  rollback' subcommands. rollback supports --list, --id <ts>, -y.
- hermes_cli/config.py: curator.backup.{enabled, keep} config block
  with sane defaults (enabled=true, keep=5).
- Docs: curator.md gets a 'Backups and rollback' section; cli-commands
  .md table gets the new rows.

Tests (new file tests/agent/test_curator_backup.py, 16 cases):
- snapshot creates tarball + manifest with correct counts
- snapshot excludes .curator_backups/ (recursion guard) and .hub/
- snapshot disabled via config returns None without creating anything
- snapshot uniquifies ids within the same second (-01 suffix)
- prune honors keep count, newest-first
- list_backups + _resolve_backup cover newest-default and unknown-id
- rollback restores a deleted skill with content intact
- rollback is itself undoable — safety snapshot shows up in list_backups
- rollback with no snapshots returns an error
- rollback refuses tarballs with absolute paths or .. components
- real curator runs take a 'pre-curator-run' snapshot; dry-runs do not

All curator tests: 210 passing locally.

2026-05-01 09:49:59 -07:00

.github

docs: publish llms.txt and llms-full.txt for agent-friendly ingestion (#18276 )

2026-04-30 23:17:14 -07:00

.plans

Merge PR #724 : feat: --yolo flag to bypass all approval prompts

2026-03-10 20:56:30 -07:00

acp_adapter

fix(compression): include system prompt + tool schemas in token estimates (#18265 )

2026-04-30 23:03:54 -07:00

acp_registry

feat: restore ACP server implementation from PR #949 (#1254 )

2026-03-14 00:09:05 -07:00

agent

fix(curator): defer first run and add --dry-run preview (#18373 ) (#18389 )

2026-05-01 09:49:59 -07:00

assets

Update banner image to new version

2026-02-25 11:53:44 -08:00

cron

fix(curator): rewrite cron job skill refs after consolidation (#18253 )

2026-04-30 23:04:50 -07:00

datagen-config-examples

feat: add WebResearchEnv RL environment for multi-step web research

2026-03-05 14:34:36 +00:00

docker

fix(docker): don't chown config.yaml after gosu drop (#15865 ) (#16096 )

2026-04-26 08:27:39 -07:00

docs

feat(kanban): durable multi-profile collaboration board (#17805 )

2026-04-30 13:36:47 -07:00

environments

refactor: remove remaining redundant local imports (comprehensive sweep)

2026-04-21 00:50:58 -07:00

gateway

fix(telegram): send seed message after creating DM topics (#18334 )

2026-05-01 15:21:56 +05:30

hermes_cli

fix(curator): defer first run and add --dry-run preview (#18373 ) (#18389 )

2026-05-01 09:49:59 -07:00

nix

fix(security): address CodeQL path-traversal and info-exposure findings

2026-04-30 20:29:37 -04:00

optional-skills

fix(paths): route achievements plugin + profile-tui through HERMES_HOME

2026-04-30 23:21:54 -07:00

packaging/homebrew

chore: prepare Hermes for Homebrew packaging (#4099 )

2026-03-30 17:34:43 -07:00

plans

fix(gemini): tighten native routing and streaming replay

2026-04-19 12:40:08 -07:00

plugins

fix: kanban button

2026-05-01 07:33:54 -04:00

scripts

fix(paths): route achievements plugin + profile-tui through HERMES_HOME

2026-04-30 23:21:54 -07:00

skills

feat(skills): add here.now as an optional skill

2026-04-30 19:48:15 -07:00

tests

fix(curator): defer first run and add --dry-run preview (#18373 ) (#18389 )

2026-05-01 09:49:59 -07:00

tinker-atropos @ 65f084ee80

Add tinker-atropos submodule and update RL training tools

2026-02-04 10:36:01 -08:00

tools

fix: coerce show_reasoning and guard_agent_created config bools

2026-04-30 20:40:46 -07:00

tui_gateway

fix: lazy session creation — defer DB row until first message (#18370 )

2026-05-01 18:39:12 +05:30

ui-tui

Merge pull request #18117 from NousResearch/austin/fix/model-selector

2026-05-01 05:30:05 -07:00

web

Merge pull request #18095 from NousResearch/austin/feat/plugins-page

2026-05-01 05:29:24 -07:00

website

fix(curator): defer first run and add --dry-run preview (#18373 ) (#18389 )

2026-05-01 09:49:59 -07:00

.dockerignore

fix: prevent tui rebuilding assets

2026-05-01 16:29:46 +10:00

.env.example

feat(teams): add Microsoft Teams platform adapter as a plugin

2026-04-30 01:19:34 -07:00

.envrc

nix: add tui lockfile update script

2026-04-10 00:46:37 -04:00

.gitattributes

feat: web UI dashboard for managing Hermes Agent (#8756 )

2026-04-12 22:26:28 -07:00

.gitignore

feat(providers): add GMI Cloud as a first-class API-key provider (#11955 )

2026-04-27 11:17:59 -07:00

.gitmodules

refactor: remove mini-swe-agent dependency — inline Docker/Modal backends (#2804 )

2026-03-24 07:30:25 -07:00

.mailmap

chore: add MestreY0d4-Uninter to AUTHOR_MAP and .mailmap

2026-04-15 15:03:28 -07:00

AGENTS.md

remove: BOOT.md built-in hook (#17093 )

2026-04-28 09:50:27 -07:00

batch_runner.py

fix: eliminate duplicate checkpoint entries and JSON-unsafe coercion

2026-04-24 14:32:21 -07:00

cli-config.yaml.example

fix(agent): make tool loop guardrails warning-first

2026-04-30 20:43:15 -07:00

cli.py

fix: lazy session creation — defer DB row until first message (#18370 )

2026-05-01 18:39:12 +05:30

constraints-termux.txt

feat: add tested Termux install path and EOF-aware gh auth

2026-04-09 16:24:53 -07:00

CONTRIBUTING.md

fix(tui): restore macOS copy behavior and theme polish (#17131 )

2026-04-28 18:47:14 -05:00

docker-compose.yml

fix(teams): pipe TEAMS_PORT through docker-compose properly

2026-04-30 19:43:32 -07:00

Dockerfile

fix: prevent tui rebuilding assets

2026-05-01 16:29:46 +10:00

flake.lock

fix nix build

2026-04-11 15:30:37 -04:00

flake.nix

feat(nix): declarative plugin installation for NixOS module (#15953 )

2026-04-28 00:18:32 +05:30

hermes

fix: use argparse entrypoint in top-level launcher (#3874 )

2026-03-29 21:54:36 -07:00

hermes_constants.py

Merge branch 'main' of github.com:NousResearch/hermes-agent into feat/ink-refactor

2026-04-13 21:17:41 -05:00

hermes_logging.py

fix(logging): attach gateway log after cli init

2026-04-26 19:01:26 -07:00

hermes_state.py

fix: lazy session creation — defer DB row until first message (#18370 )

2026-05-01 18:39:12 +05:30

hermes_time.py

refactor: extract shared helpers to deduplicate repeated code patterns (#7917 )

2026-04-11 13:59:52 -07:00

hermes-already-has-routines.md

docs: automation templates gallery + comparison post (#9821 )

2026-04-14 12:30:50 -07:00

LICENSE

fix: restore missing MIT license file

2026-03-07 13:43:08 -08:00

MANIFEST.in

chore: prepare Hermes for Homebrew packaging (#4099 )

2026-03-30 17:34:43 -07:00

mcp_serve.py

fix: point optional-dep install hints at the venv's python (#11938 )

2026-04-17 21:16:33 -07:00

mini_swe_runner.py

fix(kimi): omit temperature entirely for Kimi/Moonshot models (#13157 )

2026-04-20 12:23:05 -07:00

model_tools.py

fix(gateway): apply agent.disabled_toolsets in gateway message loop

2026-04-30 20:24:39 -07:00

package-lock.json

perf(browser): upgrade agent-browser 0.13 -> 0.26, wire daemon idle timeout

2026-04-22 16:33:36 -07:00

package.json

perf(browser): upgrade agent-browser 0.13 -> 0.26, wire daemon idle timeout

2026-04-22 16:33:36 -07:00

pyproject.toml

chore: release v0.12.0 (2026.4.30) (#18057 )

2026-04-30 11:31:01 -07:00

README.md

chore: revert docs

2026-04-26 05:46:45 -07:00

RELEASE_v0.2.0.md

chore: rebuild changelog with correct time window (Feb 25 12PM PST onwards)

2026-03-12 02:33:50 -07:00

RELEASE_v0.3.0.md

chore: release v0.3.0 (v2026.3.17)

2026-03-17 00:38:48 -07:00

RELEASE_v0.4.0.md

docs: revise v0.4.0 changelog — fix feature attribution, reorder sections

2026-03-23 22:42:22 -07:00

RELEASE_v0.5.0.md

chore: release v0.5.0 (v2026.3.28) (#3568 )

2026-03-28 13:11:39 -07:00

RELEASE_v0.6.0.md

chore: release v0.6.0 (2026.3.30) (#3985 )

2026-03-30 08:29:38 -07:00

RELEASE_v0.7.0.md

chore: release v0.7.0 (2026.4.3) (#4812 )

2026-04-03 11:14:55 -07:00

RELEASE_v0.8.0.md

docs: update v0.8.0 highlights — notify_on_complete, MiMo v2 Pro, reorder

2026-04-08 04:59:45 -07:00

RELEASE_v0.9.0.md

fix: add contributor audit script + fix missed contributors (#9264 )

2026-04-13 16:31:27 -07:00

RELEASE_v0.10.0.md

chore: release v0.10.0 (2026.4.16) (#11209 )

2026-04-16 12:53:06 -07:00

RELEASE_v0.11.0.md

chore: release v0.11.0 (2026.4.23) (#14791 )

2026-04-23 15:31:59 -07:00

RELEASE_v0.12.0.md

chore: release v0.12.0 (2026.4.30) (#18057 )

2026-04-30 11:31:01 -07:00

rl_cli.py

chore: remove unused imports and dead locals (ruff F401, F841) (#17010 )

2026-04-28 06:46:45 -07:00

run_agent.py

fix: lazy session creation — defer DB row until first message (#18370 )

2026-05-01 18:39:12 +05:30

SECURITY.md

docs: add terminal bypass test to Out of Scope section

2026-04-15 14:34:09 -07:00

setup-hermes.sh

fix(termux): make setup-hermes use android path

2026-04-09 16:24:53 -07:00

toolset_distributions.py

chore: fix 154 f-strings, simplify getattr/URL patterns, remove dead code (#3119 )

2026-03-25 19:47:58 -07:00

toolsets.py

feat(kanban): durable multi-profile collaboration board (#17805 )

2026-04-30 13:36:47 -07:00

trajectory_compressor.py

chore: remove unused imports and dead locals (ruff F401, F841) (#17010 )

2026-04-28 06:46:45 -07:00

utils.py

refactor: consolidate symlink-safe atomic replace into shared helper

2026-04-28 04:58:22 -07:00

uv.lock

fix: lazy session creation — defer DB row until first message (#18370 )

2026-05-01 18:39:12 +05:30

README.md

Hermes Agent ☤

The self-improving AI agent built by Nous Research. It's the only agent with a built-in learning loop — it creates skills from experience, improves them during use, nudges itself to persist knowledge, searches its own past conversations, and builds a deepening model of who you are across sessions. Run it on a $5 VPS, a GPU cluster, or serverless infrastructure that costs nearly nothing when idle. It's not tied to your laptop — talk to it from Telegram while it works on a cloud VM.

Use any model you want — Nous Portal, OpenRouter (200+ models), NVIDIA NIM (Nemotron), Xiaomi MiMo, z.ai/GLM, Kimi/Moonshot, MiniMax, Hugging Face, OpenAI, or your own endpoint. Switch with hermes model — no code changes, no lock-in.

A real terminal interface	Full TUI with multiline editing, slash-command autocomplete, conversation history, interrupt-and-redirect, and streaming tool output.
Lives where you do	Telegram, Discord, Slack, WhatsApp, Signal, and CLI — all from a single gateway process. Voice memo transcription, cross-platform conversation continuity.
A closed learning loop	Agent-curated memory with periodic nudges. Autonomous skill creation after complex tasks. Skills self-improve during use. FTS5 session search with LLM summarization for cross-session recall. Honcho dialectic user modeling. Compatible with the agentskills.io open standard.
Scheduled automations	Built-in cron scheduler with delivery to any platform. Daily reports, nightly backups, weekly audits — all in natural language, running unattended.
Delegates and parallelizes	Spawn isolated subagents for parallel workstreams. Write Python scripts that call tools via RPC, collapsing multi-step pipelines into zero-context-cost turns.
Runs anywhere, not just your laptop	Six terminal backends — local, Docker, SSH, Daytona, Singularity, and Modal. Daytona and Modal offer serverless persistence — your agent's environment hibernates when idle and wakes on demand, costing nearly nothing between sessions. Run it on a $5 VPS or a GPU cluster.
Research-ready	Batch trajectory generation, Atropos RL environments, trajectory compression for training the next generation of tool-calling models.

Quick Install

curl -fsSL https://raw.githubusercontent.com/NousResearch/hermes-agent/main/scripts/install.sh | bash

Works on Linux, macOS, WSL2, and Android via Termux. The installer handles the platform-specific setup for you.

Android / Termux: The tested manual path is documented in the Termux guide. On Termux, Hermes installs a curated .[termux] extra because the full .[all] extra currently pulls Android-incompatible voice dependencies.

Windows: Native Windows is not supported. Please install WSL2 and run the command above.

After installation:

source ~/.bashrc    # reload shell (or: source ~/.zshrc)
hermes              # start chatting!

Getting Started

hermes              # Interactive CLI — start a conversation
hermes model        # Choose your LLM provider and model
hermes tools        # Configure which tools are enabled
hermes config set   # Set individual config values
hermes gateway      # Start the messaging gateway (Telegram, Discord, etc.)
hermes setup        # Run the full setup wizard (configures everything at once)
hermes claw migrate # Migrate from OpenClaw (if coming from OpenClaw)
hermes update       # Update to the latest version
hermes doctor       # Diagnose any issues

📖 Full documentation →

CLI vs Messaging Quick Reference

Hermes has two entry points: start the terminal UI with hermes, or run the gateway and talk to it from Telegram, Discord, Slack, WhatsApp, Signal, or Email. Once you're in a conversation, many slash commands are shared across both interfaces.

Action	CLI	Messaging platforms
Start chatting	`hermes`	Run `hermes gateway setup` + `hermes gateway start`, then send the bot a message
Start fresh conversation	`/new` or `/reset`	`/new` or `/reset`
Change model	`/model [provider:model]`	`/model [provider:model]`
Set a personality	`/personality [name]`	`/personality [name]`
Retry or undo the last turn	`/retry`, `/undo`	`/retry`, `/undo`
Compress context / check usage	`/compress`, `/usage`, `/insights [--days N]`	`/compress`, `/usage`, `/insights [days]`
Browse skills	`/skills` or `/<skill-name>`	`/<skill-name>`
Interrupt current work	`Ctrl+C` or send a new message	`/stop` or send a new message
Platform-specific status	`/platforms`	`/status`, `/sethome`

For the full command lists, see the CLI guide and the Messaging Gateway guide.

Documentation

All documentation lives at hermes-agent.nousresearch.com/docs:

Section	What's Covered
Quickstart	Install → setup → first conversation in 2 minutes
CLI Usage	Commands, keybindings, personalities, sessions
Configuration	Config file, providers, models, all options
Messaging Gateway	Telegram, Discord, Slack, WhatsApp, Signal, Home Assistant
Security	Command approval, DM pairing, container isolation
Tools & Toolsets	40+ tools, toolset system, terminal backends
Skills System	Procedural memory, Skills Hub, creating skills
Memory	Persistent memory, user profiles, best practices
MCP Integration	Connect any MCP server for extended capabilities
Cron Scheduling	Scheduled tasks with platform delivery
Context Files	Project context that shapes every conversation
Architecture	Project structure, agent loop, key classes
Contributing	Development setup, PR process, code style
CLI Reference	All commands and flags
Environment Variables	Complete env var reference

Migrating from OpenClaw

If you're coming from OpenClaw, Hermes can automatically import your settings, memories, skills, and API keys.

During first-time setup: The setup wizard (hermes setup) automatically detects ~/.openclaw and offers to migrate before configuration begins.

Anytime after install:

hermes claw migrate              # Interactive migration (full preset)
hermes claw migrate --dry-run    # Preview what would be migrated
hermes claw migrate --preset user-data   # Migrate without secrets
hermes claw migrate --overwrite  # Overwrite existing conflicts

What gets imported:

SOUL.md — persona file
Memories — MEMORY.md and USER.md entries
Skills — user-created skills → ~/.hermes/skills/openclaw-imports/
Command allowlist — approval patterns
Messaging settings — platform configs, allowed users, working directory
API keys — allowlisted secrets (Telegram, OpenRouter, OpenAI, Anthropic, ElevenLabs)
TTS assets — workspace audio files
Workspace instructions — AGENTS.md (with --workspace-target)

See hermes claw migrate --help for all options, or use the openclaw-migration skill for an interactive agent-guided migration with dry-run previews.

Contributing

We welcome contributions! See the Contributing Guide for development setup, code style, and PR process.

Quick start for contributors — clone and go with setup-hermes.sh:

git clone https://github.com/NousResearch/hermes-agent.git
cd hermes-agent
./setup-hermes.sh     # installs uv, creates venv, installs .[all], symlinks ~/.local/bin/hermes
./hermes              # auto-detects the venv, no need to `source` first

Manual path (equivalent to the above):

curl -LsSf https://astral.sh/uv/install.sh | sh
uv venv venv --python 3.11
source venv/bin/activate
uv pip install -e ".[all,dev]"
scripts/run_tests.sh

RL Training (optional): The RL/Atropos integration (environments/) ships via the atroposlib and tinker dependencies pulled in by .[all,dev] — no submodule setup required.

Community

💬 Discord
📚 Skills Hub
🐛 Issues
🔌 HermesClaw — Community WeChat bridge: Run Hermes Agent and OpenClaw on the same WeChat account.

License

MIT — see LICENSE.

Built by Nous Research.

Languages

Python 84.1%

TypeScript 12.3%

JavaScript 1%

TeX 0.9%

Shell 0.5%

Other 1.1%