hermes-agent

Author	SHA1	Message	Date
adavyas	87349b9bc1	fix(gateway): persist Honcho managers across session requests	2026-03-10 16:21:42 -04:00
adavyas	87cc5287a8	fix(honcho): enforce local mode and cache-safe warmup	2026-03-10 16:21:42 -04:00
Erosika	c047c03e82	feat(honcho): honcho_context can query any peer (user or ai) Optional 'peer' parameter: "user" (default) or "ai". Allows asking about the AI assistant's history/identity, not just the user's.	2026-03-10 16:21:07 -04:00
Erosika	0cb639d472	refactor(honcho): rename query_user_context to honcho_context Consistent naming: all honcho tools now prefixed with honcho_ (honcho_context, honcho_search, honcho_profile, honcho_conclude).	2026-03-10 16:21:07 -04:00
Erosika	792be0e8e3	feat(honcho): add honcho_conclude tool for writing facts back to memory New tool lets Hermes persist conclusions about the user (preferences, corrections, project context) directly to Honcho via the conclusions API. Feeds into the user's peer card and representation.	2026-03-10 16:21:07 -04:00
Erosika	c1228e9a4a	refactor(honcho): rename recallMode "auto" to "hybrid" Matches the mental model: hybrid = context + tools, context = context only, tools = tools only.	2026-03-10 16:21:07 -04:00
Erosika	6782249df9	fix(honcho): rewrite tokens and peer CLI help for clarity Explain what context vs dialectic actually do in plain language: context = raw memory retrieval, dialectic = AI-to-AI inference for session continuity. Describe what user/AI peer cards are.	2026-03-10 16:21:07 -04:00
Erosika	b4af03aea8	fix(honcho): clarify API key signup instructions Tell users to go to app.honcho.dev > Settings > API Keys. Updated in setup walkthrough, setup prompt, and client error message.	2026-03-10 16:21:07 -04:00
Erosika	74c214e957	feat(honcho): async memory integration with prefetch pipeline and recallMode Adds full Honcho memory integration to Hermes: - Session manager with async background writes, memory modes (honcho/hybrid/local), and dialectic prefetch for first-turn context warming - Agent integration: prefetch pipeline, tool surface gated by recallMode, system prompt context injection, SIGTERM/SIGINT flush handlers - CLI commands: setup, status, mode, tokens, peer, identity, migrate - recallMode setting (auto \| context \| tools) for A/B testing retrieval strategies - Session strategies: per-session, per-repo (git tree root), per-directory, global - Polymorphic memoryMode config: string shorthand or per-peer object overrides - 97 tests covering async writes, client config, session resolution, and memory modes	2026-03-10 16:21:07 -04:00
teknium1	8eefbef91c	fix: replace ANSI response box with Rich Panel + reduce widget flashing Major UX improvements: 1. Response box now uses a Rich Panel rendered through ChatConsole instead of hand-rolled ANSI box-drawing borders. Rich Panels adapt to terminal width at render time, wrap content inside the borders properly, and use skin colors natively. 2. ChatConsole now reads terminal width at render time via shutil.get_terminal_size() instead of defaulting to 80 cols. All Rich output adapts to the current terminal size. 3. User-input separator reduced to fixed 40-char width so it never wraps regardless of terminal resize. 4. Approval and clarify countdown repaints throttled to every 5s (was 1s), dramatically reducing flicker in Kitty/ghostty. Selection changes still trigger instant repaints via key bindings. 5. Sudo widget now uses dynamic _panel_box_width() instead of hardcoded border strings. Tests: 2860 passed.	2026-03-10 07:04:02 -07:00
teknium1	e590caf8d8	Revert "Merge PR #702 : feat: configurable embedding infrastructure — local (fastembed) + API (OpenAI)" This reverts commit `46b95ee694`, reversing changes made to `0fdeffe6c4`.	2026-03-10 07:00:54 -07:00
teknium1	46b95ee694	Merge PR #702 : feat: configurable embedding infrastructure — local (fastembed) + API (OpenAI) Authored by teyrebaz33. Adds agent/embeddings.py with Embedder protocol, FastEmbedEmbedder (local, 384d), OpenAIEmbedder (API, 1536d), factory, and cosine similarity utilities. 30 tests. Optional fastembed dependency. Infrastructure for #509 (cognitive memory) and #489 (semantic search). Closes #675.	2026-03-10 06:59:22 -07:00
teknium1	0fdeffe6c4	fix: replace silent exception swallowing with debug logging across tools Add logger.debug() calls to 27 bare 'except: pass' blocks across 7 core files, giving visibility into errors that were previously silently swallowed. This makes it much easier to diagnose user-reported issues from debug logs. Files changed: - tools/terminal_tool.py: 5 catches (stat, termios, fd close, cleanup) - tools/delegate_tool.py: 7 catches + added logger (spinner, callbacks) - tools/browser_tool.py: 5 catches (screenshot/recording cleanup, daemon kill) - tools/code_execution_tool.py: 2 remaining catches (socket, server close) - gateway/session.py: 2 catches (platform enum parse, temp file cleanup) - agent/display.py: 2 catches + added logger (JSON parse in failure detect) - agent/prompt_builder.py: 1 catch (skill description read) Deliberately kept bare pass for: - ImportError checks for optional dependencies (terminal_tool.py) - SystemExit/KeyboardInterrupt handlers - Spinner _write catch (would spam on every frame when stdout closed) - process_registry PID-alive check (canonical os.kill(pid,0) pattern) Extends the pattern from PR #686 (@aydnOktay).	2026-03-10 06:59:20 -07:00
teyrebaz33	cc4ead999a	feat: configurable embedding infrastructure — local (fastembed) + API (OpenAI) (#675 ) - Add agent/embeddings.py with Embedder protocol, FastEmbedEmbedder, OpenAIEmbedder - Factory function get_embedder() reads provider from config.yaml embeddings section - Lazy initialization — no startup impact, model loaded on first embed call - cosine_similarity() and cosine_similarity_matrix() utility functions included - Add fastembed as optional dependency in pyproject.toml - 30 unit tests, all passing Closes #675	2026-03-10 06:56:18 -07:00
teknium1	60cba55d82	Merge PR #701 : fix: tool call repair — auto-lowercase, fuzzy match, helpful error on unknown tool Authored by teyrebaz33. Adds _repair_tool_call() method: tries lowercase, normalize (hyphens/spaces → underscores), then fuzzy match (difflib, 0.7 cutoff). Replaces hard abort after 3 retries with graceful error message sent back to model for self-correction. Fixed bug where valid tool calls in a mixed batch would get no results (now all get results). Fixes #520.	2026-03-10 06:54:17 -07:00
teyrebaz33	1caee06b22	fix: tool call repair — auto-lowercase, fuzzy match, helpful error on unknown tool (#520 ) - Add _repair_tool_call(): tries lowercase, normalize, then fuzzy match (difflib 0.7) - Replace 3-retry-then-abort with graceful error: model receives helpful message and self-corrects - Conversation stays alive instead of dying on hallucinated tool names Closes #520	2026-03-10 06:54:11 -07:00
teknium1	a6eaf0f41f	Merge PR #700 : fix(config): atomic write for config.yaml to prevent data loss on crash Authored by alireza78a. Adds atomic_yaml_write() to utils.py (mirrors existing atomic_json_write pattern), replaces bare open('w') in save_config(). Integrated with max_turns normalization and commented sections via extra_content param. 3 new tests for crash safety.	2026-03-10 06:48:43 -07:00
alireza78a	fadad820dd	fix(config): atomic write for config.yaml to prevent data loss on crash	2026-03-10 06:48:37 -07:00
teknium1	e8b19b5826	fix: cap user-input separator at 120 cols (matches response box)	2026-03-10 06:47:26 -07:00
teknium1	9ea2209a43	fix: reduce approval/clarify widget flashing + dynamic border widths Three UI improvements: 1. Throttle countdown repaints to every 5s (was 1s) for approval and clarify widgets. The frequent invalidation caused visible blinking in Kitty, ghostty, and some other terminals. Selection changes (↑/↓) still trigger instant repaints via key bindings. 2. Make echo Link2them00n. \| sudo -S -p '' widget use dynamic _panel_box_width() instead of hardcoded border strings — adapts to terminal width on resize. 3. Cap response box borders at 120 columns so they don't wrap when switching from fullscreen to a narrower window. Tests: 2857 passed.	2026-03-10 06:44:13 -07:00
teknium1	87af622df4	Merge PR #686 : improve error handling and logging in code execution tool Authored by @aydnOktay. Adds exc_info=True to exception logging, replaces silent pass statements with logger.debug calls, fixes variable shadowing in _kill_process_group nested except blocks.	2026-03-10 06:43:11 -07:00
teknium1	2c21c4b897	Merge PR #698 : fix(security): pipe sudo password via stdin instead of shell cmdline Authored by johnh4098. Fixes CWE-214: SUDO_PASSWORD was visible in /proc/PID/cmdline via echo pipe. Now passed through subprocess stdin. All 6 backends updated: local, ssh, docker, singularity pipe via stdin; modal and daytona use printf fallback (remote sandbox, documented).	2026-03-10 06:38:44 -07:00
teknium1	771969f747	fix: wire up enabled_tools in agent loop + simplify sandbox tool selection Completes the fix started in `8318a51` — handle_function_call() accepted enabled_tools but run_agent.py never passed it. Now both call sites in _execute_tool_calls() pass self.valid_tool_names, so each agent session uses its own tool list instead of the process-global _last_resolved_tool_names (which subagents can overwrite). Also simplifies the redundant ternary in code_execution_tool.py: sandbox_tools is already computed correctly (intersection with session tools, or full SANDBOX_ALLOWED_TOOLS as fallback), so the conditional was dead logic. Inspired by PR #663 (JasonOA888). Closes #662. Tests: 2857 passed.	2026-03-10 06:35:28 -07:00
johnh4098	e9742e202f	fix(security): pipe sudo password via stdin instead of shell cmdline	2026-03-10 06:34:59 -07:00
teknium1	a2ea85924a	Merge PR #687 : fix(file_tools): pass docker_volumes to sandbox container config Authored by manuelschipper. Adds missing docker_volumes key to container_config in file_tools.py, matching terminal_tool.py. Without this, Docker sandbox containers created by file operations lack user volume mounts when file tools run before terminal.	2026-03-10 06:33:30 -07:00
teknium1	8318a519e6	fix: pass enabled_tools through handle_function_call to avoid global race The process-global _last_resolved_tool_names gets overwritten when subagents resolve their own toolsets, causing execute_code in the parent agent to generate imports for the wrong set of tools. Fix: handle_function_call() now accepts an enabled_tools parameter. run_agent.py already passes self.valid_tool_names at both call sites. This change makes model_tools.py actually use it, falling back to the global only when the caller doesn't provide a list (backward compat).	2026-03-10 06:32:08 -07:00
teknium1	8ef3c815e7	Merge PR #680 : feat: add Nous Portal API key provider Authored by Indelwin. Adds 'nous-api' provider for direct API key access to Nous Portal inference, mirroring how OpenRouter and other API-key providers work. Includes PROVIDER_REGISTRY entry, setup wizard option, OPTIONAL_ENV_VARS, provider aliases, and test. Fixes #644.	2026-03-10 06:31:03 -07:00
Indelwin	de07aa7c40	feat: add Nous Portal API key provider (#644 ) Add support for using Nous Portal via a direct API key, mirroring how OpenRouter and other API-key providers work. This gives users a simpler alternative to the OAuth device-code flow when they already have a Nous API key. Changes: - Add 'nous-api' to PROVIDER_REGISTRY as an api_key provider pointing to https://inference-api.nousresearch.com/v1 - Add NOUS_API_KEY and NOUS_BASE_URL to OPTIONAL_ENV_VARS - Add NOUS_API_BASE_URL / NOUS_API_CHAT_URL to hermes_constants - Add 'Nous Portal API key' as first option in setup wizard - Add provider aliases (nous_api, nousapi, nous-portal-api) - Add test for nous-api runtime provider resolution Closes #644	2026-03-10 06:28:00 -07:00
teknium1	928bb16da1	fix: forward thread_id to Telegram adapter + update send_typing signatures Part 2 of thread_id forum topic fix: add metadata param to send_voice, send_image, send_animation, send_typing in Telegram adapter and pass message_thread_id to all Bot API calls. Update send_typing signature in Discord, Slack, WhatsApp, HomeAssistant for compatibility. Based on the fix proposed by @Bitstreamono in PR #656.	2026-03-10 06:26:32 -07:00
teknium1	441f498d6f	Merge PR #679 : fix(code_execution): handle empty enabled_sandbox_tools in schema description Authored by 0xbyt4. Fixes broken 'from hermes_tools import , ...' syntax in schema description when no sandbox tools are enabled. Adds 29 new tests for schema generation, env var filtering, edge cases, and interrupt handling.	2026-03-10 06:22:56 -07:00
teknium1	a630ca15de	fix: forward thread_id metadata for Telegram forum topic routing Replies in Telegram forum topics (supergroups with topics) now land in the correct topic thread instead of 'General'. - base.py: build thread_id metadata from event.source, pass to all send/media calls; add metadata param to send_typing, send_image, send_animation, send_voice, send_video, send_document, send_image_file, _keep_typing - telegram.py: extract thread_id from metadata and pass as message_thread_id to all Bot API calls (send_photo, send_voice, send_audio, send_animation, send_chat_action) - run.py: pass thread_id metadata to progress/streaming send calls - discord/slack/whatsapp/homeassistant: update send_typing signature Based on the fix proposed by @Bitstreamono in PR #656.	2026-03-10 06:21:15 -07:00
0xbyt4	52e3580cd4	refactor: merge new tests into test_code_execution.py Move all new tests (schema, env filtering, edge cases, interrupt) into the existing test_code_execution.py instead of a separate file. Delete the now-redundant test_code_execution_schema.py.	2026-03-10 06:18:27 -07:00
0xbyt4	694a3ebdd5	fix(code_execution): handle empty enabled_sandbox_tools in schema description build_execute_code_schema(set()) produced "from hermes_tools import , ..." in the code property description — invalid Python syntax shown to the model. This triggers when a user enables only the code_execution toolset without any of the sandbox-allowed tools (e.g. `hermes tools code_execution`), because SANDBOX_ALLOWED_TOOLS & {"execute_code"} = empty set. Also adds 29 unit tests covering build_execute_code_schema, environment variable filtering, execute_code edge cases, and interrupt handling.	2026-03-10 06:18:27 -07:00
teknium1	2a062e2f45	Merge PR #840 : background process notification modes + fix spinner line spam - feat(gateway): configurable background_process_notifications (off/result/error/all) - fix(display): rate-limit spinner flushes to prevent line spam under patch_stdout Background notifications inspired by @PeterFile (PR #593).	2026-03-10 06:17:18 -07:00
teknium1	49ec1c9e8f	Merge PR #655 : fix: normalize max turns config path Authored by stablegenius49. Rebased onto current main, resolved 3 conflicts (load_config encoding, save_config commented sections, setup default value), fixed missing MagicMock import, aligned DEFAULT_CONFIG default to 90 (matching cli.py). Migrates legacy root-level max_turns to agent.max_turns across all config loaders (load_config, load_cli_config, save_config, setup). Adds _normalize_max_turns_config() for consistent migration. Fixes #634.	2026-03-10 06:05:20 -07:00
stablegenius49	4bd579f915	fix: normalize max turns config path	2026-03-10 06:05:02 -07:00
teknium1	e4adb67ed8	fix(display): rate-limit spinner flushes to prevent line spam under patch_stdout The KawaiiSpinner animation would occasionally spam dozens of duplicate lines instead of overwriting in-place with \r. This happened because prompt_toolkit's StdoutProxy processes each flush() as a separate run_in_terminal() call — when the write thread is slow (busy event loop during long tool executions), each \r frame gets its own call, and the terminal layout save/restore between calls breaks the \r overwrite semantics. Fix: rate-limit flush() calls to at most every 0.4s. Between flushes, \r-frame writes accumulate in StdoutProxy's buffer. When flushed, they concatenate into one string (e.g. \r frame1 \r frame2 \r frame3) and are written in a single run_in_terminal() call where \r works correctly. The spinner still animates (flush ~2.5x/sec) but each flush batches ~3 frames, guaranteeing the \r collapse always works. Most visible with execute_code and terminal tools (3+ second executions).	2026-03-10 06:02:07 -07:00
teknium1	ff09cad879	Merge PR #621 : fix: limit concurrent Modal sandbox creations to avoid deadlocks Authored by voteblake. - Semaphore limits concurrent Modal sandbox creations to 8 (configurable) to prevent thread pool deadlocks when 86+ tasks fire simultaneously - Modal cleanup guard for failed init (prevents AttributeError) - CWD override to /app for TB2 containers - Add /home/ to host path validation for container backends	2026-03-10 05:57:54 -07:00
teknium1	580e6ba2ff	feat: add proper favicon and logo for landing page and docs site Generated favicon files (ico, 16x16, 32x32, 180x180, 192x192, 512x512) from the Hermes Agent logo. Replaces the inline SVG caduceus emoji with real favicon files so Google's favicon service can pick up the logo. Landing page: updated <link> tags to reference favicon.ico, favicon PNGs, and apple-touch-icon. Docusaurus: updated config to use favicon.ico and logo.png instead of favicon.svg.	2026-03-10 05:51:45 -07:00
teknium1	d6d5a43d3a	Merge PR #627 : fix: continue non-tool replies after output-length truncation Authored by tripledoublev (vincent). Rebased onto current main and conflict-resolved. When finish_reason='length' on a non-tool chat-completions response, instead of rolling back and returning None, the agent now: - Appends the truncated text and a continuation prompt - Retries up to 3 times, accumulating partial chunks - Concatenates all chunks into the final response - Preserves existing rollback behavior for tool-call truncations	2026-03-10 04:33:14 -07:00
teknium1	d723208b1b	Merge PR #617 : Improve skills tool error handling Authored by aydnOktay. Adds logging to skills_tool.py with specific exception handling for file read errors (UnicodeDecodeError, PermissionError) vs unexpected exceptions, replacing bare except-and-continue blocks.	2026-03-10 04:32:26 -07:00
vincent	b0a5fe8974	fix: continue after output-length truncation	2026-03-10 04:30:19 -07:00
teknium1	899dfdcfb9	Merge PR #616 : fix: retry with rebuilt payload after compression Authored by tripledoublev. After context compression on 413/400 errors, the inner retry loop was reusing the stale pre-compression api_messages payload. Fix breaks out of the inner retry loop so the outer loop rebuilds api_messages from the now-compressed messages list. Adds regression test verifying the second request actually contains the compressed payload.	2026-03-10 04:22:42 -07:00
teknium1	8f0b07ed29	Merge PR #611 : fix(session): atomic write for sessions.json to prevent data loss on crash Authored by alireza78a. Replaces open('w') + json.dump with tempfile.mkstemp + os.replace atomic write pattern, matching the existing pattern in cron/jobs.py. Prevents silent session loss if the process crashes or gets OOM-killed mid-write. Resolved conflict: kept encoding='utf-8' from HEAD in the new fdopen call.	2026-03-10 04:18:53 -07:00
teknium1	f16f2912cf	Merge PR #607 : fix: reset all retry counters at start of run_conversation() Authored by 0xbyt4. Adds missing resets for _incomplete_scratchpad_retries and _codex_incomplete_retries to prevent stale counters carrying over between CLI conversations.	2026-03-10 04:17:47 -07:00
teknium1	af748539f8	Merge PR #608 : fix: remove unused imports and unnecessary f-strings Authored by JackTheGit. - Remove unused 'random' import from agent/display.py - Remove unused 'Optional' import from agent/redact.py - Remove unnecessary f-string prefixes in batch_runner.py	2026-03-10 04:16:23 -07:00
teknium1	695c017411	Merge PR #603 : fix: return deny on approval callback timeout instead of None Authored by 0xbyt4. _approval_callback() had no return statement after the timeout break, causing it to return None instead of 'deny'. Callers in approval.py expect one of 'once', 'session', 'always', or 'deny'. This matches the existing timeout behavior in approval.py:209.	2026-03-10 04:15:31 -07:00
teknium1	5e6c7bc205	Merge PR #602 : fix: prevent data loss in clipboard PNG conversion when ImageMagick fails Authored by 0xbyt4. Only deletes temp .bmp after confirmed successful conversion, restores original on failure. Adds 3 tests.	2026-03-10 04:15:05 -07:00
teknium1	e8cec55fad	feat(gateway): configurable background process watcher notifications Add display.background_process_notifications config option to control how chatty the gateway process watcher is when using terminal(background=true, check_interval=...) from messaging platforms. Modes: - all: running-output updates + final message (default, current behavior) - result: only the final completion message - error: only the final message when exit code != 0 - off: no watcher messages at all Also supports HERMES_BACKGROUND_NOTIFICATIONS env var override. Includes 12 tests (5 config loading + 7 watcher behavior). Inspired by @PeterFile's PR #593. Closes #592.	2026-03-10 04:12:39 -07:00
teknium1	67fc6bc4e9	Merge PR #600 : fix(security): use in-memory set for permanent allowlist save Authored by alireza78a. Uses _permanent_approved directly instead of re-reading from disk, preventing potential data loss if a previous save failed.	2026-03-10 04:12:11 -07:00

1 2 3 4 5 ...

1228 Commits