Timmy-time-dashboard

Archived

forked from Rockachopa/Timmy-time-dashboard

Author	SHA1	Message	Date
hermes	48c8efb2fb	[loop-cycle-40] fix: use get_system_prompt() in cloud backends (#135 ) (#138 ) ## What Cloud backends (Grok, Claude, AirLLM) were importing SYSTEM_PROMPT directly, which is always SYSTEM_PROMPT_LITE and contains unformatted {model_name} and {session_id} placeholders. ## Changes - backends.py: Replace `from timmy.prompts import SYSTEM_PROMPT` with `from timmy.prompts import get_system_prompt` - AirLLM: uses `get_system_prompt(tools_enabled=False, session_id="airllm")` (LITE tier, correct) - Grok: uses `get_system_prompt(tools_enabled=True, session_id="grok")` (FULL tier) - Claude: uses `get_system_prompt(tools_enabled=True, session_id="claude")` (FULL tier) - 9 new tests verify formatted model names, correct tier selection, and session_id formatting ## Tests 1508 passed, 0 failed (41 new tests this cycle) Fixes #135 Co-authored-by: Kimi Agent <kimi@timmy.local> Reviewed-on: http://localhost:3000/rockachopa/Timmy-time-dashboard/pulls/138 Reviewed-by: rockachopa <alexpaynex@gmail.com> Co-authored-by: hermes <hermes@timmy.local> Co-committed-by: hermes <hermes@timmy.local>	2026-03-15 09:44:43 -04:00
hermes	d48d56ecc0	[loop-cycle-38] fix: add soul identity to system prompts (#127 ) (#134 ) Co-authored-by: hermes <hermes@timmy.local> Co-committed-by: hermes <hermes@timmy.local>	2026-03-15 09:42:57 -04:00
hermes	76df262563	[loop-cycle-38] fix: add retry logic for Ollama 500 errors (#131 ) (#133 ) Co-authored-by: hermes <hermes@timmy.local> Co-committed-by: hermes <hermes@timmy.local>	2026-03-15 09:38:21 -04:00
hermes	92e123c9e5	[loop-cycle-36] fix: create soul.md and wire into system context (#125 ) (#130 )	2026-03-15 08:37:24 -04:00
hermes	cf48b7d904	[loop-cycle-1] fix: lint errors — ambiguous vars + unused import (#123 ) (#124 )	2026-03-15 08:07:19 -04:00
hermes	9220732581	Merge pull request '[loop-cycle-31] feat: workspace heartbeat monitoring (#28 )' (#120 ) from feat/workspace-heartbeat into main	2026-03-14 21:52:24 -04:00
Kimi Agent	66544d52ed	feat: workspace heartbeat monitoring for thinking engine (#28 ) - Add src/timmy/workspace.py: WorkspaceMonitor tracks correspondence.md line count and inbox file list via data/workspace_state.json - Wire workspace checks into _gather_system_snapshot() so Timmy sees new workspace activity in his thinking context - Add 'workspace' seed type for workspace-triggered reflections - Add _check_workspace() post-hook to mark items as seen after processing - 16 tests covering detection, mark_seen, persistence, edge cases	2026-03-14 21:51:36 -04:00
Kimi Agent	a277d40e32	feat: Timmy authenticates to Gitea as himself - .timmy_gitea_token checked before legacy ~/.config/gitea/token - Token created for Timmy user (id=2) with write collaborator perms - .timmy_gitea_token added to .gitignore	2026-03-14 21:45:54 -04:00
Kimi Agent	a57fd7ea09	[loop-cycle-30] fix: gitea-mcp binary name + test stabilization 1. gitea-mcp → gitea-mcp-server (brew binary name). Fixes Timmy's Gitea triage — MCP server can now be found on PATH. 2. Mark test_returns_dict_with_expected_keys as @pytest.mark.slow — it runs pytest recursively and always exceeds the 30s timeout. 3. Fix ruff F841 lint in test_cli.py (unused result= variable).	2026-03-14 21:32:39 -04:00
Kimi Agent	b9b78adaa2	perf: eliminate redundant LLM calls in agentic loop (#24 ) Three optimizations to the agentic loop: 1. Cache loop agent as singleton (avoid repeated warmups) 2. Sliding window for step context (last 2 results, not all) 3. Replace summary LLM call with deterministic summary Saves 1 full LLM inference call per agentic loop invocation (30-60s on local models) and reduces context window pressure. Also fixes pre-existing test_cli.py repl test bugs (missing result= assignment).	2026-03-14 20:55:52 -04:00
Kimi Agent	65e5e7786f	feat: REPL mode, stdin support, multi-word fix for CLI (#26 )	2026-03-14 20:45:25 -04:00
Kimi Agent	547b502718	fix: smart_read_file accepts path= kwarg from LLMs (#113 ) LLMs naturally call read_file(path=...) but the wrapper only accepted file_name=. Pydantic strict validation rejected the mismatch. Now accepts both file_name and path kwargs, with clear error on missing both. Added 6 tests covering: positional args, path kwarg, no-args error, directory listing, empty dir, hidden file filtering.	2026-03-14 20:40:19 -04:00
hermes	3e7a35b3df	Merge pull request '[loop-cycle-12] feat: Kimi delegation tool for coding tasks (#67 )' (#112 ) from fix/kimi-delegation-67 into main	2026-03-14 20:31:08 -04:00
Kimi Agent	453c9a0694	feat: add delegate_to_kimi() tool for coding delegation (#67 ) Timmy can now delegate coding tasks to Kimi CLI (262K context). Includes timeout handling, workdir validation, output truncation. Sovereign division of labor — Timmy plans, Kimi codes.	2026-03-14 20:29:03 -04:00
Kimi Agent	2fb104528f	feat: add run_self_tests() tool for self-verification (#65 ) Timmy can now run his own test suite via the run_self_tests() tool. Supports 'fast' (unit only), 'full', or specific path scopes. Returns structured results with pass/fail counts. Sovereign self-verification — a fundamental capability.	2026-03-14 20:28:24 -04:00
Kimi Agent	ddb872d3b0	fix: enrich self-knowledge with architecture map and self-modification pathway - Replace flat file list with layered architecture map (config→agent→prompt→tool→memory→interface) - Add SELF-MODIFICATION section: Timmy knows he can edit his own config and code - Remove false limitation 'cannot modify own source code' - Update tests to match new section headers, add self-modification tests Closes #81 (reasoning depth) Closes #86 (self-modification awareness) [loop-cycle-11]	2026-03-14 20:15:30 -04:00
Kimi Agent	b12e29b92e	fix: dedup memory consolidation with existing memory search (#105 ) _maybe_consolidate() now checks get_memories(subject=agent_id) before storing. Skips if a memory of the same type (pattern/anomaly) was created within the last hour. Prevents duplicate consolidation entries on repeated task completion/failure events. Also restructured branching: neutral success rates (0.3-0.8) now return early instead of falling through. 9 new tests. 1465 total passing.	2026-03-14 20:04:18 -04:00
Kimi Agent	ffae5aa7c6	feat: add codebase self-knowledge to system prompts (#78 , #80 ) Adds SELF-KNOWLEDGE section to both SYSTEM_PROMPT_LITE and SYSTEM_PROMPT_FULL with: - Codebase map (all src/timmy/ modules with descriptions) - Current capabilities list (grounded, not generic) - Known limitations (real gaps, not LLM platitudes) Lite prompt gets condensed version; full prompt gets detailed. Timmy can now answer 'what does tool_safety.py do?' and give grounded answers about his actual limitations. 10 new tests. 1456 total passing.	2026-03-14 19:58:10 -04:00
hermes	0204ecc520	Merge pull request '[loop-cycle-9] fix: CLI multi-word messages (#26 )' (#107 ) from fix/cli-multiword-messages into main	2026-03-14 19:48:28 -04:00
Kimi Agent	9171d93ef9	fix: CLI chat accepts multi-word messages without quotes Changed message param from str to list[str] in chat() and route() commands. Words are joined with spaces, so 'timmy chat hello how are you' works without quoting. Single-word messages still work as before. - chat(): message: list[str], joined to full_message - route(): message: list[str], joined to full_message - 7 new tests in test_cli_multiword.py Closes #26	2026-03-14 19:43:52 -04:00
Kimi Agent	f8f3b9b81f	feat: inject session_id into system prompt for session identity awareness Timmy can now introspect which session he's running in (cli, dashboard, loop). - Add {session_id} placeholder to both lite and full system prompts - get_system_prompt() accepts session_id param (default: 'unknown') - create_timmy() accepts session_id param, forwards to prompt - CLI chat/think/status pass their session_id to create_timmy() - session.py passes _DEFAULT_SESSION_ID to create_timmy() - 7 new tests in test_session_identity.py - Updated 2 existing CLI test mocks Closes #64	2026-03-14 19:43:11 -04:00
hermes	4b553fa0ed	Merge pull request 'fix: word-boundary routing + debug route command (#31 )' (#102 ) from fix/routing-patterns into main	2026-03-14 19:24:16 -04:00
hermes	342b9a9d84	Merge pull request 'feat: JSON status endpoints for briefing, memory, swarm (#49 , #50 )' (#101 ) from fix/api-consistency into main	2026-03-14 19:24:15 -04:00
Kimi Agent	b3809f5246	feat: add JSON status endpoints for briefing, memory, swarm (#49 , #50 )	2026-03-14 19:23:32 -04:00
Kimi Agent	67497133fd	fix: word-boundary routing + debug route command (#31 ) - Replace substring matching with word-boundary regex in route_request() - "fix the bug" now correctly routes to coder - Multi-word patterns match if all words appear (any order) - Add "timmy route" CLI command for debugging routing - Add route_request_with_match() for pattern visibility - Expand routing keywords in agents.yaml - 22 new routing tests, all passing	2026-03-14 19:21:30 -04:00
hermes	c1ec43c59f	Merge pull request '[loop-cycle-8] fix: replace 59 bare except clauses with proper logging (#25 )' (#99 ) from fix/bare-except-clauses into main	2026-03-14 19:08:40 -04:00
Kimi Agent	fdc5b861ca	fix: replace 59 bare except clauses with proper logging (#25 ) All `except Exception:` now catch as `except Exception as exc:` with appropriate logging (warning for critical paths, debug for graceful degradation). Added logger setup to 4 files that lacked it: - src/timmy/memory/vector_store.py - src/dashboard/middleware/csrf.py - src/dashboard/middleware/security_headers.py - src/spark/memory.py 31 files changed across timmy core, dashboard, infrastructure, integrations. Zero bare excepts remain. 1340 tests passing.	2026-03-14 19:07:14 -04:00
Kimi Agent	9c59b386d8	feat: add OLLAMA_NUM_CTX config to cap context window (#83 ) - Add ollama_num_ctx setting (default 4096) to config.py - Pass num_ctx option to Ollama in agent.py and agents/base.py - Add OLLAMA_NUM_CTX to .env.example with usage docs - Add context_window note in providers.yaml - Fix mock_settings in test_agent.py for new attribute - qwen3:30b with 4096 ctx uses ~19GB vs 45GB default	2026-03-14 18:54:43 -04:00
Kimi Agent	bce6e7d030	fix: log Ollama disconnections with specific error handling (#92 ) - BaseAgent.run(): catch httpx.ConnectError/ReadError/ConnectionError, log 'Ollama disconnected: <error>' at ERROR level, then re-raise - session.py: distinguish Ollama disconnects from other errors in chat(), chat_with_tools(), continue_chat() — return specific message 'Ollama appears to be disconnected' instead of generic error - 11 new tests covering all disconnect paths	2026-03-14 18:40:15 -04:00
hermes	8a14bbb3e0	Merge pull request '[loop-cycle-5] fix: warmup model on cold load (#82 )' (#95 ) from fix/warmup-cold-model into main	2026-03-14 18:26:48 -04:00
Kimi Agent	86956bd057	fix: warmup model on cold load to prevent first-request disconnect Add _warmup_model() that sends a minimal generation request (1 token) before returning the Agent. 60s timeout handles cold VRAM loads. Warns but does not abort if warmup fails. Closes #82	2026-03-14 18:24:00 -04:00
Kimi Agent	b3a1e0ce36	fix: prune dead web_search tool — ddgs never installed (#87 ) Remove DuckDuckGoTools import, all web_search registrations across 4 toolkit factories, catalog entry, safety classification, prompt references, and session regex. Total: -41 lines of dead code. consult_grok is functional (grok_enabled=True, API key set) and opt-in, so it stays — but Timmy never calls it autonomously, which is correct sovereign behavior (no cloud calls unless user permits). Closes #87	2026-03-14 18:13:51 -04:00
Kimi Agent	7132b42ff3	fix: model introspection uses exact match, queries /api/ps first _get_ollama_model() used prefix match (startswith) on /api/tags, causing qwen3:30b to match qwen3.5:latest. Now: 1. Queries /api/ps (loaded models) first — most accurate 2. Falls back to /api/tags with exact name match 3. Reports actual running model, not just configured one Updated test_get_system_info_contains_model to not assume model==config. Fixes #77. 5 regression tests added.	2026-03-14 18:03:59 -04:00
hermes	74e426c63b	[loop-cycle-2] fix: suppress confirmation tool WARNING spam (#79 ) (#89 )	2026-03-14 17:54:58 -04:00
hermes	09fcf956ec	Merge pull request '[loop-cycle-1] feat: tool allowlist for autonomous operation (#69 )' (#88 ) from fix/tool-allowlist-autonomous into main	2026-03-14 17:41:56 -04:00
Kimi Agent	d28e2f4a7e	[loop-cycle-1] feat: tool allowlist for autonomous operation (#69 ) Add config/allowlist.yaml — YAML-driven gate that auto-approves bounded tool calls when no human is present. When Timmy runs with --autonomous or stdin is not a terminal, tool calls are checked against allowlist: matched → auto-approved, else → rejected. Changes: - config/allowlist.yaml: shell prefixes, deny patterns, path rules - tool_safety.py: is_allowlisted() checks tools against YAML rules - cli.py: --autonomous flag, _is_interactive() detection - 44 new allowlist tests, 8 updated CLI tests Closes #69	2026-03-14 17:39:48 -04:00
Kimi Agent	94cd1a9840	fix: make model fallback chains configurable (#53 ) Move hardcoded model fallback lists from module-level constants into settings.fallback_models and settings.vision_fallback_models (pydantic Settings fields). Can now be overridden via env vars FALLBACK_MODELS / VISION_FALLBACK_MODELS or config/providers.yaml. Removed: - OLLAMA_MODEL_PRIMARY / OLLAMA_MODEL_FALLBACK from config.py - DEFAULT_MODEL_FALLBACKS / VISION_MODEL_FALLBACKS from agent.py get_effective_ollama_model() and _resolve_model_with_fallback() now walk the configurable chains instead of hardcoded constants. 5 new tests guard the configurable behavior and prevent regression to hardcoded constants.	2026-03-14 17:26:47 -04:00
Kimi Agent	061c8f6628	fix: brevity tuning — plain text prompts, markdown=False, front-loaded brevity Closes #71: Timmy was responding with elaborate markdown formatting (tables, headers, emoji, bullet lists) for simple questions. Root causes fixed: 1. Agno Agent markdown=True flag explicitly told the model to format responses as markdown. Set to False in both agent.py and agents/base.py. 2. SYSTEM_PROMPT_FULL used ## and ### markdown headers, bold (**), and numbered lists — teaching by example that markdown is expected. Rewritten to plain text with labeled sections. 3. Brevity instructions were buried at the bottom of the full prompt. Moved to immediately after the opening line as 'VOICE AND BREVITY' with explicit override priority. 4. Orchestrator prompt in agents.yaml was silent on response style. Added 'Voice: brief, plain, direct' with concrete examples. The full prompt is now 41 lines shorter (124 → 83). The prompt itself practices the brevity it preaches. SOUL.md alignment: - 'Brevity is a kindness' — now front-loaded in both base and agent prompt - 'I do not fill silence with noise' — explicit in both tiers - 'I speak plainly. I prefer short sentences.' — structural enforcement 4 new tests guard against regression: - test_full_prompt_brevity_first: brevity section before tools/memory - test_full_prompt_no_markdown_headers: no ## or ### in prompt text - test_full_prompt_plain_text_brevity: 'plain text' instruction present - test_lite_prompt_brevity: lite tier also instructs brevity	2026-03-14 17:15:56 -04:00
rockachopa	927e25cc40	Merge pull request 'fix: replace print() with proper logging (#29 , #51 )' (#59 ) from fix/print-to-logging into main	2026-03-14 16:50:04 -04:00
Kimi Agent	64fd1d9829	voice: reinforce brevity at top of system prompt	2026-03-14 16:32:47 -04:00
Kimi Agent	f0b0e2f202	fix: WebSocket 403 spam and missing /swarm endpoints - CSRF middleware now skips WebSocket upgrade requests (they don't carry tokens) - Added /swarm/live WebSocket endpoint wired to ws_manager singleton - Added /swarm/agents/sidebar HTMX partial (was 404 on every dashboard poll) Stops hundreds of 403 Forbidden + 404 log lines per minute.	2026-03-14 16:29:59 -04:00
Kimi Agent	b30b5c6b57	[loop-cycle-6] Break thinking rumination loop — semantic dedup (#38 ) Add post-generation similarity check to ThinkingEngine.think_once(). Problem: Timmy's thinking engine generates repetitive thoughts because small local models ignore 'don't repeat' instructions in the prompt. The same observation ('still no chat messages', 'Alexander's name is in profile') would appear 14+ times in a single day's journal. Fix: After generating a thought, compare it against the last 5 thoughts using SequenceMatcher. If similarity >= 0.6, retry with a new seed up to 2 times. If all retries produce repetitive content, discard rather than store. Uses stdlib difflib — no new dependencies. Changes: - thinking.py: Add _is_too_similar() method with SequenceMatcher - thinking.py: Wrap generation in retry loop with dedup check - test_thinking.py: 7 new tests covering exact match, near match, different thoughts, retry behavior, and max-retry discard +96/-20 lines in thinking.py, +87 lines in tests.	2026-03-14 16:21:16 -04:00
Kimi Agent	79edfd1106	feat: persist chat history in SQLite — survives server restarts Replace in-memory MessageLog with SQLite-backed implementation. Same API surface (append/all/clear/len) so zero caller changes needed. - data/chat.db stores messages with role, content, timestamp, source - Lazy DB connection (opened on first use, not at import time) - Retention policy: oldest messages pruned when count > 500 - New .recent(limit) method for efficient last-N queries - Thread-safe with explicit locking - WAL mode for concurrent read performance - Test isolation: conftest redirects DB to tmp_path per test - 8 new tests: persistence, retention, concurrency, source field Closes #46	2026-03-14 16:09:26 -04:00
Kimi Agent	f426df5b42	feat: add --session-id option to timmy chat CLI Allows specifying a named session for conversation persistence. Use cases: - Autonomous loops can have their own session (e.g. --session-id loop) - Multiple users/agents can maintain separate conversations - Testing different conversation threads without polluting the default Precedence: --session-id > --new > default 'cli' session	2026-03-14 16:05:00 -04:00
Kimi Agent	70d5dc5ce1	fix: replace eval() with AST-walking safe evaluator in calculator Fixes #52 - Replace eval() in calculator() with _safe_eval() that walks the AST and only permits: numeric constants, arithmetic ops (+,-,,/,//,%,*), unary +/-, math module access, and whitelisted builtins (abs, round, min, max) - Reject all other syntax: imports, attribute access on non-math objects, lambdas, comprehensions, string literals, etc. - Add 39 tests covering arithmetic, precedence, math functions, allowed builtins, error handling, and 14 injection prevention cases	2026-03-14 15:51:35 -04:00
Kimi Agent	db129bbe16	fix: replace print() with proper logging (#29 , #51 )	2026-03-14 15:07:07 -04:00
Kimi Agent	591954891a	fix: sanitize dynamic innerHTML in templates (#47 )	2026-03-14 15:07:00 -04:00
Kimi Agent	bb287b2c73	fix: sanitize WebSocket data in HTML templates (XSS #47 )	2026-03-14 15:01:48 -04:00
Kimi Agent	efb1feafc9	fix: replace print() with proper logging (#29 , #51 )	2026-03-14 15:01:34 -04:00
Hermes Agent	fa838b0063	fix: clean shutdown — silence MCP async-generator teardown noise Swallow anyio cancel-scope RuntimeError and BaseExceptionGroup from MCP stdio_client generators during GC on voice loop exit. Custom unraisablehook + loop exception handler + warnings filter.	2026-03-14 14:12:05 -04:00

1 2 3 4 5

229 Commits