All checks were successful
Smoke Test / smoke (pull_request) Successful in 15s
tests/tool_call_regression.py: - 10 test cases covering 5 hermes tools: read_file, web_search, terminal, execute_code, delegate_task - Schema validation (OpenAI-compatible tool call format) - Argument validation (correct tool + expected args) - Parallel tool calling test (multiple tools in one response) - Dry-run mode for CI (schema validation without server) - Full server mode with latency tracking - Markdown report generation with results matrix - JSON results output for programmatic consumption - 95% accuracy threshold gate (exit code 1 on failure) benchmarks/tool-call-regression.md: - Results template with model/preset matrix - Tool coverage tracking table .gitea/workflows/smoke.yml: - Added dry-run tool call schema validation step