Timmy_Foundation/compounding-intelligence

Go to file

Alexander Payne ceb7e0bd0c

Test / pytest (pull_request) Failing after 30s

Details

feat: add doc link validator script (closes #103 )

Add scripts/validate_doc_links.py — scans all markdown files in the
repository, extracts inline and autolinks, and verifies each URL via
HTTP HEAD request (with GET fallback for servers that reject HEAD).

Features:
  --root           : repository root to scan (default: repo root)
  --fail-on-broken : exit 1 if any broken links found
  --json           : emit JSON report for CI consumption
  --ignore         : comma-separated URL prefixes to skip

Ignores non-HTTP URLs, localhost/127.0.0.1, and private IP ranges.
Requires only Python stdlib — no external dependencies.

Smoke-tested against this repo: 2 unique URLs checked, 0 broken.
Addresses 4.8: Doc Link Validator acceptance criteria.

Closes #103

2026-04-25 20:55:19 -04:00

.gitea/workflows

ci: add pytest workflow for #190

2026-04-15 11:29:23 -04:00

knowledge

feat(#10 ): knowledge file format schema + example knowledge files

2026-04-14 14:21:21 -04:00

metrics

Initial structure: knowledge store, scripts, metrics, templates

2026-04-14 11:17:01 -04:00

scripts

feat: add doc link validator script (closes #103 )

2026-04-25 20:55:19 -04:00

templates

feat: build bootstrapper.py - pre-session context assembler

2026-04-14 14:05:30 -04:00

test_sessions

Add test session 5: Session with questions

2026-04-14 19:01:03 +00:00

tests

Merge pull request 'feat: knowledge deduplication — content hash + token similarity (#196 )' (#228 ) from burn/196-1776306000 into main

2026-04-21 15:28:50 +00:00

.gitignore

fix: implement refactoring_opportunity_finder API (#210 )

2026-04-21 07:29:44 -04:00

GENOME.md

fix(#676 ): update GENOME.md for compounding-intelligence

2026-04-21 04:43:54 +00:00

Makefile

ci: add pytest workflow for #190

2026-04-15 11:29:23 -04:00

quality_gate.py

feat: quality gate — score and filter knowledge entries (#198 )

2026-04-20 20:31:04 -04:00

README.md

Initial structure: knowledge store, scripts, metrics, templates

2026-04-14 11:17:01 -04:00

requirements.txt

ci: add pytest workflow for #190

2026-04-15 11:29:23 -04:00

README.md

Compounding Intelligence

Turn 1B+ daily tokens into durable, compounding fleet intelligence.

The Problem

20,991 sessions on disk. Each one starts at zero. Every agent rediscover the same HTTP 405 is a branch protection issue. The intelligence from a million tokens of work evaporates when the session ends.

The Solution

Three pipelines that form a compounding loop:

SESSION ENDS → HARVESTER → KNOWLEDGE STORE → BOOTSTRAPPER → NEW SESSION STARTS SMARTER
                              ↓
                         MEASURER → Prove it's working

Architecture

Pipeline 1: Harvester

Reads finished session transcripts. Extracts durable knowledge: facts, pitfalls, patterns, tool quirks. Stores in knowledge/.

Pipeline 2: Bootstrap

Before a session starts, queries knowledge store for relevant facts. Assembles compact 2k-token context. Injects into session so it starts with full situational awareness.

Pipeline 3: Measure

Tracks whether compounding is happening. Knowledge velocity, error reduction, hit rate, task completion. Daily report proves the loop works.

Directory Structure

├── knowledge/
│   ├── index.json          # Machine-readable fact index
│   ├── global/             # Cross-repo knowledge
│   ├── repos/{repo}.md     # Per-repo knowledge
│   └── agents/{agent}.md   # Agent-type notes
├── scripts/
│   ├── harvester.py        # Post-session knowledge extractor
│   ├── bootstrapper.py     # Pre-session context loader
│   ├── measurer.py         # Compounding metrics
│   └── session_reader.py   # JSONL parser
├── metrics/
│   └── dashboard.md        # Human-readable status
└── templates/
    ├── bootstrap-context.md
    └── harvest-prompt.md

The 100x Path

Month 1:  15,000 facts, sessions 20% faster
Month 2:  45,000 facts, sessions 40% faster, first-try success up 30%
Month 3:  90,000 facts, fleet measurably smarter per token

Each new session is better than the last. The intelligence compounds.

Issues

See all issues for the full roadmap.

Epics:

EPIC 1: Session Harvester (#2)
EPIC 2: Knowledge Store & Bootstrap (#3)
EPIC 3: Compounding Measurement (#4)
EPIC 4: Retroactive Harvest (#5)