The scenario the project exists for: capture and seal a run, archive it, then six months later verify it is unmodified, replay it offline, and hand over an Evidence Bundle — with an explicit section on what it does not prove
Hands-on tour of every capability — proxies, all providers, LangChain, evidence bundles, knowledge graph, capture-level policy, GDPR erasure, dashboard reports, topology, compliance exports, the zero-token eval loop, intervention replay, the Accountability Spine (energy / ledger / safety case), incident tracking, plus the Langfuse-parity cohort: prompt lifecycle and labels, offline analytics (query / views / trends / pricing), sessions and the execution graph, and the team evaluation workflow
Where capsules live (always filesystem), why the current single-cluster design doesn't reach 1M agents, the self-contained rule, the four-layer architecture, the three hard problems (storage, lineage, identity), and the phased build plan