Thor Henning Hetland’s Post

Every agent monitor I've seen tries to capture what the model was thinking.  That's the wrong thing to capture.  A model's chain-of-thought is an unreliable narrator. It will tell you it reasoned carefully about the evidence. It may have. But emitted reasoning is a plausible reconstruction, not a faithful record. You can't audit a narrator.  We built something different.  kcp-dashboard doesn't scrape thoughts. It reconstructs the decision graph from the governance layer — not what the model said it considered, but which documents it was actually handed, and which it wasn't, and where each candidate failed.  pci-scope — skipped at temporal. Out of date.  vendor-intel — skipped at payment. Not cleared.  prod-secrets — skipped at access. Restricted.  Same inputs, same cascade, same verdict — every time. That's not a guess about reasoning. That's a record of the information environment the agent was operating in.  The thought graph: not what the model was thinking. What it was given to think with.  ---  Honest limit: this works because kcp-harness runs a fixed, deterministic governance cascade and emits a content-free trace of every verdict. If your agent doesn't have a deterministic governance layer, there's nothing faithful to reconstruct. You're back to the unreliable narrator.  That's not a caveat. It's the actual point.  Five scenarios. Three complexity levels. Runs in your browser. https://lnkd.in/eJwK-XeJ

  • No alternative text description for this image

To view or add a comment, sign in

Explore content categories