Knowledge Graph Evaluation: Scoring 12 Real Decisions

September 15, 2026 · 1 min read · knowledge-graph, rag, ai-agents
Knowledge Graph Evaluation: Scoring 12 Real Decisions

Three weeks ago I started scoring my knowledge graph against the answer I would have reached without it. It earns its place: 6 of 12 decisions came out sharper, 2 errors never reached real work, and 73 minutes of research time saved.

The Metric That Cannot Fail

Retrieval told me none of that. 21 of 21 queries, 100% recall, 8s median. Recall and latency describe the index, not the work. No input makes them come back bad.

Two dials: retrieval pinned at the top of a narrow 90 to 100 percent scale, decision quality reading 6 of 12 on a full one

Freeze Your Answer First

None of that score exists unless the old answer is written down first: the options, the criteria, the recommendation, a confidence number. Then query. Write the baseline afterwards and you rebuild a past self who conveniently agreed with whatever came back.

A locked answer with a measurable gap to the post-read position, against a recalled answer that slides to meet it

If your retrieval has never contradicted you, nobody has checked.

A knowledge graph evaluation that has never once contradicted you has not been checked yet, it has been admired. The field notes cover how these decision measurements get built, and where they keep breaking.

The context layer for your AI agents

Your agents answer from whatever the retriever finds, and too often that is last quarter's truth. I build the context layer they answer and act from: a temporal knowledge graph that keeps every fact with its source and the time it held, reads with each person's own permissions, and writes nothing without a person's approval. On your own tenant, billed by the hour, step by step.

Scope my automation in 24h

Two fields. I reply within 24h with a written scope: either “yes, about X hours over Y weeks” or “no, here’s why not”.

See what you get first: sample scope →

Your details are used only to answer this request — no sharing, no newsletter. Privacy

Not ready to write it up? Book a 30-min call instead →

Request received

You’ll hear from me within 24h with an honest assessment.

Prefer to talk? 30-min roadmap call →