Grounded Confidence

A method for making AI confidence depend on where the system has historically performed well, measured against expert-labeled evals.

Runtime confidence component using grounded confidence, case-local confidence, and historical evals to publish or abstain.
Grounded Confidence uses historical evals and previous agent runs to decide when confidence is earned.
- article coming on July 13 -