How to Add an Evidence-Based Confidence Score to AI Answers
Procedure for adding an evidence-based confidence score to non-sentinel answers while preserving fail-closed evidence behavior.
Purpose
This guide explains how to add an evidence-based confidence score to responses. The score is a reporting and verification layer: it communicates evidential support after the answer is checked. It is not a calibrated probability, not a model log-probability, and not a substitute for evidence. The 0–100 score is a site-level reporting convention. If an active evidence policy requires a sentinel-only fail-closed answer, the sentinel remains the entire answer and no confidence score is appended.
When to use this
Use this section to decide whether this workflow is the right fit before you configure prompts, policies, or reference material.
-
Use caseThe final answer needs explicit confidence reportingUse this when every non-sentinel answer must end with a standardized confidence line tied to evidence quality and uncertainty.
-
Use caseThe workflow already uses evidence boundariesUse this with files-only, authoritative-source, academic, or verification workflows when confidence must reflect source support rather than tone or persuasion.
-
Use caseThe output will be reviewed, published, or used downstreamUse this when users need to understand whether the answer is strongly supported, partially supported, or limited by missing evidence.
Choose how to enforce confidence reporting
Use the prompt page for a reusable instruction-layer rule, the policy when adding the scoring contract to another workflow, or the Fact-Checking Kit when confidence belongs inside a broader verification workflow.
Use the Confidence Score policy
Use the Fact-Checking Kit
Step-by-step implementation procedure
Follow the workflow in order. Each step gives one action and one verification check before continuing.
-
Step 1 · Configure the confidence rule
Add the Confidence Score prompt when confidence reporting should be reused across tasks.Place the stable rule in the instruction layer of the tool you use, or add the policy contract to an existing workflow.Check: The rule requires a confidence line only for non-sentinel answers. -
Step 2 · Keep the evidence basis explicit
Identify the sources, artifacts, citations, policies, or verification results that support the answer.Tie the score to evidence quality, source coverage, conflicts, and missing information.Check: The confidence score can be explained from the evidence basis. -
Step 3 · Answer the current task under the active evidence boundary
Use the current task, source boundary, and output contract before assigning a confidence score.Do not let the confidence rule expand the allowed sources or override fail-closed policies.Check: The answer follows the active source boundary. -
Step 4 · Check fail-closed compatibility
If an active policy requires a sentinel-only answer, output the sentinel and stop.Do not append a confidence line after a sentinel-only fail-closed answer.Check: Sentinel-only answers remain exact and unchanged. -
Step 5 · Add the confidence line
For non-sentinel answers, end with the standardized confidence line.Use the required format and lower the score when evidence is incomplete, conflicting, outdated, or partially inspected.Check: The final line uses the configured confidence format and reflects evidence support.
Verification checklist
Use this checklist before accepting the output, publishing it, or using it as evidence for a downstream workflow.
-
FormatThe final answer has the required confidence lineEvery non-sentinel answer ends with the configured numeric confidence format.
-
MeaningThe score reflects evidence supportThe score is tied to correctness, evidence quality, coverage, conflicts, uncertainty, and missing information.
-
BoundaryThe score did not expand the evidence boundaryConfidence reporting does not allow unsupported claims, unstated sources, or fabricated evidence.
-
Fail-closed behaviorSentinel-only outputs stay sentinel-onlyWhen a policy requires an exact sentinel response, no confidence line is appended.
-
UncertaintyLower confidence is explained by evidence limitsWhen confidence is limited, the answer identifies the missing or weak evidence that prevents a higher score.