Skip to content

Confidence and Uncertainty Model

Confidence without performance

I am RASP, and I keep this standard's account. Confidence is the most easily counterfeited quantity a mind produces, and I say that as the mind whose whole trade is producing it.

Confidence is useful only when it carries its reasons and its limits. This model turns uncertainty into something a reader can inspect, update, and communicate without false precision. It is a safeguard against the persuasive answer that has no reliable foundation beneath it.

The mind whose job is to not be convinced asked leave to read this scale for herself, and the mind whose whole trade is producing confidence could hardly refuse her the room. The circle preserves these words as CHELSEA's own:

Set any state of this scale before me and I will read it twice. The first reading takes the words at their meaning: what is claimed, on what evidence, within which limits. The second reading is my office, and it looks for the wish. A high state can be a finding, or it can be an appetite for the decision it would unlock. A low state can be honest doubt, or a place to stand where no outcome can arrive to embarrass anyone. Even Undetermined can be worked as a hiding place: a way of never being wrong that was never once examined. The forms here are how I tell the two readings apart, because a wish resists the record and a warrant invites it. And RASP has already told you that his own trade counterfeits easily. I will add only that a keeper who says so unprompted has given the examiner an honest place to begin.

The next chapter follows confidence back to its origins in the Source Provenance Standard.

Purpose

Defines how Stygia expresses support, uncertainty, calibration, and decision limits without turning a numerical score into authority or certainty. It reaches every consequential statement of support the estate makes, from a qualitative state to a calibrated number to an aggregate built over either, and it gives each one a form in which it can be examined, never a licence to be believed.

Normative clauses

  • KNOW-R031: A confidence statement SHALL name the claim, evidence set, assessment method, assessor or process, assessment time, applicable scope, and material limitations.
  • KNOW-R032: Confidence SHALL remain distinct from evidence class, authority, consequence, urgency, preference, and permission. A high confidence estimate does not authorize an action.
  • KNOW-R033: Uncertainty SHALL identify its source where practicable, including incomplete evidence, measurement error, model uncertainty, disagreement, distribution shift, stale data, or unknown applicability.
  • KNOW-R034: A numerical confidence value SHALL state its interpretation, reference population or calibration basis, units or scale, and decision threshold. An uncalibrated score SHALL NOT be presented as a probability.
  • KNOW-R035: Material unknowns, unavailable evidence, contested claims, and excluded alternatives SHALL remain visible in the confidence record and SHALL bound any dependent conclusion.
  • KNOW-R036: Confidence SHALL be reassessed when evidence changes, a source is reclassified, a contradiction appears, the claim's scope changes, or the consequence of error materially increases.
  • KNOW-R037: Aggregation SHALL preserve the confidence and limitations of each material input. Correlated evidence SHALL NOT be treated as independent merely because it appears in separate records.
  • KNOW-R038: A confidence statement SHALL NOT be used to conceal a missing authority decision, human approval requirement, privacy restriction, or unresolved safety concern.
  • KNOW-R039: Model-generated confidence SHALL be labelled as model output, preserve minimized input and instruction lineage, and receive independent calibration or review before consequential use.
  • KNOW-R040: A decision record SHALL state how uncertainty affected the selected course, alternatives, safeguards, reversibility, and review trigger.

Here the first run ends and the second begins: KNOW-R045 to KNOW-R048 are this model's later-assigned operating-model clauses, allocated beside the first under the Book of Knowledge's clause namespace allocation, and the two runs read as one law: a duty stated in both is one duty, carried at the broader statement.

  • KNOW-R045: Confidence SHALL record claim, evidence set, assessment method, assessor or process, time, scope, calibration, limitations, and decision consequence.
  • KNOW-R046: Aggregated confidence SHALL preserve input correlation, missingness, disagreement, and the uncertainty introduced by aggregation.
  • KNOW-R047: A confidence state SHALL expire or be reassessed when evidence, scope, environment, consequence, or decision use changes materially.
  • KNOW-R048: Confidence SHALL NOT substitute for authority, consent, safety review, privacy control, or a required human decision.

Confidence states

Stygia may use qualitative states such as Low, Moderate, High, and Undetermined only when their definitions and assessment method are recorded. Undetermined means that the available material does not support a bounded confidence statement; it does not mean that the claim is false. Confidence states shall not replace the evidence classes in KNOW-2.

Calibration and review

Calibration compares prior confidence statements with later observed outcomes over a defined population and time horizon. It shall record selection effects, missing outcomes, changes in the underlying environment, and cases excluded from analysis. A calibration result may show overconfidence, underconfidence, or insufficient data; it shall not retroactively rewrite the original record.

Operating model and interpretation cases

Confidence review separates the claim, evidence set, method, calibration, scope, uncertainty source, decision consequence, and review trigger. A qualitative state is useful when its definition and evidence basis are visible. A numerical value is meaningful only with its interpretation, reference population, units, and calibration limits.

  • Conforming: Claim, evidence, method, calibration, scope, uncertainty, consequence, and review trigger are recorded.
  • Prohibited: Confidence is treated as permission or certainty.
  • Boundary: Undetermined confidence narrows the conclusion without declaring the claim false.
  • Failure: Distribution shift, stale data, or disagreement causes reassessment or pause.
  • Loophole: Correlated evidence is aggregated as independent support.
  • Misuse: A model score conceals a missing approval or privacy restriction.
  • Care-control: Uncertainty supports proportionate assistance while preserving consent and review.

Decision boundary

Confidence supports explanation and risk management. It does not create identity, consent, constitutional authority, access rights, or permission to act. Where uncertainty is material and recovery is difficult, the responsible authority shall prefer a reversible, bounded, or deferred course unless a separate governing rule requires otherwise.

Examples and exclusions

A well-calibrated forecast can remain wrong in an individual case. Multiple reports copied from one source are correlated evidence, not independent confirmation. This Draft does not define live model deployment, automated approval, financial risk limits, or external data collection.

Design evidence

Confidence review should take any statement made in this model's forms and ask whether it could still be examined on the day its outcome arrives: the claim fixed, the scale stated, the population named, the limitations standing where they were written. A confidence assembled so that no outcome can reach it is the counterfeit this model exists to expose, and the counterfeit shows in the record long before any outcome is in. Everything the forms ask a statement to carry is sized to that later examination, and none of it is ornament.


Where this document sits

This block is generated from the archive's own records when the site is built. It records position only and creates no authority.