Skip to content

Dimension registry

A score that travels without its status eventually gets quoted as a result. The registry is the artifact that stops that: one machine-readable entry per dimension, carrying what the dimension measures, how it is scored, and what evidence stands behind it. A consumer that cannot read a status can refuse to render the number.

FieldHoldsWhy it is required
KeyThe immutable identifier, for example xtask\xtaskJoins a score back to its definition across versions and datasets
NameThe human labelWhat a surface prints; may change freely, unlike the key
QuestionThe single question the dimension answersA dimension whose question cannot be stated in one sentence is more than one dimension
Value spaceThe permitted values, including \abstLets a consumer distinguish a real zero from an absent measurement
MethodModel-judged, rule-scored, or a fold of bothDetermines what calibration would even mean for this dimension
Evidence statusWhat the evidence is, and what it was measured on: fitted, preview, or noneThe claim about validity, defined on the scoring standard
Promotion statusWhether the dimension is in force: live or provisionalA separate axis from the evidence. A dimension can carry strong evidence and not be promoted, and a promoted one can rest on none
Lifecycle statusCurrent or supersededWhich definition is in force
Calibration referenceA pointer to the record, or an explicit nullAn absent record is a fact worth publishing, not a blank
SupersedesThe prior key and version, where one existsKeeps an old figure traceable to the definition that produced it

What a calibration record itself must contain is specified on calibration, not here. The registry points at that record; it does not restate it.

A key is immutable and never reused. Renaming a dimension in place silently rewrites the meaning of every score already published under that key. A changed meaning is a new key that supersedes the old one.

A superseded dimension stays in the registry. Deleting it orphans every figure that cited it. Retention is what lets an old result stay readable without being mistaken for a current one.

Every dimension carries both statuses, always. The common failure is a dimension that is current and provisional at once, published with only the lifecycle half visible, so a live definition reads as a validated measurement. Both axes are independent and both are required.

A null calibration reference is written, not omitted. The same discipline as abstention: the absence of evidence is recorded as an absence, never as a missing field a reader fills in optimistically.

Promotion runs one direction, and only measurement moves it. Shipping a stronger model does not promote anything, because the missing thing is evidence rather than capability. The conditions a dimension has to clear live on the validation protocol, which is their one home. This page does not restate them, because a second copy of a rule is how the weaker version gets quoted.

The registry is where that rule becomes enforceable rather than aspirational: promotion is an edit to a status field with a required calibration reference, so a dimension cannot be promoted without naming the evidence that promoted it.

The registry describes dimensions, not the tasks they are computed over. A biased task distribution produces well-formed entries with sound statuses and misleading scores, and nothing in the registry can see it.

Registry entries are also per-dimension, not per-surface. Two consumers can read the same entry and present it differently, so a status shown correctly in one place can still be dropped in another. The registry makes the status available; it cannot make a surface honor it.

The lifecycle rule above is in force and unmet, which is recorded rather than softened.