Concept

Uncertainty — where it appears

How much a published number could have been otherwise, which has at least two independent parts that are routinely conflated. A standard error says how far the answer would move on a different sample of the same inputs; the arbitrariness of the construction says how far it would move under a different defensible choice of formula, set or model, and the second is usually the larger.

Named by 12 essays across 5 fields — each of them below, with the objects they name alongside it.

Which steps of the census ranking a change of unit reverses. Every adjacent pair in the published census ranking that at least one unit puts the other way round. The bar counts how many of the five other units reverse it. The marker on the left says whether the test set had already declared the pair unresolved — a gap smaller than twice its own paired standard error, which is a statement about sampling over 125 surfaces and shares no arithmetic with a change of ruler. The two pairs every unit reverses are both flagged, which is the agreement. The pair at the bottom is the disagreement: the test set resolves it at 9.1 standard errors and four of the five units reverse it anyway, because a sampling error cannot see a change of ruler and a change of ruler cannot see a sampling error.

Two instruments and one ranking

A sampling error over a hundred and twenty-five surfaces and a change of colour-difference formula share no arithmetic at all, and they were asked the same question of the same table. Every adjacency the whole menu reverses had already been flagged as unresolved. And one the test set settles at nine standard errors is reversed by four of the five formulae, which is what makes them two instruments rather than one.

limits · Limits
Where on the scale the units disagree. The reference pairs split into bands by how far apart they are in ΔE2000, with each unit's root-mean-square relative departure from the published one plotted per band. Every unit is calibrated once, over the whole sample, so a band is not refitted and the shape is the effect rather than an artefact of fitting. Every one of the five falls: the disagreement is proportionally largest on the pairs that are closest together, which is the opposite of what being fitted to threshold data would suggest. The appearance unit is the extreme case, at 91 per cent on the narrowest band and 17 on the widest, because CAM16-UCS raises its distance to the power 0.63 and a power below one inflates small differences against large ones. In absolute terms every curve here runs the other way — the widest band disagrees by 1.16 to 2.37 ΔE₀₀-equivalent against 0.14 to 0.68 on the narrowest — so which reading is right depends on whether the published quantity is a level or a ratio. This is the mechanism behind the census's own behaviour, where the mildest rows spread furthest across the menu.

The disagreement is at the near end

Every colour-difference formula on the menu was fitted to threshold data, so the expectation is that they agree about pairs an observer can only just tell apart and diverge on large differences. They do the opposite. Proportionally the disagreement is largest at the near end, by a factor of six for the appearance unit, and the cause is an exponent of 0.63.

difference · Metric
What six of this collection's published numbers do when the unit changes. Six quantities, from six calculations that share nothing: a change of light after an observer has adapted, a camera profile's error, the gap between the two standard observers, a metameric pair under the lamp that breaks it, the same image on two papers, and an observer two seconds into a new room. Each is recomputed under all six units and every unit is calibrated onto ΔE2000's scale first, so the bar is not a change of units in the ordinary sense. The bar is the ratio of the largest reading to the smallest, and it runs from 1.71 to 2.30. Five of the six are printed in ΔE2000 by the essays that report them; the sixth is printed in CAM16-UCS, because the model it comes out of defines that unit.

The observers differ by a unit's worth

The gap between the 1931 and 1964 standard observers is the one quantity in this collection's audit with no published number under it — nothing here reports it as a single figure over a stated set. It also has the second-largest dependence on which colour-difference formula is used, running from 1.60 to 3.55 across the menu.

matching · Gamut
What six of this collection's published numbers do when the unit changes. Six quantities, from six calculations that share nothing: a change of light after an observer has adapted, a camera profile's error, the gap between the two standard observers, a metameric pair under the lamp that breaks it, the same image on two papers, and an observer two seconds into a new room. Each is recomputed under all six units and every unit is calibrated onto ΔE2000's scale first, so the bar is not a change of units in the ordinary sense. The bar is the ratio of the largest reading to the smallest, and it runs from 1.71 to 2.30. Five of the six are printed in ΔE2000 by the essays that report them; the sixth is printed in CAM16-UCS, because the model it comes out of defines that unit.

Three choices reached

Two rounds ago this collection named three things it rested on and could not audit — a unit, a diagonal and a template. All three are now reached, and the interesting part is not the three answers but that four of the round's own predictions were refused by the arithmetic and one of its measurements was wrong in a way only a cross-check caught.

limits · Limits
The six arguments a surface's response has, and the one this model keeps. A surface's response to light is a function of six arguments: the wavelength, direction and place the light arrives with, and the wavelength, direction and place it leaves with. The model every colour here is computed from keeps one number per wavelength, which means it takes the diagonal of the first pair, integrates the second away, and assumes the third pair equal. Each departure drawn here restores one of them. The fourth departure is not on the diagram: the wavelength grid is the range of the index that was kept rather than an index that was dropped, which is why it is the cheapest of the four to fix and was still not fixed.

The audit that changed the object

Three rounds audited this collection's numbers, its sets and its conventions, and each one ended by naming the same thing it could not reach — that a surface is a reflectance. This round reached it, found that four departures share one algebraic form, had two of its predictions refused, and discovered that its own hemisphere quadrature had been passing a convergence check by luck.

limits · Limits
The angle between two departures, which nothing in the audit records. Every pair of the six departures on every surface — 2520 pairs — binned by the angle between their two deviations in the local metric. The distribution reaches both ends: 440 pairs sit under thirty degrees and point almost the same way, and 615 sit above a hundred and fifty and point almost opposite. On 1095 of the 2520 the two together cost less than the larger of them alone. A table of magnitudes cannot say which case it is in.

A size is not a direction

An audit that reports magnitudes cannot say what two of them cost together. Over two and a half thousand pairs of observer departures the angle between them in the local metric runs from one degree to a hundred and seventy-nine, and on forty-three per cent of them the two together cost less than the larger of the two alone. Adding the angle predicts the composition to one and a quarter per cent; Pythagoras is out by twenty-eight.

matching · Gamut
What a code lattice costs, and where. Twelve thousand colours quantised to 8 bits per channel through the sRGB transfer function and read back, with lightness across the bottom and the colour difference the rounding cost up the side. The mean is 0.191 and the worst case is 1.15, a factor of 6.0. The bars are band means, and they rise: the encoding spends its codes in the shadows, so the top of the ramp is where the lattice is coarsest against a metric that does not compress as hard.

A lattice has no derivative

Every departure priced here was priced by perturbing something and reading the answer, which requires the thing being perturbed to have a derivative. A file written on a code lattice does not have one — its output is flat almost everywhere and jumps on a set of measure zero — so quantisation can be bounded and never propagated. The bound is 1.15 colour differences at eight bits per channel against a mean of 0.19, and it is worst where the encoding spends fewest codes.

matching · Gamut
The audit's six departures, read twice. Each departure priced in the matching unit it was published in and in the appearance unit an appearance prediction would be judged by. The model does not scale them by one factor: it amplifies the smallest by 1.58 and the largest by 1.03, so the range between them narrows from 3.34 to 2.17. The mean ranking is unchanged and 14 of the 42 surfaces reorder.

The audit, read as appearances

The previous round priced six observer departures in a matching unit and left them there. Read in the appearance unit an appearance prediction would actually be judged by, they are not the same six numbers scaled by a constant — the model amplifies the smallest by 1.58 and the largest by 1.03, so the range between the largest and smallest term narrows from 3.34 to 2.17. The published ranking survives in the mean and changes on fourteen of the forty-two surfaces it was averaged over.

brain · Appearance
What a stated lightness pins down, and where. A lightness quoted to 0.05 of a unit, inverted, and the luminance it fixes. Read as a fraction of the colour's own luminance the requirement is 0.90 per cent at J 10 and 0.097 at J 95, a factor of 9.3. Read in absolute luminance it is the other way round, by a factor of 6.5. Both readings are true and they answer different questions.

A stated lightness is two requirements

A specification quotes an appearance to a stated precision — a lightness to a tenth of a unit, say — and the same precision everywhere. Inverted, a twentieth of a unit of lightness fixes the luminance to nine tenths of one per cent at the bottom of the scale and to a tenth of one per cent at the top, a factor of 9.3. Read in absolute luminance it is the other way round by a factor of 6.5, and both readings are true.

brain · Appearance
The budget's three numbers, and the units they are in. The published three-stage budget's own figures, with each one's unit named, beside the same stage re-measured in a single unit over the same colours. Two of the three are colour differences between stimuli and the third is a distance between appearances, and the budget adds them. The fourth row is a stage the budget has no entry for: the colours the separation cannot reach even after the mapping has moved them, which comes to 1.45.

The budget adds two units

This collection publishes a three-stage error budget for a colour-management chain and prints its sum. Two of the three stages are colour differences between stimuli and the third is a distance between appearances, and the conversion between those is not a constant. Re-measured in one unit over the same colours the chain has four stages rather than three, its end-to-end error is 5.05 against a sum of 7.66, and the profile's contribution where a job actually lands is 3.2 times its published average.

applied · Delivery
The angle between the lens and the macular pigment, before and after adaptation. On each of 120 surfaces the angle, in the local metric, between what an older lens does to the reading and what a denser macular pigment does, binned in ten-degree steps. Read without adaptation, where both filters yellow the observer's white along with everything else, the two point nearly the same way: a median of 8 degrees. Read after each observer has adapted to its own white, the median is 156, and the two together cost less than the larger alone on 115 of the 120.

Two yellow filters cancel on a slope

An older lens and a denser macular pigment both take blue out of the light, and read before adaptation they move a colour in nearly the same direction, eight degrees apart. Once each eye has adapted to its own white they point a median 156 degrees apart on smooth reflectances and together cost less than the lens alone. On surfaces with a narrow absorption band they still sit 26 degrees apart and add. What decides it is the width of the surface's own features.

difference · Metric
How far the chain falls short of its sum, and how much of that is the unit. For each rendering intent, three ratios of the chain's end-to-end error to the sum of its four stages. The lowest bar is the published one, in the power-corrected unit. The middle bar is what that unit reports for a chain whose stages point the same way and add exactly — the exponent on its own. The top bar is the same chain measured in the model's own Euclidean space. Under the colorimetric intent the published ratio is 0.66, the exponent alone gives 0.74 and the chain in a space that can add 0.84: 77 per cent of the shortfall is the unit.

A chain measured in a unit that cannot add

A delivery chain's four stages, measured in the power-corrected appearance difference, sum to 7.66 while the chain end to end measures 5.05, and the shortfall was read as the stages partly cancelling. A chain whose four stages lay in a straight line and added exactly would still read 0.74 of its sum in that unit, because a distance raised to the power 0.63 cannot add. Measured in the model's own Euclidean space the same chain reaches 0.84 of its sum. Three quarters of the published shortfall was the exponent.

applied · Delivery

Named alongside it

The objects these essays reach for when they reach for this one.

Colour differenceDeclared inputSensitivitySpecificationTest setAuditCalibrationChromatic adaptationCIECAM16Structural choiceWorst caseAppearance model

All concepts