Concept

Measurement uncertainty — where it appears

The spread of values an instrument would report on repeated measurement, distinct from the bias that makes them all wrong together. Only the second can be corrected, which is why the two are budgeted separately in any specification worth reading.

Named by 19 essays across 6 fields — each of them below, with the objects they name alongside it.

One reflectance, four different infrared tails, and what the camera makes of each. The same visible reflectance continued past 780 nm to four different near-infrared values. A spectrophotometer reports only the left-hand part; an unfiltered sensor integrates all of it. The channel spread falls from 0.14 at a tail of 0.05 to 0.02 at 0.85, with nothing about the visible half changed.

Most things are pale in the infrared

A spectrophotometer stops at 780 nanometres and a black cotton shirt reflecting five per cent of visible light reflects more than half the near infrared. The measurement everybody has and the quantity a camera integrates are different quantities, and nothing in the first says so.

imaging · Capture
A camera's spectral sensitivities, with the filter removed. Silicon quantum efficiency times the colour-filter dye times nothing else, per channel, on a grid running to 1100 nm rather than to 780. With the filter removed, 68 per cent of the area under the three curves lies beyond the visible band, and all three curves are the same curve out there.

The grid outside every figure

Every figure here is computed on 380 to 780 nanometres, which is exactly right for an eye and insufficient for a sensor. This field met the first subject the grid cannot hold, and the decision was not to widen it — because widening it honestly is impossible.

imaging · Capture
Everything between the photons and the picture, and what each stage decides. The 8 stages of a camera pipeline. Only the second is physics; every one after it is a decision somebody made, and the reason two cameras pointed at the same scene disagree is that they made different ones.

A photograph is not a measurement

A photograph is a measurement made by an instrument whose kernel nobody published, under an illuminant nobody recorded, corrected by a matrix fitted to somebody else's surfaces, with two thirds of every pixel invented. It supports relative claims well and absolute ones badly, and it is used for the second.

imaging · Capture
How wrong a profile is between its entries. A device profile is a lattice of measured patches and an interpolation. At every node of every grid drawn here the error is exactly zero, which is why a profile checked at its own patches always looks perfect. Sampled between the nodes, a 3-step grid is worst by ΔE00 = 2.93 and a 17-step grid by 0.30. Both are honest tables of the same press.

A profile is a table

An ICC profile does not contain a model of the device it describes. It contains a lattice of patches that were printed and measured, and everything between them is interpolation — which is exactly zero at every patch, so a profile checked against its own target reports perfection for ever.

applied · Delivery
One tolerance decision, about pairs that agree less and less about the spectrum. Every point is a pair of samples that the reference observer reports as exactly ΔE00 1.0 apart — the same number, the same decision, the same line in the same specification. Along the axis is how far apart their two reflectances are. Up the side is the 95th percentile of what two hundred other eyes report. It runs from 1.13 to 2.32. The document records the horizontal line and not the axis it is plotted against.

A tolerance needs a second number

Two samples one colour difference apart can be read almost identically by everybody or two units apart by the worst-off twentieth, and which of those it is depends on how far apart their spectra are — a quantity every spectrophotometer has already measured and none of them prints. Adding it as a second field predicts the population three times better, and at the tolerances where it matters it is right about a third of the decisions the difference alone gets wrong.

difference · Metric
A tolerance of one unit, re-measured under every light in the census. Every one of the 23 pairs behind this figure is at exactly ΔE00 1.000 under D65 by construction. Each bar is what those same pairs measure under another light, after the observer has adapted to it: the line is the median and the bar spans the pairs. A tolerance is written as a property of a pair and it is not one — the light multiplies both members, and the difference between two products is not the product of the difference. The widest row is a lens at twenty against a lens at seventy, spanning 0.81 to 1.57.

One unit in another room

Twenty-three pairs built at exactly ΔE00 1.000 under D65, re-measured under every change of light this site models with the observer adapted to each, come out anywhere between 0.64 and 1.57. A tolerance is written as a property of a pair and it is a property of a pair and a room.

difference · Metric
What each fitted thing in these essays carries, what its data fix, and what is left. Three columns per row: how many numbers the model has, how many the stated data determine, and the difference — the dimension of the family that fits equally well. The third column is the one nobody publishes. A zero there does not mean the model is right; it means it is determined, which is a much weaker property and is compatible with being determined badly, as the camera row is.

A fit can be exact and empty

Every fitted object here reports one number, the residual on the data it was fitted to, and every one of them has two more that nobody publishes — how many of its parameters the data actually determine, and how large the family of equally good answers is. The third column is where the failures live.

limits · Limits
Four departures from the model equation, each at an ordinary strength. What each of the four assumptions inside a colour integral costs, in ΔE₀₀, on a stated sample under a stated light. The wavelength index is a coated printing paper measured with and without the ultraviolet of D50; the range is the same paper integrated from 300 nanometres and from 380; the place index is a pigmented plastic through a four-millimetre radius; the direction index is an eggshell paint beside a window. The spread is a factor of 7.0. This is a ranking of four examples rather than of four departures — each of them can be made larger by choosing a more extreme sample, and the marble in the same collection of materials reaches 12.7 on the index that comes third here.

The departures are larger than the tolerance

A delivery tolerance is written around one ΔE₀₀ and every one of this round's four departures is above it on ordinary material. A specification that names an illuminant, an observer and a tolerance, and does not name a measurement condition, an aperture and a field, has written a number that two honest laboratories can miss each other on by more than the number itself.

difference · Metric
Each departure over forty-two surfaces rather than one. The same six departures measured over a family of forty-two analytic reflectances — an absorption band of stated centre, width and depth — with the smallest, the median, the ninety-fifth percentile and the largest marked. Every one of them spans more than a factor of three, and the ranking between them is not stable across the family: what decides a departure's size is which sample it is asked about, because a departure is a pairing and the sample is one of the two factors. Quoting any single number for what an observer's age is worth is quoting a choice of example.

The ranking is not stable

On a red pigment under daylight the six observer departures run from 2.38 down to 1.20 ΔE₀₀. Over forty-two surfaces two of them change places, the top two separate, and every one spans between a factor of ten and a factor of thirty-five. A chart of six bars is a chart of one example.

eye · Cones
The three cone absorptances at two settings of the age of the lens. Solid and dashed are the same construction at the two ends of twenty years old against seventy. The curves are built from one pigment template through its ocular media, which is the same model its population of two hundred eyes is drawn from. The largest difference between the two sets is 25.5 per cent of the peak, and where it sits along the wavelength axis is what decides which stimuli the two observers disagree about — a departure concentrated in the blue is invisible on a sample with no blue in it.

The observer has no age

Five of the six arguments in this round are spreads — a population differs about them and the mean is a reasonable summary. The lens is not. Everybody's lens yellows in the same direction at about the same rate, so a standard observer with no age is not an average over a population; it is a snapshot of one moment in every reader's life.

eye · Cones
Each departure over forty-two surfaces rather than one. The same six departures measured over a family of forty-two analytic reflectances — an absorption band of stated centre, width and depth — with the smallest, the median, the ninety-fifth percentile and the largest marked. Every one of them spans more than a factor of three, and the ranking between them is not stable across the family: what decides a departure's size is which sample it is asked about, because a departure is a pairing and the sample is one of the two factors. Quoting any single number for what an observer's age is worth is quoting a choice of example.

A tolerance with an observer in it

A delivery tolerance is written in ΔE₀₀ against the 1931 observer, and six departures of that observer combine to about three of the same units on an ordinary saturated sample. A one-unit tolerance is being asked to contain a three-unit uncertainty that nothing in its budget mentions.

difference · Metric
The conditions under which an observer's departure is exactly zero. A departure of the observer is the pairing of something belonging to the observer with something belonging to the stimulus, so emptying either factor empties the product. The axis is logarithmic in what is left when the condition is imposed. Six rows empty the stimulus's factor — a perfectly neutral sample is the same colour for every observer, at any age and any field size — and two empty the observer's, since a gain on each cone and a change of basis are both absorbed exactly. All eight are identities rather than small numbers. The last two are the same two conditions imposed in a published cone space rather than in the observer's own, and they are worth eight and thirteen units: the identity is about the eye, and the arithmetic everybody uses is in somebody else's coordinates.

The instrument is one observer exactly

A spectrophotometer has no observer uncertainty. It computes through the 1931 tables to the last bit, reproducibly, forever — which is its whole value and is also why it cannot report the three colour differences of observer spread between itself and whoever is looking at the sample.

applied · Delivery
Which of the collection's published quantities a departure can be pushed through. The six quantities the previous round recomputed under six different colour-difference units, and whether the same treatment works for a departure. Two do: the adaptation census and the metameric pair both take reflectances and a light, which is what a departure acts on. Four do not, and the reasons are different in each case rather than a single obstacle. A unit is a function applied to the answers, so it can be swapped at the end of any computation; a departure changes the object at the start, so it has to be accepted by every stage in between. That is the practical difference between auditing a convention and auditing a structure.

What this round could not reach

Three audits, about twenty departures and ten conditions, and a longer list of things that were named and not measured. Every item on it is specific, most of them are an afternoon's work, and the reason none was done is the same in every case — the round ran out of round.

limits · Limits
One instrument, one uncertainty, six places on the scale. What an absolute uncertainty of 0.001 in measured reflectance is worth in lightness, at six levels from a four-colour solid to the paper. Nothing about the instrument changes between the rows: what changes is the slope of the lightness function, which is a straight line of 903 units per unit of luminance factor below a luminance of 0.0089 and a cube root above it. At the solid the uncertainty is 0.903 lightness units and at the paper 0.043 — 21 times as much, for the same measurement.

The scale hangs from one measurement

Black point compensation is a straight line between two blacks, and the destination's is a measurement of one patch at the darkest place a spectrophotometer is ever asked to read. The lightness scale's slope is 903 units per unit of luminance factor there and 43 at the paper, so a thousandth of a reflectance is worth nine tenths of a lightness unit at the black and four hundredths at the white. That one number moves a mid grey by nearly a quarter of a delivery tolerance, and no specification names it.

applied · Delivery
Three bounds against the error they bound, over 68 notches under a fluorescent tube. Each notch placed across by its actual colour error from blurring the lamp and the sample separately, and up by a bound on that error, both on logarithmic scales; the dashed diagonal is where a bound equals the error, and a valid bound sits above it. Cauchy–Schwarz with the true window variances is above the diagonal on every notch, a median 14.7 times the error. Estimated from the blurred tables it falls below on 6 of 68, as low as 0.45 of the error. The Bhatia–Davis bound from the tables and declared ranges is above on every notch and a median 196 times the error.

The tables cannot bound what they discarded

A colour computed from a lamp's blurred table and a sample's blurred table is wrong by the covariance the two blurs threw away, and Cauchy–Schwarz bounds a covariance by two variances. With the true variances the bound always holds and sits fifteen times above the error. With variances read from the tables it fails on six of sixty-eight notches under a fluorescent tube and twenty-two under a laser projector — on the line, where the error is largest. A blurred table does not carry the width of a line, and the covariance depends on it.

light · Light
What declaring a narrowest feature buys, and where it stops being true. The median looseness of a Cauchy–Schwarz bound whose lamp variance is bounded by a declared narrowest feature, against the width declared, for a fluorescent tube and a three-laser projector. Each lamp's own Bhatia–Davis bound — the peak declared and nothing else — is the upper dashed line, and the bound with the true variances is the lower one. The marks are the width each lamp's lines actually have. Declaring it truly takes the tube from ×196 to ×86 and the projector from ×30 to ×14. The open circles are declarations the lamp does not meet, where the bound falls below the error.

A declared width buys a factor of two

A colour engine given two separately blurred spectral tables cannot bound its own error from them, and the bound that always holds — the peak declared and nothing else — sits a median 196 times above the error under a fluorescent tube. Adding one number, the width of the lamp's narrowest feature, brings that to 86. It never fails on any declaration the lamp truly meets, it fails on 47 of 68 notches on one it does not, and its rank correlation with the error it bounds is 0.27.

light · Light
Which two lamps to stand a mesopic match between. Every pair of the five lights, by how far a match made under one and set under the other moves as the rod signal's weight into the S channel goes from nothing to equal — the median over forty-two surfaces, which is the signal an experiment has to resolve. The best pair is daylight against phosphor LED at 0.58 ΔE₀₀; the worst is tungsten against fluorescent tube at 0.09, a factor of 6. The count at the right is how many settings it takes to resolve the weight to a tenth at half a colour difference of scatter per setting.

The reference lamp must not move

To measure an uncertain weight, use the condition in which the answer depends on it most. That is right about half of an asymmetric colour match and exactly wrong about the other half: a match measures a difference of two displacements, and a reference field that also moves with the weight cancels the signal the test field carries. Daylight is the least sensitive of five lamps and belongs in every one of the three best pairs — 75 settings against a phosphor LED, 2,804 against the pair of lamps the principle as stated would have chosen.

eye · Cones
How wrong a declared veil makes the unit, for four true veils. On a 100 cd/m² display, the worst error over grey steps from L 2 to L 90 — the size of the natural logarithm of the declared reading over the true one — against the veil declared, for rooms putting 0.1, 0.3, 1 and 3 cd/m² on the screen. Each curve reaches nought at its own true veil and rises on both sides. The flat stretch at the left is declaring almost nothing, which is declaring none: 0.22 for a true veil of 0.1, 0.54 for a true veil of 0.3, 1.14 for a true veil of 1, 1.89 for a true veil of 3.

A guessed veil halves the error

A colour difference that takes a display's absolute luminance leaves out the light a room reflects off the screen, and on an ordinary display in an ordinary room that makes it wrong about the darkest greys by a factor of three. Giving the unit the veil as a declared argument fixes that when the veil is known. The worry was that it never would be — that a guessed argument is no better than none. It is better: any declared veil up to about twice the true one beats declaring none, and one middling guess for every room halves the worst error. What a guess cannot do is reach ten per cent; that needs the veil known within a sixth, which is what a luminance meter aimed at a black screen gives.

difference · Metric
How far a mesopic match moves as the S weight opens: two rooms against one field. For every pair of the five lamps, the median over forty-two surfaces of how far a match moves as the rod signal's weight into the S channel goes from nothing to equal: pale for the two-room match, each half adapted to its own lamp; dark for a bipartite field whose two halves share one adaptation. The field's signal is larger for every pair, by ×1.5 to ×7.7.

One field keeps what two rooms divide out

An asymmetric match can measure how strongly the rods feed the blue-yellow pathway, but set with the observer adapted to each lamp in turn it needs seventy-five settings on the best pair of lamps and hours of waiting between them. Putting the two lamps on the two halves of one field was proposed as the quick version, at the cost of a weaker signal. The signal is not weaker. Under one shared adaptation it is three times stronger for daylight against a white LED, and the best pair needs six settings. The adaptation that makes the slow version slow is also what was dividing the rods' contribution out of each half.

eye · Cones

Named alongside it

The objects these essays reach for when they reach for this one.

SpecificationQuality controlToleranceAcceptabilityΔEIndividual variationObserver metamerismSpectrophotometryStandard observerIlluminantIntegrationInter-instrument agreement

All concepts