Concept

Elasticity — where it appears

The proportional change in a computed answer for a proportional change in one of its inputs, taken over a stated finite span rather than as a derivative. An elasticity of one means an answer that doubles when the input does; a tenth means an input a conclusion barely notices.

Named by 14 essays across 6 fields — each of them below, with the objects they name alongside it.

Every width at the wide end of its span, and at the narrow end. One line per published quantity, each spanning the value it takes when all four declared widths are read at the narrow end of their reported ranges to the value at the wide end, with a marker at the value as declared. The largest span is the deutan margin at a factor of 2.62; the smallest is 1.22. This is the reading the population model's own documentation promised for four phases and nothing ever took. It is not a confidence interval — the four ends are not quantiles and the widths are not independent draws — it is what a reader who distrusts all four at once sees.

A width nobody varied

Five numbers say how much people differ from one another, and every conclusion drawn here about a population rests on them. Each was written down with the range the literature reports beside it, so that a result could be re-read at the pessimistic end. Nothing ever was.

eye · Cones
Which width carries the answer, and which carries the doubt. Two columns of bars over the four things that differ between two pairs of eyes. On the left, the share of the population's disagreement each one accounts for — the attribution quoted here since the population was built, which puts the lens first at 81%. On the right, how much of the doubt each one puts on everything published here, which is its elasticity multiplied by how badly the width itself is known. The macular pigment comes first there, at 0.65 against the lens's 0.60 — a lead of 8%. The two lists agree exactly below the top.

Which measurement is worth making

Four things about an eye differ between people, and they have been ranked here by how much of the answers they carry since the population was built. Ranking them by how much doubt they carry gives a different order, and ranking them by which one takes a published claim closest to failing gives a third.

eye · Cones
A mid-grey's lightness across a continuum of rooms. The lightness a mid-grey is predicted to have, plotted along the continuous surround parameter running from an average room to a dark one. The three rooms the standard tabulates are marked on it: average at the left, dark at the right, and dim 61% of the way between them rather than halfway. The whole span is 9.71 units of lightness and the step from average to dim is 5.64 of it — 58% — so choosing one of the three rows is a decision worth most of the range.

The surround is three rows of a table

An appearance model takes the room as three constants, and the standard tabulates three rooms. Every appearance figure in this collection is drawn at one of them. The parameter they are three points of is continuous, and the middle row is not in the middle.

brain · Appearance
Room is not safety: two orderings of the same three claims. Three pairs of bars, one pair per published statement about the confusion points. The upper bar in each pair is the margin — how far the measured number is from the threshold that makes the statement true, as a ratio. The lower bar is the headroom — the factor by which one declared width of the population model would have to be wrong for the statement to fail. Both start at one, which is the line. Ordered by margin the three read the protan margin, the tritan margin, the deutan margin; ordered by headroom they read the protan margin, the deutan margin, the tritan margin, and the middle two change places. Every one of the three is inside a factor of two of failing, which the margins do not say.

What would have to be wrong

A great many statements here have thresholds written into them, which turns out to make an audit possible — for each one, the smallest change in a declared input that would stop it holding. Most are unreachable. One is inside a factor of one and a third.

limits · Limits
The winner survives the census's own construction; the middle of it does not. One row per perturbation of a constant the adaptation census is built from — the imaginary wall's centre wavelength, its width, its depth, its base, the macular filter's density and the two lens ages — each moved by an amount plausible for that quantity in its own units, up and down, and then all of them together. Each row shows where the five published transforms rank under it. Bradford holds the first column in all 14 rows. The second and third columns, which the table as built separates by six parts in a thousand, change places in 2 of them — so that ordering was never a fact about the transforms.

The census is a construction too

Five of the fourteen changes of light this collection scores adaptation transforms against are not measurements of anything — they are a wall somebody invented, at a wavelength somebody chose. Moving those constants by amounts plausible in their own units moves the mean residual by two fifths and never changes which transform wins.

light · Light
How far each census row moves when the test set's own description does. A grid of bars, one row per change of light in the census and one bar in each row per number that describes the region the test surfaces are drawn from: how saturated they are, how bright, and how far the two modulations may go together. A bar's length is the elasticity — the proportional change in the published residual for a proportional change in that number. Saturation runs from 0.49 to 0.91 and brightness averages 0.104, so a test set's chroma range is nearly everything and its lightness range is nearly nothing. For scale, the largest elasticity found anywhere among this collection's five declared population widths is about a half — and those at least have declared ranges, while these three numbers have never been quoted with one.

Saturation is nearly everything

The set of test surfaces has three numbers describing it, and only one of them matters. How saturated the surfaces are carries an elasticity of about 0.7 on every result computed over them; how bright they are carries 0.10. A test chart's chroma range decides its answer and its lightness range does not.

difference · Metric
How far each census row moves when the test set's own description does. A grid of bars, one row per change of light in the census and one bar in each row per number that describes the region the test surfaces are drawn from: how saturated they are, how bright, and how far the two modulations may go together. A bar's length is the elasticity — the proportional change in the published residual for a proportional change in that number. Saturation runs from 0.49 to 0.91 and brightness averages 0.104, so a test set's chroma range is nearly everything and its lightness range is nearly nothing. For scale, the largest elasticity found anywhere among this collection's five declared population widths is about a half — and those at least have declared ranges, while these three numbers have never been quoted with one.

The input nobody declared

An audit that swept every declared width in this collection found the largest elasticity anywhere to be about a half. The most elastic input turns out to be one that was never declared, never quoted with a range and never varied — and being undeclared is exactly why it escaped the audit that was looking for it.

limits · Limits
How far each census row moves when the test set's own description does. A grid of bars, one row per change of light in the census and one bar in each row per number that describes the region the test surfaces are drawn from: how saturated they are, how bright, and how far the two modulations may go together. A bar's length is the elasticity — the proportional change in the published residual for a proportional change in that number. Saturation runs from 0.49 to 0.91 and brightness averages 0.104, so a test set's chroma range is nearly everything and its lightness range is nearly nothing. For scale, the largest elasticity found anywhere among this collection's five declared population widths is about a half — and those at least have declared ranges, while these three numbers have never been quoted with one.

A chart decides what a camera scores

A camera profile's reported error changes by a factor of five when the test chart's saturation changes, with the camera and its matrix untouched. The elasticity is 0.67 — the same figure, to two per cent, that an entirely unrelated measurement over an entirely unrelated set of surfaces gives.

imaging · Capture
What six of this collection's published numbers do when the unit changes. Six quantities, from six calculations that share nothing: a change of light after an observer has adapted, a camera profile's error, the gap between the two standard observers, a metameric pair under the lamp that breaks it, the same image on two papers, and an observer two seconds into a new room. Each is recomputed under all six units and every unit is calibrated onto ΔE2000's scale first, so the bar is not a change of units in the ordinary sense. The bar is the ratio of the largest reading to the smallest, and it runs from 1.71 to 2.30. Five of the six are printed in ΔE2000 by the essays that report them; the sixth is printed in CAM16-UCS, because the model it comes out of defines that unit.

A choice with no magnitude

An audit can multiply a width by 1.25 and report an elasticity. It cannot multiply CIEDE2000 by anything. Auditing a structural choice needs a different instrument, and building one shows that six published numbers in this collection each carry a factor of about two of unit-choice — after the change of scale has been taken out.

limits · Limits
How far each unit is from being a rescaling of the one this collection publishes in. One row per unit on the menu. The bar is the root-mean-square scatter about that unit's own best rescaling of ΔE2000, over 374 pairs of surfaces differing by a fraction of a unit to about ten. A bar of zero would mean the unit is ΔE2000 in different money — every printed number would change and no conclusion would. ΔE2000's own row is zero by construction and is the check that the table is computed the right way round. The two units that divide a chroma difference by the chroma it was measured at, ΔE94 at 15 per cent and CAM16-UCS at 24, are closer to it than the three that do not, which run from 28 to 35. The split is by weighting and not by whether the unit is a matching difference or an appearance one.

The weighting is the disagreement

Five colour-difference formulae, three decades and two committees, and the single property that predicts which of them agree is whether a chroma difference gets divided by the chroma it was measured at. It sorts the menu exactly, it cuts across the distinction between a matching difference and an appearance one, and it halves the census's largest sensitivity.

difference · Metric
The census along the line from ΔE76 to ΔE94, and past it. ΔE94 is ΔE76 with two weighting constants in it, and at zero those constants make every weight exactly one, so the two formulae are joined by a line rather than separated by a choice. The horizontal axis is how much of the published weighting is applied: 0 is exactly ΔE76, 1 is exactly ΔE94, and 3 is three times more weighting than anybody has proposed. The falling curve is the census's mean elasticity to how saturated its test set is, which drops from 1.11 to 0.74 — most of the fall happening before the published value is reached. The other curve is Kendall's τ against ΔE2000's ranking, and it peaks at w = 0.5, not at 1: the weighting that best reproduces the published ordering is about half the published weighting. There is no value of this dial that reaches ΔE2000, whose rotation term is not on this line at all.

A dial through a discrete menu

ΔE*94 is ΔE*ab with two weighting constants in it, and at zero those constants make every weight exactly one — so the two ends of the oldest disagreement in colour difference are joined by a line rather than separated by a choice. Walking it gives a derivative where a menu gives only a spread, and the derivative says the published weighting is on the far side of the interesting part.

difference · Metric
Two sensitivities from two libraries, under every unit. Two quantities that share no code, no test set and no physical question: how much the adaptation census's residual depends on how saturated its surfaces are, and how much a camera profile's reported error depends on how saturated its test chart is. The first is a mean over fourteen changes of light built from cosine combinations; the second is one number about one silicon sensor scored on Gaussian bumps. Under the published unit they sit at 0.687 and 0.656. Across the whole menu they move together, from about 0.5 under the appearance unit to about 1.15 under plain CIELAB, staying within 12 per cent of each other at the worst point. Two numbers agreeing once is a coincidence; two curves agreeing at six points across a factor of two and a half is a shared mechanism, and the mechanism is the compression the unit applies to a chroma difference.

The coincidence was a mechanism

Two sensitivities from two libraries with no shared code came out two per cent apart, and the claim made about them was that they share a mechanism rather than a number. That claim has a colour-difference formula inside it, so it can be tested by changing the formula — and both curves move together across the whole menu, from 0.5 to 1.15.

imaging · Capture
What six of this collection's published numbers do when the unit changes. Six quantities, from six calculations that share nothing: a change of light after an observer has adapted, a camera profile's error, the gap between the two standard observers, a metameric pair under the lamp that breaks it, the same image on two papers, and an observer two seconds into a new room. Each is recomputed under all six units and every unit is calibrated onto ΔE2000's scale first, so the bar is not a change of units in the ordinary sense. The bar is the ratio of the largest reading to the smallest, and it runs from 1.71 to 2.30. Five of the six are printed in ΔE2000 by the essays that report them; the sixth is printed in CAM16-UCS, because the model it comes out of defines that unit.

Three choices reached

Two rounds ago this collection named three things it rested on and could not audit — a unit, a diagonal and a template. All three are now reached, and the interesting part is not the three answers but that four of the round's own predictions were refused by the arithmetic and one of its measurements was wrong in a way only a cross-check caught.

limits · Limits
Which of the collection's published quantities a departure can be pushed through. The six quantities the previous round recomputed under six different colour-difference units, and whether the same treatment works for a departure. Two do: the adaptation census and the metameric pair both take reflectances and a light, which is what a departure acts on. Four do not, and the reasons are different in each case rather than a single obstacle. A unit is a function applied to the answers, so it can be swapped at the end of any computation; a departure changes the object at the start, so it has to be accepted by every stage in between. That is the practical difference between auditing a convention and auditing a structure.

A departure is not a unit

The previous round audited six published quantities by swapping the unit they were quoted in — a function applied at the end of each computation. Nothing of that shape works here. A departure changes the object at the start, so every stage in between has to accept it, and only two of the same six quantities can take one at all. The fourth cannot even be expressed in the interface.

limits · Limits

Named alongside it

The objects these essays reach for when they reach for this one.

Test setSensitivityColour differenceChromatic adaptationDeclared inputCalibrationChromaDegrees of freedomCamera profileCIEDE2000Rank correlationResidual

All concepts