Concept

Sensitivity — where it appears

How far a computed answer moves when one of its inputs does, usually reported as an elasticity over a stated span. It is what says which inputs are worth measuring better, and it ranks them differently from how much each contributes to the answer itself.

Named by 23 essays across 7 fields — each of them below, with the objects they name alongside it.

How far each census row moves when the test set's own description does. A grid of bars, one row per change of light in the census and one bar in each row per number that describes the region the test surfaces are drawn from: how saturated they are, how bright, and how far the two modulations may go together. A bar's length is the elasticity — the proportional change in the published residual for a proportional change in that number. Saturation runs from 0.49 to 0.91 and brightness averages 0.104, so a test set's chroma range is nearly everything and its lightness range is nearly nothing. For scale, the largest elasticity found anywhere among this collection's five declared population widths is about a half — and those at least have declared ranges, while these three numbers have never been quoted with one.

Saturation is nearly everything

The set of test surfaces has three numbers describing it, and only one of them matters. How saturated the surfaces are carries an elasticity of about 0.7 on every result computed over them; how bright they are carries 0.10. A test chart's chroma range decides its answer and its lightness range does not.

difference · Metric
How far each census row moves when the test set's own description does. A grid of bars, one row per change of light in the census and one bar in each row per number that describes the region the test surfaces are drawn from: how saturated they are, how bright, and how far the two modulations may go together. A bar's length is the elasticity — the proportional change in the published residual for a proportional change in that number. Saturation runs from 0.49 to 0.91 and brightness averages 0.104, so a test set's chroma range is nearly everything and its lightness range is nearly nothing. For scale, the largest elasticity found anywhere among this collection's five declared population widths is about a half — and those at least have declared ranges, while these three numbers have never been quoted with one.

The input nobody declared

An audit that swept every declared width in this collection found the largest elasticity anywhere to be about a half. The most elastic input turns out to be one that was never declared, never quoted with a range and never varied — and being undeclared is exactly why it escaped the audit that was looking for it.

limits · Limits
How far each census row moves when the test set's own description does. A grid of bars, one row per change of light in the census and one bar in each row per number that describes the region the test surfaces are drawn from: how saturated they are, how bright, and how far the two modulations may go together. A bar's length is the elasticity — the proportional change in the published residual for a proportional change in that number. Saturation runs from 0.49 to 0.91 and brightness averages 0.104, so a test set's chroma range is nearly everything and its lightness range is nearly nothing. For scale, the largest elasticity found anywhere among this collection's five declared population widths is about a half — and those at least have declared ranges, while these three numbers have never been quoted with one.

A chart decides what a camera scores

A camera profile's reported error changes by a factor of five when the test chart's saturation changes, with the camera and its matrix untouched. The elasticity is 0.67 — the same figure, to two per cent, that an entirely unrelated measurement over an entirely unrelated set of surfaces gives.

imaging · Capture
Every census row under five constructions of the same test set. A slope chart with 5 columns — lattice, coarse, fine, uniform, natural — and one line per change of light in the census, each line joining that row's mean residual under each construction. Four of the five columns describe the same region of surfaces walked at different densities or against different measures; the last is the clamped, realistic family, which is not linear in its parameters and is therefore answering a slightly different question. The levels move: between the coarse and fine lattices every row shifts by seven to nine per cent, in the same direction, which is a common-mode factor no published residual here has ever carried. The order almost survives. Inside the region exactly one pair crosses, and it is the pair the standard error had already flagged; under the clamped set two more cross, including one the error separates by nearly nine standard errors. The crossing lines are drawn heavy.

What the audit still cannot reach

Two rounds have now swept every declared width in this collection and one of its structural choices. Three structural choices remain, none of them has a multiplier to sweep, and the reason each resists is different — which makes the list a description of where this kind of audit ends rather than a queue of work.

limits · Limits
What six of this collection's published numbers do when the unit changes. Six quantities, from six calculations that share nothing: a change of light after an observer has adapted, a camera profile's error, the gap between the two standard observers, a metameric pair under the lamp that breaks it, the same image on two papers, and an observer two seconds into a new room. Each is recomputed under all six units and every unit is calibrated onto ΔE2000's scale first, so the bar is not a change of units in the ordinary sense. The bar is the ratio of the largest reading to the smallest, and it runs from 1.71 to 2.30. Five of the six are printed in ΔE2000 by the essays that report them; the sixth is printed in CAM16-UCS, because the model it comes out of defines that unit.

A choice with no magnitude

An audit can multiply a width by 1.25 and report an elasticity. It cannot multiply CIEDE2000 by anything. Auditing a structural choice needs a different instrument, and building one shows that six published numbers in this collection each carry a factor of about two of unit-choice — after the change of scale has been taken out.

limits · Limits
The census along the line from ΔE76 to ΔE94, and past it. ΔE94 is ΔE76 with two weighting constants in it, and at zero those constants make every weight exactly one, so the two formulae are joined by a line rather than separated by a choice. The horizontal axis is how much of the published weighting is applied: 0 is exactly ΔE76, 1 is exactly ΔE94, and 3 is three times more weighting than anybody has proposed. The falling curve is the census's mean elasticity to how saturated its test set is, which drops from 1.11 to 0.74 — most of the fall happening before the published value is reached. The other curve is Kendall's τ against ΔE2000's ranking, and it peaks at w = 0.5, not at 1: the weighting that best reproduces the published ordering is about half the published weighting. There is no value of this dial that reaches ΔE2000, whose rotation term is not on this line at all.

A dial through a discrete menu

ΔE*94 is ΔE*ab with two weighting constants in it, and at zero those constants make every weight exactly one — so the two ends of the oldest disagreement in colour difference are joined by a line rather than separated by a choice. Walking it gives a derivative where a menu gives only a spread, and the derivative says the published weighting is on the far side of the interesting part.

difference · Metric
Two sensitivities from two libraries, under every unit. Two quantities that share no code, no test set and no physical question: how much the adaptation census's residual depends on how saturated its surfaces are, and how much a camera profile's reported error depends on how saturated its test chart is. The first is a mean over fourteen changes of light built from cosine combinations; the second is one number about one silicon sensor scored on Gaussian bumps. Under the published unit they sit at 0.687 and 0.656. Across the whole menu they move together, from about 0.5 under the appearance unit to about 1.15 under plain CIELAB, staying within 12 per cent of each other at the worst point. Two numbers agreeing once is a coincidence; two curves agreeing at six points across a factor of two and a half is a shared mechanism, and the mechanism is the compression the unit applies to a chroma difference.

The coincidence was a mechanism

Two sensitivities from two libraries with no shared code came out two per cent apart, and the claim made about them was that they share a mechanism rather than a number. That claim has a colour-difference formula inside it, so it can be tested by changing the formula — and both curves move together across the whole menu, from 0.5 to 1.15.

imaging · Capture
How far this site's median observer sits from the 1931 standard, by template. Twenty-four natural reflectances under D65, each given a tristimulus value twice: once by the 1931 colour-matching functions and once by this site's median member, with each judged against its own white. The bar is the mean difference, which is the residual this collection bounds and calls inescapable. It is inescapable, and it is smallest for the simpler template: Lamb's 1995 nomogram gives 0.9006 against Govardovskii's 0.9522, and removing Govardovskii's secondary band brings it down again to 0.9354. Neither is an argument for changing template — a nomogram is fitted to measurements of individual receptors, not to colour matches, so agreement with the standard observer is not what either was trying to achieve. What it says is that the residual is a mismatch between two kinds of observer rather than a shortfall a better pigment model would close. The number after each bar is the template's tail ratio: how far the L cone's half-maximum reaches below its peak against how far it reaches above.

The population rests on a template

Two hundred observers here are built from one formula fitted to microspectrophotometry in 2000. The obvious alternative — the tabulated cone fundamentals — is not available, and the reason is the finding. A tabulated fundamental has no peak wavelength to move, so the moment it is used the population collapses to a single observer.

eye · Cones
A template's asymmetry against what its observer costs. The horizontal axis is the tail ratio of the L cone's pigment absorbance — how far the curve reaches below its peak at half maximum against how far it reaches above — and the vertical is how far the observer built from that template sits from the 1931 standard. A real visual pigment has a long short-wavelength tail, so the three curves derived from a published nomogram sit above 1.1 and the two Gaussians sit below. The four are matched in width, so nothing here is about size. The ordering is the point: the two caricatures cost between two and four times what either nomogram does, and the axis they are separated on is the one feature the caricatures do not have.

A template is mostly its tail

Four pigment templates matched to the same width at the same peak, differing only in which side of the peak their half-maximum reaches further. Ordered by that one number, the observers they build are ordered by how far they sit from the standard one — and the two with the tail on the wrong side cost two and four times what either real nomogram costs.

eye · Cones
The secondary band, dialled from nothing to twice what Govardovskii published. Govardovskii's template has a second, smaller absorption band below 400 nm, published at 0.26 of the α-band's peak. It is the one coefficient in the whole template whose contribution is somewhere else in the spectrum than the peak, and it is the part of the template a reader is least likely to have heard of. Sweeping it from nothing to twice the published value moves the median observer's distance from the 1931 standard from 0.9354 to 0.9790 ΔE₀₀, monotonically upwards. That is a small effect — about a fiftieth of the residual — and its being monotone is the interesting part: there is no interior optimum, so nothing here recommends the published value over any other, and a coefficient whose measured value is not the one that best fits an unrelated agreement is a coefficient to leave where the measurement put it.

The band below four hundred

The pigment template every observer here is built from carries a second, smaller absorption band in the ultraviolet, published at 0.26 of the main one. Dialling it from nothing to twice that moves the median observer monotonically away from the standard one, with no interior optimum — which is what a physical constant looks like when nothing downstream is pulling it anywhere.

eye · Cones
What six of this collection's published numbers do when the unit changes. Six quantities, from six calculations that share nothing: a change of light after an observer has adapted, a camera profile's error, the gap between the two standard observers, a metameric pair under the lamp that breaks it, the same image on two papers, and an observer two seconds into a new room. Each is recomputed under all six units and every unit is calibrated onto ΔE2000's scale first, so the bar is not a change of units in the ordinary sense. The bar is the ratio of the largest reading to the smallest, and it runs from 1.71 to 2.30. Five of the six are printed in ΔE2000 by the essays that report them; the sixth is printed in CAM16-UCS, because the model it comes out of defines that unit.

Three choices reached

Two rounds ago this collection named three things it rested on and could not audit — a unit, a diagonal and a template. All three are now reached, and the interesting part is not the three answers but that four of the round's own predictions were refused by the arithmetic and one of its measurements was wrong in a way only a cross-check caught.

limits · Limits
Where a sample's colour goes as the aperture closes. The a and b of three translucent materials as the measuring aperture narrows from forty millimetres to one. Each track starts at the open circle, which is the colour the model says the sample has, and ends at the filled one. The axes cross at the neutral point. pale marble passes through neutral at a radius of 5.32 millimetres and comes out on the other side; candle wax passes through neutral at a radius of 7.07 millimetres and comes out on the other side; skin passes through neutral at a radius of 0.76 millimetres and comes out on the other side. Nothing about the sample changed: the aperture is a filter with a colour of its own, and the colour is decided by how the sample scatters rather than by what it absorbs.

The hue the hole decides

A piece of pale marble measured through a wide aperture is faintly yellow. Measured through a narrow one it is faintly blue, and between the two there is an aperture at which it is exactly neutral. Nothing about the stone changes; the aperture is a filter with a colour, and what decides that colour is the size of the particles rather than the pigment between them.

scene · Scene
Which of the collection's published quantities a departure can be pushed through. The six quantities the previous round recomputed under six different colour-difference units, and whether the same treatment works for a departure. Two do: the adaptation census and the metameric pair both take reflectances and a light, which is what a departure acts on. Four do not, and the reasons are different in each case rather than a single obstacle. A unit is a function applied to the answers, so it can be swapped at the end of any computation; a departure changes the object at the start, so it has to be accepted by every stage in between. That is the practical difference between auditing a convention and auditing a structure.

A departure is not a unit

The previous round audited six published quantities by swapping the unit they were quoted in — a function applied at the end of each computation. Nothing of that shape works here. A departure changes the object at the start, so every stage in between has to accept it, and only two of the same six quantities can take one at all. The fourth cannot even be expressed in the interface.

limits · Limits
Each departure over forty-two surfaces rather than one. The same six departures measured over a family of forty-two analytic reflectances — an absorption band of stated centre, width and depth — with the smallest, the median, the ninety-fifth percentile and the largest marked. Every one of them spans more than a factor of three, and the ranking between them is not stable across the family: what decides a departure's size is which sample it is asked about, because a departure is a pairing and the sample is one of the two factors. Quoting any single number for what an observer's age is worth is quoting a choice of example.

The census under another observer

This collection's largest computed result is an adaptation census — fourteen changes of light judged over a hundred and twenty-five constructed surfaces. Every number in it was computed through one observer, and the observer's own departures are between one and two and a half units on the same surfaces, which is the size of the effects the census reports.

brain · Appearance
The two tabulation choices over forty-two surfaces, under a 6500 K thermal radiator. Each column is one choice, measured over a family of forty-two analytic reflectances rather than on a single example: an absorption band of stated centre, width and depth. The four marks are the smallest, the median, the ninety-fifth percentile and the largest cost in ΔE₀₀, logarithmically. Under a smooth light the range is worth 9.1 times the step at the median, so a collection wanting one repair should widen its range rather than refine its step — and under a fluorescent tube the ranking reverses outright.

The grid under the census

The adaptation census is computed on eighty-one wavelengths from 380 to 780 nanometres. Under the daylight and blackbody sources it uses, the range is worth about half a colour difference on ordinary surfaces and the step about a twentieth — so the census carries a tabulation term as well as an observer one, and they are not the same size.

brain · Appearance
Which of the collection's published quantities a departure can be pushed through. The six quantities the previous round recomputed under six different colour-difference units, and whether the same treatment works for a departure. Two do: the adaptation census and the metameric pair both take reflectances and a light, which is what a departure acts on. Four do not, and the reasons are different in each case rather than a single obstacle. A unit is a function applied to the answers, so it can be swapped at the end of any computation; a departure changes the object at the start, so it has to be accepted by every stage in between. That is the practical difference between auditing a convention and auditing a structure.

What this round could not reach

Three audits, about twenty departures and ten conditions, and a longer list of things that were named and not measured. Every item on it is specific, most of them are an afternoon's work, and the reason none was done is the same in every case — the round ran out of round.

limits · Limits
One surface dimmed sixteen times, and the two things that happen to it. A single surface, dimmed by successive halvings, with the lens age departure measured on it at every level. The tristimulus deviation falls by exactly the dimming factor — 16 times over the sweep, to the last bit, because the colour integral is linear in the stimulus. What a unit of that deviation is worth rises by 8.2 times over the same sweep. The colour difference the audit reports is the product of the two, and it falls by only 1.95.

A deviation is not a difference

The previous round priced twenty departures in colour differences and treated each number as a property of the thing that departed. Every one of them is a product of two factors — how far the reading moved, which is linear and belongs to the departure, and what a unit of that movement is worth where it landed, which is not linear and belongs to the colour. Dimming one surface sixteen times scales the first by exactly sixteen and the second by eight.

matching · Gamut
The angle between two departures, which nothing in the audit records. Every pair of the six departures on every surface — 2520 pairs — binned by the angle between their two deviations in the local metric. The distribution reaches both ends: 440 pairs sit under thirty degrees and point almost the same way, and 615 sit above a hundred and fifty and point almost opposite. On 1095 of the 2520 the two together cost less than the larger of them alone. A table of magnitudes cannot say which case it is in.

A size is not a direction

An audit that reports magnitudes cannot say what two of them cost together. Over two and a half thousand pairs of observer departures the angle between them in the local metric runs from one degree to a hundred and seventy-nine, and on forty-three per cent of them the two together cost less than the larger of the two alone. Adding the angle predicts the composition to one and a quarter per cent; Pythagoras is out by twenty-eight.

matching · Gamut
The straight piece under the cube root, and where it stops. CIELAB's lightness against relative luminance, over the bottom 5.0 per cent of the range. Below Y = 0.008856 it is a straight line of slope 7.787; above it, a cube root. The two meet at L* 8 in value and in slope, exactly — the CIE's two constants are chosen to make that true. The dashed curve is the pure cube root, which reaches negative lightness before it reaches zero luminance and has an infinite slope there. The break is marked, and the axis runs to Y = 0.050.

The straight piece under the cube root

CIELAB's lightness is described everywhere as a cube root and below a stated luminance it is a straight line, spliced on with two constants chosen so the join is exact in value and in slope. Every black a delivery chain reaches is inside that straight piece — a press black at L* 2.4, a projected black at 1.1 — where the compression does not compress, the price of a deviation is flat to two parts in a thousand, and the second derivative the composition of two departures needs does not exist.

matching · Gamut
What a code lattice costs, and where. Twelve thousand colours quantised to 8 bits per channel through the sRGB transfer function and read back, with lightness across the bottom and the colour difference the rounding cost up the side. The mean is 0.191 and the worst case is 1.15, a factor of 6.0. The bars are band means, and they rise: the encoding spends its codes in the shadows, so the top of the ramp is where the lattice is coarsest against a metric that does not compress as hard.

A lattice has no derivative

Every departure priced here was priced by perturbing something and reading the answer, which requires the thing being perturbed to have a derivative. A file written on a code lattice does not have one — its output is flat almost everywhere and jumps on a set of measure zero — so quantisation can be bounded and never propagated. The bound is 1.15 colour differences at eight bits per channel against a mean of 0.19, and it is worst where the encoding spends fewest codes.

matching · Gamut
Where averaging the model's answers is and is not averaging its argument. Each row takes a spread of situations, averages the model's predictions across them, and compares that against the model's prediction for the average situation. The bar is the gap as a share of the spread itself. Over the differences between observers it is 1.4 per cent — the model is very nearly linear there. Over the range of adapting luminance one room covers in a day it is 59 per cent, and from indoors to outdoors 74.

Where the model's curve does not matter

This round has been about what a nonlinearity does to an average, and CIECAM16 is the most nonlinear thing in the collection. Over the spread a population of observers produces, the average of its predictions is its prediction for the average to within 1.4 per cent — the nonlinearity is there and the excursion is too small to reach it. Over the range of adapting luminance one room covers in a day, the same gap is 59 per cent of the spread, and from indoors to outdoors 74.

brain · Appearance
The audit's six departures, read twice. Each departure priced in the matching unit it was published in and in the appearance unit an appearance prediction would be judged by. The model does not scale them by one factor: it amplifies the smallest by 1.58 and the largest by 1.03, so the range between them narrows from 3.34 to 2.17. The mean ranking is unchanged and 14 of the 42 surfaces reorder.

The audit, read as appearances

The previous round priced six observer departures in a matching unit and left them there. Read in the appearance unit an appearance prediction would actually be judged by, they are not the same six numbers scaled by a constant — the model amplifies the smallest by 1.58 and the largest by 1.03, so the range between the largest and smallest term narrows from 3.34 to 2.17. The published ranking survives in the mean and changes on fourteen of the forty-two surfaces it was averaged over.

brain · Appearance
What a stated lightness pins down, and where. A lightness quoted to 0.05 of a unit, inverted, and the luminance it fixes. Read as a fraction of the colour's own luminance the requirement is 0.90 per cent at J 10 and 0.097 at J 95, a factor of 9.3. Read in absolute luminance it is the other way round, by a factor of 6.5. Both readings are true and they answer different questions.

A stated lightness is two requirements

A specification quotes an appearance to a stated precision — a lightness to a tenth of a unit, say — and the same precision everywhere. Inverted, a twentieth of a unit of lightness fixes the luminance to nine tenths of one per cent at the bottom of the scale and to a tenth of one per cent at the top, a factor of 9.3. Read in absolute luminance it is the other way round by a factor of 6.5, and both readings are true.

brain · Appearance

Named alongside it

The objects these essays reach for when they reach for this one.

Colour differenceTest setDeclared inputAuditElasticitySpecificationChromatic adaptationMeasurement errorChromaLightnessStructural choiceUncertainty

All concepts