Test set — where it appears
Named by 46 essays across 9 fields — each of them below, with the objects they name alongside it.
A mean has a set under it
Every adaptation number this collection publishes is an average over a hundred and twenty-five surfaces that were written down once, in one file, with no argument for how many there should be or how saturated. The average runs from exactly zero to twice itself across them, and the set has never been varied.
A lattice is a quadrature rule
Walking a set of test surfaces more finely does not converge on a better answer, because refining a lattice under a constraint changes which corners of the region get sampled and not only how densely. The lattice used here turns out to be a two per cent biased estimate of the integral it stands for.
The error on a gap is not the errors at its ends
Comparing two rows of a table by looking at whether their error bars overlap is the wrong comparison, and here it is wrong by a factor of up to 3.4. The same 125 surfaces score both rows, so the difference between them is quieter than either — and how much quieter is a measurement of how alike the two rows are.
Four steps the test set cannot order
The adaptation census prints fourteen numbers to four figures and its ranking is asked to say which lamps adaptation handles worst. Nine of its thirteen steps are established beyond any doubt the test set can raise; the other four are not, and three of them are consecutive — a tungsten lamp, a halogen lamp and a white LED are simply not ordered.
The instrument named the pair that moved
A standard error over a test set flagged four steps of the adaptation census as unresolved. Rebuilding the set three different ways reversed exactly one pair, and it was one of the four. Rebuilding it to a different rule reversed a pair the error separated by nearly nine standard errors — which is not a failure of the instrument but a statement of what it is about.
A theorem about a family
A change of light acts on the test surfaces used here as an exact 3×3 matrix with no residual whatsoever, and the whole adaptation argument is built on that being exact. It is exact because the surfaces span exactly three dimensions, and they span exactly three dimensions because three basis functions were written down.
A fourth dimension has a shape
How much a fourth reflectance dimension costs spans a factor of nine across four equally plausible shapes at one amplitude, and the expensive ones are not the shapes a variance figure would identify. The band that hurts is set by the illuminant rather than by the eye, which is why a triphosphor tube and a tungsten lamp disagree about it.
The row a fourth dimension improves
Giving the test surfaces one more degree of freedom makes almost every change of light harder for an adapted observer — but not all of them, and which one it helps depends entirely on what the extra dimension looks like. A triphosphor tube is improved by one shape and hurt more than anything else in the census by another.
Saturation is nearly everything
The set of test surfaces has three numbers describing it, and only one of them matters. How saturated the surfaces are carries an elasticity of about 0.7 on every result computed over them; how bright they are carries 0.10. A test chart's chroma range decides its answer and its lightness range does not.
The input nobody declared
An audit that swept every declared width in this collection found the largest elasticity anywhere to be about a half. The most elastic input turns out to be one that was never declared, never quoted with a range and never varied — and being undeclared is exactly why it escaped the audit that was looking for it.
The surfaces that answer nothing
Five of the hundred and twenty-five test surfaces contribute exactly zero to every number the adaptation census reports — not approximately, exactly — and the reason is the one fact about von Kries adaptation that makes it worth having at all. Counting the set by how much it contributes gives about a hundred members rather than a hundred and twenty-five.
A mean is not a worst case
Every adaptation number this collection publishes is an average over objects, and the reader asking whether adaptation will fail them is asking about the object it fails on. That object costs between 1.9 and 4.0 times the published figure, and how uneven a change of light is across objects turns out to be a property of the change rather than a constant.
An extremum is still not a sample
Two rounds ago three measurements turned up that took a maximum over a sample of a set and were short by up to a factor of two. The same error was live in a fourth place the whole time, on the set of surfaces every adaptation number is averaged over, and it is short by up to a third.
Every worst surface sits on a declaration
Bounding the wall in a painted room produced a real worst case — the residual turns over at a band six nanometres wide because a narrower band returns too little light. Bounding the surfaces the residual is averaged over produces nothing of the kind, because all fourteen answers sit exactly on two numbers somebody typed and the one constraint that comes from the world never binds at all.
A chart decides what a camera scores
A camera profile's reported error changes by a factor of five when the test chart's saturation changes, with the camera and its matrix untouched. The elasticity is 0.67 — the same figure, to two per cent, that an entirely unrelated measurement over an entirely unrelated set of surfaces gives.
A budget drawn through one hue
The three-stage error budget this collection publishes for a colour-management chain is computed over twenty-four colours of a single hue at a single lightness. The quantity that actually varies with hue — how many distinguishable colours a rendering intent destroys — runs from nothing at all to more than a third, and the hue the budget uses is near the bottom of that range.
What the audit still cannot reach
Two rounds have now swept every declared width in this collection and one of its structural choices. Three structural choices remain, none of them has a multiplier to sweep, and the reason each resists is different — which makes the list a description of where this kind of audit ends rather than a queue of work.
A choice with no magnitude
An audit can multiply a width by 1.25 and report an elasticity. It cannot multiply CIEDE2000 by anything. Auditing a structural choice needs a different instrument, and building one shows that six published numbers in this collection each carry a factor of about two of unit-choice — after the change of scale has been taken out.
The census in six units
Recomputing every change of light in the adaptation census under six colour-difference formulae, with the scale factor divided out, leaves a table whose levels move by up to a factor of three point seven. The rows that move most are the mild ones, which is the opposite of what a reader would guess and is a property of where each formula was fitted.
The weighting is the disagreement
Five colour-difference formulae, three decades and two committees, and the single property that predicts which of them agree is whether a chroma difference gets divided by the chroma it was measured at. It sorts the menu exactly, it cuts across the distinction between a matching difference and an appearance one, and it halves the census's largest sensitivity.
A dial through a discrete menu
ΔE*94 is ΔE*ab with two weighting constants in it, and at zero those constants make every weight exactly one — so the two ends of the oldest disagreement in colour difference are joined by a line rather than separated by a choice. Walking it gives a derivative where a menu gives only a spread, and the derivative says the published weighting is on the far side of the interesting part.
Two instruments and one ranking
A sampling error over a hundred and twenty-five surfaces and a change of colour-difference formula share no arithmetic at all, and they were asked the same question of the same table. Every adjacency the whole menu reverses had already been flagged as unresolved. And one the test set settles at nine standard errors is reversed by four of the five formulae, which is what makes them two instruments rather than one.
The disagreement is at the near end
Every colour-difference formula on the menu was fitted to threshold data, so the expectation is that they agree about pairs an observer can only just tell apart and diverge on large differences. They do the opposite. Proportionally the disagreement is largest at the near end, by a factor of six for the appearance unit, and the cause is an exponent of 0.63.
The observers differ by a unit's worth
The gap between the 1931 and 1964 standard observers is the one quantity in this collection's audit with no published number under it — nothing here reports it as a single figure over a stated set. It also has the second-largest dependence on which colour-difference formula is used, running from 1.60 to 3.55 across the menu.
The coincidence was a mechanism
Two sensitivities from two libraries with no shared code came out two per cent apart, and the claim made about them was that they share a mechanism rather than a number. That claim has a colour-difference formula inside it, so it can be tested by changing the formula — and both curves move together across the whole menu, from 0.5 to 1.15.
The objective nobody chose
Every camera profile here, and as far as can be told everywhere, is a linear least-squares solve in tristimulus space. That is an objective and it is on nobody's menu — it weights an error by how bright the patch is. Refitting the same matrix to minimise a real colour-difference formula improves the fit in all six, and moves the matrix, which means different pixels rather than a different report.
Three numbers the scene supplies
An adaptation model's parameters are not all the same kind of thing. Some are numbers an observer must estimate from the room it is standing in; others could have been settled once by evolution. Counting them separately turns the diagonal gain from a crude approximation into the only model of the set that gets a large answer from information the observer can actually have.
A partial correction is worth its fraction
Between a diagonal gain and the exact matrix there is a line, and a bounded observer's natural hope is that the first part of it is worth a disproportionate share. It is not. On all fourteen changes of light, at every setting, the share of the residual removed matches the share of the correction applied to within 2.2 percentage points — which closes the last way the gap could have been cheap.
Three choices reached
Two rounds ago this collection named three things it rested on and could not audit — a unit, a diagonal and a template. All three are now reached, and the interesting part is not the three answers but that four of the round's own predictions were refused by the arithmetic and one of its measurements was wrong in a way only a cross-check caught.
A departure is not a unit
The previous round audited six published quantities by swapping the unit they were quoted in — a function applied at the end of each computation. Nothing of that shape works here. A departure changes the object at the start, so every stage in between has to accept it, and only two of the same six quantities can take one at all. The fourth cannot even be expressed in the interface.
A neutral has no grid
A perfectly flat reflectance computes to exactly the same colour on every wavelength grid, through every slit, at every origin, and for every observer — not nearly, but to the last bit of a floating-point number. The condition is an identity rather than a limit, and what makes it one is the white point.
Where the grid starts
Holding the step at five nanometres and sliding the grid's origin through one cell moves a fluorescent tube's computed colour by 3.18 ΔE₀₀ and a laser projector's by 35.0. Refining the step does not fix it and averaging over origins hides it. It is the one tabulation fault with no smooth error to cancel against.
Two ends and one is empty
Extending this collection's wavelength range down to 300 nanometres moves a red pigment under daylight by 0.502 ΔE₀₀. Extending it up to 830 moves the same colour by 0.00015. A fifth of a thermal source's power lies outside the range and almost none of its colour does, and confusing those two shares is how a range gets argued about instead of measured.
Which end to buy
Refining a five-nanometre grid to one buys a daylight calculation 0.05 ΔE₀₀ and widening its range buys 0.54. Under a fluorescent tube the same two purchases are worth 0.83 and 0.0001. The ranking reverses completely, and what decides it is one length compared against one other length.
The normaliser carries the error too
A five-nanometre sum gets a red pigment's tristimulus value wrong by two hundredths of a per cent and its colour wrong by six hundredths of a unit. Those two numbers are not the same size because the grid appears twice in a colour — once in the sample and once in the white — and the two errors are largely the same error.
A grid is not a resolution
Ten essays into this collection there is one sentence about wavelength sampling, and it is that five nanometres is enough. Ten measurements later there are three decisions, three mechanisms, three repairs and two rankings, and the word resolution names none of them.
The lens is worst under tungsten
An ageing lens costs 4.38 ΔE₀₀ under a tungsten lamp and 2.20 under a fluorescent tube. The pigment peaks cost 2.84 under a three-emitter LED and 1.87 under the same tube. Six departures across six lights do not form a single ordering, which is what a pairing looks like and what a dominant term would not.
The ranking is not stable
On a red pigment under daylight the six observer departures run from 2.38 down to 1.20 ΔE₀₀. Over forty-two surfaces two of them change places, the top two separate, and every one spans between a factor of ten and a factor of thirty-five. A chart of six bars is a chart of one example.
The macular is a band, not a filter
The lens absorbs everything below 500 nanometres with a long tail; the macular pigment absorbs forty nanometres either side of 460 and nothing else. The two have similar sizes and completely different distributions, and the reason is that one is broad and one is narrow — which is what decides whether a sample is affected at all.
A tolerance with an observer in it
A delivery tolerance is written in ΔE₀₀ against the 1931 observer, and six departures of that observer combine to about three of the same units on an ordinary saturated sample. A one-unit tolerance is being asked to contain a three-unit uncertainty that nothing in its budget mentions.
The census under another observer
This collection's largest computed result is an adaptation census — fourteen changes of light judged over a hundred and twenty-five constructed surfaces. Every number in it was computed through one observer, and the observer's own departures are between one and two and a half units on the same surfaces, which is the size of the effects the census reports.
The grid under the census
The adaptation census is computed on eighty-one wavelengths from 380 to 780 nanometres. Under the daylight and blackbody sources it uses, the range is worth about half a colour difference on ordinary surfaces and the step about a twentieth — so the census carries a tabulation term as well as an observer one, and they are not the same size.
The conditions are the result
The round measured about twenty departures and established ten conditions. The departures are numbers that depend on a sample, a light and a construction; the conditions are exact, they hold under every substitution tried, and they are what a reader can act on. A size is a measurement and a condition is a mechanism.
What this round could not reach
Three audits, about twenty departures and ten conditions, and a longer list of things that were named and not measured. Every item on it is specific, most of them are an afternoon's work, and the reason none was done is the same in every case — the round ran out of round.
Two yellow filters cancel on a slope
An older lens and a denser macular pigment both take blue out of the light, and read before adaptation they move a colour in nearly the same direction, eight degrees apart. Once each eye has adapted to its own white they point a median 156 degrees apart on smooth reflectances and together cost less than the lens alone. On surfaces with a narrow absorption band they still sit 26 degrees apart and add. What decides it is the width of the surface's own features.
The average surface does not look average
Over a population of observers the appearance model is so nearly linear that the mean of its answers is its answer for the mean, to 1.4 per cent of the spread. Over the surfaces in a scene it is not. On 240 smooth reflectances the gap is 21 per cent of the spread, and the mean surface looks 4.2 units lighter than the surfaces look on average. The grey that matches the average light is a 47 per cent reflectance; the grey that matches the average look is a 43 per cent one.
Named alongside it
The objects these essays reach for when they reach for this one.
Chromatic adaptationColour differenceResidualReflectanceSensitivitySpecificationCalibrationMeanAuditElasticitySamplingStandard observer