The basins have widths as well as depths
Assumes Where the valley breaks in two.
The height of the pass between two basins measured the saddle directly, found it did not explain the exponent it was supposed to explain, and recorded a shortfall in one sentence: the shape of the aspect score surface has widths in it as well as heights, and the sweep that finds the pass was computing the widths on its way past and throwing them away.
This rung reads them out. Three things come back, and only the first was expected.
What a width is here, and how it is got
The objective is a score over the three angles an aspect is: where the rotated pole goes, in longitude and latitude, and how far the page is then turned.
Near a minimum a smooth function is a quadratic form, so its second derivative at the optimum carries the whole local shape. The width along a principal direction is the displacement at which the score rises to twice the optimum’s — √(2 f₀/λ) for the eigenvalue λ in that direction — and it is in degrees of rotation, because that is what the parameters are.
Nineteen evaluations of the objective give the three-by-three second derivative, and its eigenvalues give three widths.
That is a great deal cheaper than the alternative, and the alternative was tried first. Reading the width off the sublevel set — sweeping the score in ascending order, unioning cells as they arrive and taking the cube root of the component’s volume — is the natural thing to do with machinery that already exists. On the 24 × 13 × 12 grid the whole ladder runs on, the near-optimal component is a single cell at every threshold up to half the fracture, so the cube root is a measurement of the grid spacing. Refining the grid enough to fix that is thirty seconds a region, six regions a sweep, and the answer would still be quantised.
The first finding: it is not round
A search that steps the same distance in all three of its parameters is therefore stepping across the basin in one direction and along it in another, by up to a factor of eleven. That is not a small remark about tuning. It is the reason the parameters a search reports are not reproducible while the map is: along the widest direction the objective barely changes, so where the walk stops is decided by its step schedule rather than by the surface, and that direction is eleven times longer than the direction the surface actually constrains.
The three widths are the quantitative version of what that rung established qualitatively, and they say which direction is the loose one.
The second finding, and it is the wrong way round
The plan for this rung assumed the basin narrows as the region grows. A larger region sees more of the projection’s own variation, so it should care more about where the aspect is put, so the near-optimal set should shrink.
It does the opposite.
The mechanism is in the definition and is not a defect of it. A width is measured where the score doubles, and what a larger region does is raise the whole surface: the best achievable distortion over a forty-degree region is far worse than over a six-degree one. The floor rises faster than the curvature does, so twice the floor is further away.
That is worth stating as a general caution rather than as a local fact. A relative threshold on an objective whose floor is moving measures the floor as much as the shape. Any near-optimal set defined as “within so many per cent of the best” inherits it, and this collection has now met the same trap in the tolerance that decides the verdict, where a fixed tolerance was found to be measuring the arithmetic’s noise floor rather than the map.
The threshold the widths are being compared against
The quantity the widths have to explain is the one where the valley breaks in two measured, and it is worth having it on the page rather than only in a fitted exponent.
Two numbers now exist for the same six regions, measured by routes that share nothing: a threshold from a union-find sweep over a grid of scores, and a width from nineteen evaluations of the objective at a single point. Neither knows anything about the other, which is what makes comparing them a test rather than a restatement.
The third finding: the prediction overshoots
The reason the shortfall was recorded at all is that the width was the suspect for a missing exponent. The argument is two lines.
The set below a relative threshold t reaches a distance w₀√t from the optimum. Two near-optimal pieces become separate answers when their sets touch, so if the distance between them were fixed, w₀√t* would be a constant and
The width exponent is +1.086, so the fracture exponent should be −2.172. Over the same regions it measures −1.479.
The width does not close the gap. It opens one on the other side, and by more than the original: the argument was short by 0.48 and the width overshoots by 0.69.
That is a real result and it is worth being plain about why it is not a disappointment. The previous rung’s suspect has been measured, and measuring it has ruled it out as the explanation while establishing that it is a large part of the mechanism — the prediction moved from 1 to 2.17 and the truth is between them. What is left over is now a smaller and much better-specified question than “why is the exponent not one”.
Why it overshoots, which the same sweep answers
The quadratic argument has two quantities in it and only one of them has been measured. The other is the distance between the two pieces, which the argument assumes is fixed by the parameterisation.
There are not two pieces.
At the threshold the near-optimal set of the Robinson aspect over Japan is in twelve pieces. Where the valley breaks in two already knew this — it reports one connected sheet at a loose threshold and fourteen basins at a tight one — and the two-basin language it used, which this rung inherited, is a description of the first thing that happens rather than of the thing that is being measured.
So the fracture threshold is not the merging of two basins at a fixed separation. It is a percolation: the threshold at which a set of many small components first connects into one. Percolation thresholds have their own scaling, set by the dimension and by the correlation structure of the field, and there is no reason for it to be the two-body exponent −2.
That reframing costs nothing already established. Everything the ladder has measured — the threshold, its region dependence, the pass height, the widths — is unchanged. What changes is which model those numbers should be fed to.
What the separation actually does
It is worth showing why the two-piece model could not have been rescued by measuring its separation instead of assuming it.
Taken just under the fracture, the distance between the two lowest pieces comes back at 151°, 151°, 122°, 66.5° and 216° across the sweep, with no order in it. Fed into the same quadratic model those give predicted thresholds of 47.5, 15.7, 0.68 and 5.9 against measured values of 2.13, 1.33, 0.69 and 0.27 — agreement at one region out of four and disagreement by a factor of twenty at another.
The numbers are not noisy because the measurement is poor. They are meaningless because “the two lowest pieces” is not a well-defined pair when there are twelve, and which two the sweep happens to find depends on the grid.
What is left of the two-line argument
It is worth separating what the crumbling destroys from what it leaves standing, because it is less than it looks.
The relation between the width and the threshold survives: the two move in opposite directions, at fits of R² 0.944 and 0.948, and they must, because a wider basin reaches its neighbours at a lower threshold whatever the neighbours are. That is asserted here and it is the part of the quadratic model that does not depend on how many pieces there are.
What does not survive is the factor of two. It comes from the square in w₀√t, which is a statement about a single pair of bodies approaching each other, and a percolation is not that. Twelve components arriving at a common level is a connectivity problem on a graph, and its exponent depends on how the field’s values are correlated between neighbouring cells rather than on the local curvature at one point.
So the honest reading of this rung is: the width is measured, it is a large part of the mechanism, and the exponent it predicts is wrong for a reason that the same sweep can see and that the ladder’s own previous rung had already reported without noticing what it was reporting.
What a search should do with all this
Three consequences, and the first two are immediately usable.
Step anisotropically. The three widths are known, cheaply, from nineteen evaluations at the current best point. A compass walk that scales its step to each width converges in the same number of steps in every direction instead of grinding along the loose one — and the loose one is up to eleven times longer.
Do not report the parameters. With a width of 32° along the loosest direction, two searches that agree about the map to a part in a thousand can report pole positions half a world apart, which is exactly what the reproducibility rung found and could not previously explain in one number.
And treat the threshold as a percolation. A search that wants to enumerate the genuinely distinct near-optimal aspects should not look for the height at which two basins meet; it should look for the height at which the largest component stops growing faster than the number of components falls, which is a different reading of the same sweep.
Three numbers rather than one
What this rung leaves the ladder with is a short list, and it is worth setting out because the next rung has to start from it.
A depth — how good the best aspect is, which is what every search reports.
Three widths — how well determined it is, direction by direction, from nineteen evaluations of the objective at the point the search stopped.
And a piece count — how many genuinely different answers there are at a stated tolerance, which is what the sweep produces and what nobody reads out of it.
The three are independent, they are all cheap, and only the first is ever printed. A search that reported all three would say what it found, how precisely, and whether it was the only one — which is the whole of what a reader of an optimisation result needs and is more than any published aspect search supplies.
Where the model stops
The Hessian is taken at 2.5° in each of three angles, which is coarse against a narrowest width of 5.2° at the smallest region — so the second difference there is averaging over a third of the basin. Halving the step moves the widths by under two per cent on every region tested, which is the reason the coarse step is kept, but the smallest region is the one to distrust.
One of the six regions produces no Hessian at all. At a twenty-degree span the search’s stopping point has a negative curvature in one direction, so it is on a ridge rather than in a bowl, and a local polish before the differencing does not rescue it. The row is dropped rather than repaired, and dropping it is why the fits below use four points rather than six.
The parameterisation is degenerate at a pole latitude of ±90°, where a change in pole longitude is a change in the third angle. None of the optima found here is within twenty-five degrees of that, and a region whose optimum was would need the widths read in a different chart.
The generalisation
Strip out the cartography and what is left is a statement about reading an optimisation landscape.
A basin has a depth, three widths and a shape, and every summary of it as one number throws away the part that decides how a search behaves. The depth says how good the answer is. The widths say how well determined it is, direction by direction. The anisotropy says whether stepping uniformly is sensible. And the number of components at a threshold says whether “the basin” is a thing at all.
The site’s own instrument is doing the same work here that it does on projections: the second derivative of an objective is the same kind of object as the second derivative of a map, and the flexion ladder is the essay about what a first-order description leaves out. An optimisation landscape described by its minimum alone is Tissot’s ellipse with the ellipse left off.
Who found it, and when
Reading a basin’s shape off the Hessian is the oldest idea in numerical optimisation and is the whole content of Newton’s method: a step scaled by the inverse second derivative is a step that treats every direction alike. That the eigenvalues of the Hessian are the reciprocal squares of the widths is a restatement, and quoting them as a condition number is standard.
Percolation as the thing that happens when a level set of a random-looking field first connects is younger — the theory dates from the late 1950s — and its arrival here is a consequence of the measurement rather than a hypothesis brought to it. Nobody chose to look for a percolation threshold; the piece count refused the alternative.
The width is the precision the answer should be printed to
There is a small use for these numbers that requires no further work and that the field gets wrong routinely.
A basin’s width is the distance over which the objective does not meaningfully change. So it is exactly the resolution at which the optimum is determined: two parameter values inside one width are two names for one answer, and reporting the difference between them is reporting noise.
That gives a rule for how many digits an optimal aspect should be published with. If the basin is three degrees across in one direction, the optimum in that direction is known to about a degree, and a paper reporting it to four decimal places is asserting a precision the landscape does not contain. The extra digits are a property of where the optimiser happened to stop.
And the widths are not equal, which is the first finding above, so the precision is not the same in every direction. An answer quoted to one degree in the narrow direction and one degree in the wide one is over-precise in one and under-precise in the other, and the honest report gives a different number of digits to each — or, better, gives the widths beside the optimum and lets the reader see the shape of what was found.
This is the cheap half of reporting the map rather than the parameters, and it is available to anybody whose optimiser can be run three more times. The expensive half is establishing that the parameters are identifiable at all; the cheap half is refusing to print digits the objective cannot support.
Where the ladder goes next
The exponent is now bracketed by two arguments rather than approached by one, and neither is right. The next thing to measure is the piece-count curve itself — how the number of components rises and falls through the threshold — because that curve has an exponent of its own and it is the one the percolation model actually predicts.
What this makes readable
Essays that name this one as a prerequisite.
Named alongside this one
Essays reaching for the same objects. Nobody chose these; they are what the concept index makes visible.
- Fitting the aspect to the region aspect · optimisation
- The cheapest map that meets its areas optimisation · shortfall
- The landscape the search walks on aspect · optimisation
- The maps with no family are simply better aspect · optimisation
- The pooled score abandons a region aspect · optimisation
- The rule of thumb, scored aspect · optimisation
What links here
Every essay whose body links to this one.
The objects this essay names
Each one links to every other essay that touches it.
AnisotropyAspectAspect searchBasinExponentFracture thresholdHessianLevel setOptimisationOptimisation landscapePercolationShortfall