The second statistic was the first
Worth reading first: The sequence has a memory · The survey this site cannot do.
This collection has been assembling a specification for an experiment nobody has run: what to measure on a real shoot, how much of it, and what each measurement would settle. The specification is now long enough to be worth auditing, and one of its lines has just failed.
The line was a second statistic. It promised that a divergence sequence long enough to read a comb from is long enough to read a slow wander from as well, and that the two disagree about whether a plant’s errors are inherited between touching organs. It was the cheapest thing in the specification, because it required no additional measurements at all.
It is not a second statistic. Both readings are the autocorrelation function of the same disturbance, one evaluated at a lag and one at a block size, and the identity between them holds across every disturbance this collection has built.
What the specification said, line by line
Worth setting out before amending, because most of it is unaffected.
The counts. Record the parastichy pair at a stated radius or height. This is the field everybody reports and the qualifier is the field nobody does.
The radius. A count without the place it was taken is a statement about an annulus with the annulus missing, and this collection’s own measurements put four distinct pairs in one head.
The symmetry and the hand. Two fields that separate a genuine whorled pattern from an ordinary lattice whose families share a factor, and that separate a lattice from its mirror. Neither is expensive; both are absent from the literature.
The sequence. Divergence angles organ by organ, about nine hundred of them at half a degree of precision, which is what it takes to read a comb.
The intervention. Remove one primordium at a stated offset and record whether the next one moves. Repeated across offsets it returns the larger parastichy number with no protractor anywhere.
The line that fails
The sequence line carried a rider: two statistics, one sequence, and no extra cost. A comb at the contact offsets says the errors are shared between touching organs; a slow wander in the block means says they are also passed on and on.
The wander is real and it is in the disturbance. A divergence is a difference of two organs’ errors, and differencing removes low-frequency power, so the sequence a plant hands over is the one place the wander cannot appear. The inherited disturbance’s own block means carry forty-nine times an independent stream’s variance and the divergences it produces carry 0.83.
So the specification loses a discriminating statistic, and the sequence’s price does not fall: it was nine hundred organs for the comb and it is nine hundred organs for the comb.
What replaces it, and it is cheaper
An inherited disturbance does leave something in the block-mean statistic, and it is not a wander. It leaves a hole: at block sizes equal to the offsets it couples at, the statistic falls well below one — 0.22 at eight and 0.21 at thirteen against a null of one, and 0.06 and 0.12 on stems a placement rule grew.
That is cheaper than the comb it comes from, and the difference is worth spelling out for somebody who would have to compute it. A comb is a set of scores across many lags, judged against a sampling band, and it needs the parastichy pair either known or recovered from the sequence itself. The hole is a single ratio at a stated block size, and the block sizes to use are the two parastichy numbers, which a count of the shoot supplies before the sequence is even taken.
The null is not one, and that has to go into the specification too. A stem placed by a rule with independent errors still digs a shallow hole at its own contact offsets — 0.20 at a block of eight — because the rule’s placements are correlated where its neighbours are. A test calibrated against a null of one would have an inflated false-positive rate, and the calibration is available.
And a question the specification could not previously ask
The withdrawal comes with an addition, and the addition is not a consolation prize.
A wander in a divergence sequence says the disturbance has a memory in time: that the conditions under which organ five hundred was placed resemble those of organ four hundred more than those of organ one. That is a question about a shoot’s environment and its own development, and this collection had no observable for it at all.
It is also readable quantitatively rather than as a yes or no. For a disturbance that remembers the last error with coefficient ρ, the block-mean statistic climbs to one over one minus ρ and stops — measured at 1.38, 1.90, 3.15, 9.51 and 30.94 against closed-form values of 1.43, 2.00, 3.33, 10.00 and 33.33. So the height of the plateau names the memory’s length, and a sequence long enough to see the plateau is a sequence that measures it.
Why “free” was the word that went wrong
The rider’s actual failure was not in the statistics. It was in the word “free”, and the way that word is used about measurements is worth taking apart because it recurs.
A statistic computed from data already collected is free in one sense: no additional specimen, no additional protractor work, no additional season. It is not free in the sense that matters for an argument, which is whether it carries information the other statistic does not. Those two senses of free were run together in a single sentence, and the sentence read as though the second implied the first.
They come apart in both directions, which is why the confusion is not simply carelessness. A statistic can cost nothing extra and add a great deal — the symmetry of a point set is read off a photograph already taken and settles a question the counts cannot touch. And a statistic can cost nothing extra and add nothing at all, which is this case.
The way to tell them apart is not intuition about the statistics. It is to compute both on the same simulated data and look at whether one predicts the other. Here that took a few minutes and produced an exact identity; had it produced a scatter plot with structure in it, the rider would have stood.
The specification as it now stands
Amended in three places and unchanged everywhere else.
The sequence line loses the wander as a test of inheritance and keeps the comb. It gains the hole at the contact offsets, with a stated null of about 0.2 rather than 1, and it gains the plateau as a measurement of environmental drift.
The intervention line gains the two-organ variant, which reaches coarse arrangements a single removal cannot disturb, and gains the whorled variant, where removing either organ of one node must give the same answer — a control internal to a single shoot.
And a new line: report the scatter and the block-mean curve together. A small organ-to-organ scatter reads as a well-behaved shoot and can equally be a rule correcting hard against a large slow disturbance. The two numbers separate those, and neither does it alone.
What the specification is for
It is worth restating, because a specification that keeps being amended can start to look like an end in itself.
Every frequency, every threshold and every angle in this collection is about the geometry or about a model, and says so. The census of divergences is a census of what the arithmetic hands over, not of what grows in a field. The rate at which a pattern climbs the ladder is a property of a rule. The boundary at which a lattice stops being a lattice is a boundary in a simulation. None of that is a defect — the whole discipline here is to state a claim and give it a test it could fail — but it means the collection has exactly one way of being wrong that it cannot detect from the inside, which is that the rule is not what a plant does.
A specification is the list of things a botanist could record that would put a number from a field beside a number from a model. It is not a wish list: every line on it is a field that is cheap to record, absent from the literature, and attached to a specific claim that would be settled by it.
That is why an amendment is worth an essay. Removing a line means one claim has lost its route to a measurement. Adding one means a claim has gained one. And this round does both, which is the ordinary shape of finding out that an instrument measures something slightly different from what it was said to.
The list is now five fields and two interventions, none of them requiring equipment a plant physiology laboratory does not have, and the whole of it would fit on a page. Its cost is dominated by the sequence, which is a season of patient measurement on one shoot; everything else is a photograph and a count.
How an observable gets mispriced
The general lesson is short and this collection is going to keep needing it.
An observable is a function of what can be measured. Between a hypothesis about a plant and a number a botanist can write down there is always at least one operator, and every operator here destroys something. Positions are differenced to get divergences. Divergences are wrapped modulo a turn, and modulo a fraction of a turn on a whorled shoot. Cell areas are read against a background. Counts are taken in an annulus.
The wander was proposed by measuring it on the hidden side of a difference operator and asserting it on the visible side. Nothing about the model was wrong and nothing about the measurement was wrong; the sentence between them did not survive the operator, and nobody had computed the statistic on the side a plant supplies.
The protection costs minutes: compute the proposed statistic on the visible side before writing the sentence. That is now a rule here, and it is the third time a version of it would have saved a claim.
What this does not say
It does not say the specification is weaker overall. It loses one test of one hypothesis and gains a cheaper test of the same hypothesis plus a new test of a different one. The intervention half has gained two variants in the same round.
It does not say the comb is in doubt. The comb is a separate statistic, measured separately, and it is untouched — including this collection’s earlier concession that it is evidence of a re-transmitted disturbance rather than of a placement rule.
It does not say the identity makes the block-mean statistic useless. A quantity that equals the autocorrelation at one lag is still worth computing if it is easier to compute or easier to interpret, and at large block sizes it is both. What it is not is independent information.
It does not say the amendment is the last one. The specification has been amended in every round it has existed, and the amendments have so far been about what a measurement can settle rather than about what to measure. That is the shape of a specification that is being tested rather than written.
And it does not say a real sequence will behave like these. Everything here is measured on stems grown by a rule and on lattices with stated errors added. What the specification does is say what to record so that a real sequence could be compared with them.
The three amendments in one place
For anybody reading the specification rather than this essay, the changes are:
Delete the rider on the sequence line that promises a slow wander as a second test of inheritance. It is not a second test and the sequence cannot show it.
Add to the sequence line: the block-mean statistic at block sizes equal to the two parastichy numbers, with a null of about 0.2 rather than 1, as a cheap confirmation of the comb. And, separately, the shape of the same statistic across block sizes as a measurement of environmental drift, with the plateau naming the memory’s length.
Add to the reporting line: the per-organ scatter and the block-mean curve are to be given together, because a small scatter is ambiguous between a quiet environment and hard correction against a loud one.
The check that would refuse it
The claims in this essay are about what a measurement would cost and what it would settle, and two of them are checkable arithmetic.
The identity between the block-mean statistic and the autocorrelation has to hold at every disturbance shape and every block size, within the sampling error of the noisier estimator. If it failed anywhere the two statistics would be independent after all and the line would not need amending.
The hole at the contact offsets has to be well below one at those offsets and near one away from them. Both halves are required: a statistic that was below one everywhere would be a statistic with a scale error rather than a signature.
And the null has to be measured rather than assumed. The shallow hole a rule-grown stem digs under independent noise is what a real test would be read against, and asserting that it is not one is what keeps the specification honest about its own false-positive rate.
Shares its objects with
Essays that name at least two of the same things, and that neither author linked.
- The control a survey would need — both name autocorrelation, discrimination, evidence, falsifiability, honest limits, measurement, measurement error, sample size, specimen, survey
- The survey loses its second outcome — both name autocorrelation, discrimination, evidence, falsifiability, honest limits, measurement, measurement error, sample size, specimen, survey
- The test a plant could settle — both name autocorrelation, discrimination, evidence, falsifiability, measurement, measurement error, sample size, specimen, summary statistic, survey
- The experiment this site can specify — both name discrimination, evidence, falsifiability, honest limits, measurement, measurement error, sample size, specimen, survey
- What a quiet plant is worth — both name autocorrelation, discrimination, evidence, honest limits, measurement, measurement error, sample size, specimen, survey
- What a refusal does not say — both name autocorrelation, discrimination, evidence, falsifiability, honest limits, measurement, sample size, specimen, survey
Named objects
A flat tag is an object no other essay names yet.
AutocorrelationClaim testingCounting radiusDiscriminationEvidenceFalsifiabilityHonest limitsMeasurementMeasurement errorNegative resultSample sizeSpecimenSummary statisticSurveyUntested claim