The collection

Every essay — page 15

Essays 337 to 360 of 436, in the same order.

The diagnostic after the trial

The test for whether a balanced-assignment walk can reach the whole admissible set is run on a covariate function, before any outcome exists. Run instead on the difference in arm means it is the same test — under the sharp null the outcome is a fixed column — and it is about the statistic the p-value is actually built from. Two findings: an outcome is a probe nobody chose, and on a set that is genuinely split three in ten of them see nothing at all; and the published two-sided p-value is exactly right on half a reference distribution, while a one-sided one is not.

The other half of the dependence

What a balancing rule can remove of an interaction was measured across six marginals with one joint law of the ranks held fixed. Change the copula instead and the two exact zeros come apart for two different reasons. A median split's zero is arithmetic — a centred median split squares to a quarter identically — and holds under every copula there is. A mean's zero needs the copula to be symmetric under reflection as well as the covariate to be symmetric: it is exactly nothing under a Gaussian, t or Frank copula and 7.71% under a Clayton, at the same rank correlation and with a normal covariate throughout.

A block length chosen from the data

Two block windows were compared at each window's own best block length, which is the argmin of a quantity that needs the truth. Re-run on rules a practitioner could actually run, the ordering reverses: the taper wins at the best available length and at one estimated from the sample's own persistence, and the rectangle wins at a length written into a protocol and at the rule of thumb, at every sample size measured. And the whole argument is a third of the size of what estimating the length costs.

A charge for a covariance's own dimension

The two charges that pick a band's width were derived for a regression coefficient, and a band's numbers are neither free nor entered the same way. Measured as an optimism against a second independent sample, a Bartlett band costs 0.374 of a log-likelihood unit a lag where a criterion charges one. The window's own weights are the mechanism — scale them and the charge scales with them, to within two per cent across three windows — and they are not the arithmetic: the level is three quarters of the sum, a different shape at the same sum costs more, and the charge rises with the law's persistence. Levying the right one moves every width in the table by a factor of four and the error by half a per cent.

The rate and the size of a disagreement

A sweep of what it costs to let every candidate choose its own tuning parameter reported a product and called it a cost. Separated — and the decomposition is exact, because a draw on which nothing disagreed carries exactly zero — the rate rises by half across the list and what a disagreement is worth does not move at all. A third quantity explains why: a disagreement costs something only when it changes which candidate the table selects, that happens on about an eighth of draws, and the eighth does not depend on the list.

Two searches over different features

A break search and a window search on one sample manufacture less likelihood together than separately, and both read the same residual series. Put four more pairs beside them on a scale whose zero is two searches over independent columns and whose one is a search that contains the other, and the shortfall is a property of the pair: 0.00 for the control, 0.76 for the pair that reads one series twice, exactly 1 for containment — and −0.31 for a break paired with an independent column, where charging the two separately under-charges rather than over-charging.

A probe chosen rather than picked

A diagnostic that reports on what a balancing rule was not handed is run through a column that is 92% inside the span the rule balanced — because orthogonality in the population is not orthogonality on fourteen units. Projecting the probe off that span costs one least-squares fit and triples the separation it carries; choosing the direction from the design's own leverage is worth another factor of four over a random one; and the projection-pursuit direction the argument invites is worse than random. On a short chain the raw probe misses 44% of the sets that are split and the projected one misses 12%.

Both halves of the dependence at once

One field varies the marginal with the copula held Gaussian and another varies the copula with the marginal held normal, and the natural guess is that the two leaks compound. They do not add in either direction: eleven of twenty cells cancel and nine compound, a mildly skewed covariate under a lower-tail copula leaks 0.002% where adding the two gives 16.6%, and a heavy-tailed symmetric covariate that leaks exactly nothing on its own doubles what an asymmetric copula leaks. A median split's zero survives all thirty combinations at under 10⁻¹⁶.

FieldsThreadsSeriesConceptsFigure librarySearch