The coin a clock supplies
Worth reading first: More data is not monotonically better.
The coin that is already there found that the order in which a sequence of trials’ successes arrived can play the part of the randomised interval’s coin: given the count, every arrangement of the successes is equally likely whatever the proportion, so the arrangement’s rank is a uniform that nobody drew and every analyst can recompute. Between proportions of 0.2 and 0.8 it delivered the drawn coin’s exact 95% to within a tenth of a point. At the ends it failed, because a count of zero has one arrangement and a count of one has only as many as there are trials, and a coin with one face or twenty cannot imitate a uniform.
The essay ended on a process observed continuously rather than in trials. Events that arrive at random over a window — failures over a year of service, cases over a season, decays over a counting interval — come with their times, and the times have the property the arrangement had, in a stronger form. Whether that coin removes the hole at the rare end, and what a count of zero, which has no times at all, forces the interval to give up, are both exact sums. The answer to the second turns out to decide the first.
Times that say nothing about the rate
Suppose events arrive as a Poisson process over a window of unit length, at a rate whose expected count over the window is . The count is Poisson with mean , and it is what every interval for uses. Given , the arrival times are distributed as independent uniforms on the window, sorted, and that distribution has no in it. The times are ancillary given the count in exactly the sense the arrangement of trials was, but where the arrangement took one of values, the times take a continuum.
That makes any function of the times that is uniform given an exact coin. The simplest is built from the last arrival. If is the latest of uniform times, , so
is uniform on the unit interval given , at every from one upwards. The randomised interval then contains every for which lies between 2.5% and 97.5%, exactly as with a drawn coin, and two analysts reading the same times get the same interval.
At a count of one the coin is the single event’s position in the window, and it has a continuum of values where the order of twenty trials offered twenty. Every count that has an event in it now has a coin as good as a drawn one. Only a count of zero is left without one, because no events means no times, and the interval at a count of zero has to be a stated number.
Exact everywhere the count has a time
With a continuous coin at every count from one up, the coverage at an expected count is the drawn coin’s 95% plus a correction that comes entirely from the count of zero. A drawn coin at zero would have covered with some probability; the stated interval at zero either covers or it does not; the difference, weighted by , is the whole of the departure from 95%.
The hero figure shows what that leaves. Given the exact interval’s limit at zero, , the arrival-time interval covers exactly 97.5% at every expected count between 0.025 and , and exactly 95% at every expected count beyond it. Its worst coverage anywhere is 95.00%. There is no oscillation and no hole: the curve is two flat segments joined by a step.
The step is easy to account for. Below a count of zero always covers, where a drawn coin would have covered with probability ; the excess is , which is exactly 2.5 points whatever is. Above the stated interval at zero no longer covers, and neither would a drawn coin, so the excess is zero. The only thing on the curve that is not exactly 95% is the stated interval at zero being conservative where a coin would have gambled.
Beside it, mid-p’s coverage oscillates as every interval for a count does, and reaches its worst, 91.66%, at an expected count of 3.03. The exact interval never falls below 95.62% on this range and spends most of it above 98%, which is the price a guaranteed minimum always charged for a count. The score interval for a Poisson mean, the count’s analogue of Wilson’s, has the hole a hole no sample size fills found for a proportion: 83.86% at an expected count of 0.176.
The limit a count of zero needs
What the zero count is given is not a detail. Given mid-p’s limit, , the same arrival-time coin produces the dashed curve in the hero figure: exact at 97.5% up to an expected count of 3.00, and then a drop to 92.50%, recovering towards 95% as the probability of seeing nothing dies away. The coin is the same in both curves; only the interval at zero changed, and it moved the worst case by two and a half points.
The frontier is in closed form. With a limit shorter than , an expected count just above sees a count of zero with probability ; the stated interval misses, while a drawn coin would have covered with probability , so the coverage there is , which is . Mid-p’s limit, , puts the worst at exactly 92.5%. Lengthening the zero interval raises the worst case steadily until, at , it reaches 95% and stops: any longer limit leaves the minimum where it is and adds coverage only between and the new limit.
So the exact interval’s limit at zero is not one choice among many. It is the shortest limit for which the arrival-time interval never covers below 95%, and it is forced: a reader who wants no hole has to accept as the report on seeing nothing. The score interval’s zero limit, , is longer than it needs to be and buys nothing on the minimum. Mid-p’s is shorter than it can afford.
That is the part a count of zero forces the interval to give up. The arrival-time interval is not exactly 95% everywhere: below an expected count of 3.69 it is exactly 97.5%, conservative by a fixed two and a half points, because the report on seeing nothing has to be long enough to cover the expected counts that often produce nothing. That is a smaller concession than it sounds. The exact interval is conservative by three to five points across most of the same range, and it is conservative above 3.69 as well, where the arrival-time interval is not.
The hole was the count of zero’s all along
The same arithmetic applies to the coin read from the order of a fixed number of trials, and it says something about the order of arrival that the earlier essay did not separate. That coin had two weaknesses at the rare end: a count of none had a single face, so its interval was mid-p’s; and a count of one had only faces. Its worst coverage, 92.60% at twenty trials, was attributed to both.
Give the counts of none and of all the exact interval’s one-sided limit, as the arrival-time interval’s zero count was given, and the order’s worst coverage rises from 92.51% to 94.56% at ten trials, from 92.60% to 94.79% at twenty and from 92.72% to 94.93% at fifty. Most of the hole was the single face at zero. What remains — a fifth of a point at twenty trials, seven hundredths at fifty — is the count of one’s faces, which cannot quite imitate a continuum. Read from a clock, the count of one has a continuum of faces and the remainder is zero.
The arrival times are the order’s limit in the plainest sense: as the trials become more numerous and each less likely to succeed, with the expected count held fixed, the positions of the successes among the trials become uniform times in a window, which is the Poisson limit of a binomial count applied to the coin as well as to the count. So the figure is one sequence. The order of ten trials is a coin with ten faces at a count of one, of fifty trials one with fifty, and the clock is the coin with all of them, and in every case the decisive choice is what is reported when nothing happened.
When the rate drifts across the window
The times are ancillary only if the rate is constant across the window. If it drifts — a hazard rising with age, a season building towards its peak — events crowd towards the end, the last arrival tends to be late, and the coin built from it is no longer uniform given the count.
The last-arrival coin tilts the way the order of trials did. At an expected count of five, with the rate rising from zero at the start of the window to twice its average at the end, the interval leaves the average rate below it 2.07% of the time and above it 3.34%, against 2.50% and 2.50% with no drift, and covers 94.59%. At half that drift the split is 2.26% and 2.92%. A late last arrival gives a large coin value, a large coin value moves the interval down, and the interval errs low. A coin built from the first arrival tilts in the same direction by a different amount, since the first arrival is late too when the rate rises.
A clock has more than one coin in it, though, and some are immune. Fold each time about the middle of the window, , and build the coin from the largest: . Under a constant rate the are uniform, so the coin is exact. Under a linearly drifting rate they are still uniform, because a linear intensity adds exactly as much probability at distance on one side of the middle as it removes at the same distance on the other. The folded coin’s misses stay at 2.50% and 2.50% at every drift, exactly, and its coverage does not move.
The folded coin is not immune to everything. A rate that peaks in the middle of the window, or at both ends, moves probability between distances from the middle, and a folded coin would tilt under it as the last-arrival coin tilts under a trend. What the folding buys is protection against the one departure most likely to go unnoticed in a single window, and it buys it at no cost under the null, since every one of these coins is exactly uniform when the rate is constant.
What the interval costs in width
A coin that covers less where the exact interval covers more should give a shorter interval, and it does.
At an expected count of one the arrival-time interval is 90.6% as wide as the exact interval on average; at two, 89.6%; at twelve, 93.8%. Its width falls back towards the exact interval’s only at the smallest expected counts, 95.8% of it at a quarter, where a count of zero is the likeliest outcome and the two report the same interval on it. Mid-p is narrower still at small expected counts, 87.1% of the exact width at one, and the difference is precisely its shorter zero interval — the length that leaves the hole at 92.5%. The score interval is 99.3% of the exact width at one and pays for that width with the deepest hole of the three.
So the comparison the figure draws is the one the frontier implied. Among intervals that never cover below 95%, the arrival-time interval with the exact zero limit gives back a tenth of the exact interval’s width through the middle of the range. Among intervals that are shorter still, every one buys its extra narrowness at the rare end, by reporting less than on seeing nothing.
The objection, and where the clock leaves it
The coin that is already there set out the objection that reproducibility does not answer: two records with the same count have the same likelihood for the rate, and an interval that differs between them depends on something the likelihood ignores. The clock makes that sharper. Two windows with one event each, one early and one late, carry the same evidence about the rate in the likelihood’s sense and get different intervals.
The defence is the same one and no stronger: a 95% interval is a promise about the procedure, and this is a procedure that keeps the promise exactly wherever an event was seen, with a coin anyone can recompute from the record. The clock adds a second commitment the order of trials needed too, and the essay on the order measured what breaking it costs: which function of the times is the coin has to be stated before the times are read. The last arrival, the first, the folded maximum and infinitely many others are all exact coins, and an analyst free to choose among them after looking is a coin re-drawn until it lands well. The ceiling on that failure is the same as for the order — the test that rejects whenever any coin value would — and the remedy is the same: write the coin down in advance.
What the arrival times can be used for
As the coin of the randomised interval for a rate, whenever the times of the events in the window are recorded. Given the exact interval’s limit at a count of zero, the interval covers exactly 95% at every expected count above 3.69 and exactly 97.5% below it, and is about a tenth narrower than the exact interval through the middle of the range.
With the zero count’s limit chosen on purpose. The worst coverage is for any limit below , so mid-p’s limit costs 2.5 points at an expected count of three and the exact limit costs nothing. The same choice lifts the coin read from the order of a fixed number of trials from 92.6% to 94.8% at twenty trials.
With the folded coin where a trend is possible, since it is exact under any linear drift and costs nothing when there is none.
Every coverage here is a finite sum over counts with, at each count, the probability of the coin values that cover — computed exactly, because each coin’s distribution given the count is known in closed form, with and without a drift. The coverage with the exact zero limit is checked never to fall below 95% and to equal 95% to nine decimals past , and the folded coin’s coverage is checked to be unchanged by drift to twelve decimals. The worst coverages for the order of trials are taken over a thousand proportions. Mid-p’s zero limit is refused as harmless at the rare end: the clock coin with it covers 92.50% just past .
Still open: two windows and a ratio
Everything here is one window and one rate. The comparisons people actually make with counts are between two — a rate of adverse events in two arms, cases in two seasons — and the interval for a ratio of two Poisson means is conditional on the total: given that events occurred, the number in the first window is binomial, with a proportion that depends only on the ratio. That turns the two-rate problem back into the proportion problem, with an order of arrival available again, now in the form of which window each event fell in and when.
Whether the arrival times in both windows supply a continuous coin for the conditional binomial — so that the interval for the ratio is exact at every total except zero — and what a total of zero then forces the ratio’s interval to report, is computable exactly on the same terms as the sums here. Whether the conditional test’s familiar conservatism at small totals is entirely the zero total’s, as the hole here was entirely the zero count’s, has not been computed.
Shares its objects with
Essays that name at least two of the same things, and that neither author linked.
- The coin that makes it exact — both name clopper–pearson, coverage, discreteness, mid-p, randomised interval, reproducibility, wilson interval
- An interval that covers and says nothing — both name clopper–pearson, conservative interval, coverage, interval width, wilson interval
- The shortest interval is the one that misses — both name clopper–pearson, conservative interval, coverage, discreteness, interval width
- A ratio that changes between blocks — both name conservative interval, coverage, interval width
- Marginal is not conditional — both name coverage, exchangeability, interval width
- The condition that cannot be dropped — both name conservative interval, coverage, interval width
Named objects
A flat tag is an object no other essay names yet.
Clopper–PearsonConservative intervalCoverageDiscretenessExchangeabilityInterval widthMid-pPoisson limitRandomised intervalReproducibilityWilson interval