Least favourable configuration — where it appears
Named by 2 essays across 2 fields — each of them below, with the objects they name alongside it.
The models that were never in the running
A reference distribution for a set has to assume something about every candidate in it. Assuming that all of them are as good as the benchmark is what makes the reality check honest, and it is what sixteen hopeless candidates use to destroy it.
The corner the test is calibrated at
"No candidate is better than the benchmark" is not a null but a face of a region, and a reality check is calibrated at one corner of it. Fill the table with candidates that are hopeless rather than equal and the test finds a genuine improvement 0.0% of the time.
Named alongside it
The objects these essays reach for when they reach for this one.
Benchmark forecastComposite nullError rateMonte CarloReality checkStatistical powerBootstrapConservative testData snoopingFamilywise error rateLoss differentialMultiple comparisons