v0.3.0
Intervals over repeated items are wider. A bundle sealed by 0.2.1 with replicates above one holds numbers this version does not reproduce from the same rows.
Asking an item twice does not give two independent observations of the system, but the rollup counted one row per item-trial and gave that denominator to Wilson. Over 200 items, a nominal 95% interval held:
| replicates | 0.2.1 | 0.3.0 |
|---|---|---|
| 1 | 95% | 95% |
| 5 | 79% | 96% |
| 20 | 54% | 95% |
The interval narrowed with every replicate added while the rate it covers did not move.
What changed
- The interval is computed over items. An outcome with repeated items reports
estimator: wilson_clusteredand carrieseffective_nanddesign_effect. - The BCa bootstrap resamples items rather than rows.
- The printed rate is the interval the bundle stores, rather than one recomputed from
kandn, which would have disagreed with the file once any interval was widened. - The confident-and-wrong rate counts an item once rather than once per replicate.
A run with one replicate per item is unchanged, and still reports estimator: wilson.
Upgrading
Re-read a bundle rather than comparing a 0.2.1 estimate to a 0.3.0 one. The rows are untouched and estimate recomputes from them, so an old bundle can be brought forward without rerunning the evaluation.
See Rates and Wilson and Status.
Still classified Development Status :: 2 - Pre-Alpha, which is accurate.