Semi-tail Units in Statistical Testing
- Semi-tail units are statistical measures defined as s = -log2(p), offering a universal scale for quantifying extremeness in hypothesis tests.
- They enable direct additivity by summing s-values, which facilitates the combination of evidence from independent tests.
- They standardize critical thresholds and measure efficiency improvements, providing clearer comparisons across various test statistics.
A semi-tail unit is a quantile-based measure for expressing extremeness, efficiency, and evidence in statistical hypothesis testing. Each semi-tail unit corresponds to a halving of the tail area under a null distribution, and this logarithmic base-2 scale permits direct additivity, uniform critical value progression, and natural efficiency quantification. Semi-tail units have been proposed as a universal, interpretable alternative to p-values, z-scores, and test-specific quantiles, applicable to any ordered statistical distribution (Vos, 28 Jun 2025).
1. Definition and Mathematical Formulation
Let be a real-valued test statistic, with the right-tail probability under the null hypothesis. The semi-tail unit, denoted , is defined by
which quantifies how many doublings of rarity have occurred in the observed tail probability. For example, corresponds to , to , and to . In other words, each increment of 1 in 0 represents a halving of 1. For two-tailed tests, the 2-value is employed: 3 where 4 is the median under 5.
This formulation enables direct mapping between percentile, tail probability, and semi-tail unit via
6
for upper-tail percentile 7. For instance, a result at the 99th percentile yields 8.
2. Interpretive Properties and Practical Benchmarks
The primary interpretive property is additivity and halving: every 9 increment in 0 divides the tail probability by two. Table 1, mapping selected probabilities and percentiles to 1:
| Tail Probability 2 | Semi-tail 3 | Percentile |
|---|---|---|
| 4 | 5 | 6 |
| 7 | 8 | 9 |
| 0 | 1 | 2 |
| 3 | 4 | 5 |
For two-tailed tests, the two-tailed 6-value is 7. This allows writing all critical values as equally spaced points on the 8 (or 9) scale, eliminating the distribution-specific tables common to classical hypothesis testing.
3. Conversion from and to Classical Test Statistics
To convert any test statistic:
- Compute the classical 0-value under the null.
- Set 1 (one-tailed) or 2 (two-tailed).
Examples:
- Normal 3-test, 4: two-tailed 5.
- 6-test with 7, 8: two-tailed 9.
- 0 test, 1: one-tailed 2.
- Poker hand "royal flush": 3 (tail prop 4).
All critical 5-levels become 6, so achieving a more stringent level is an arithmetic progression in 7.
4. Evidence Combination and Additivity
If 8 are independent one-tailed 9-values, their joint tail probability is 0, so the aggregate 1-value is
2
This allows for direct addition of semi-tail units when accumulating evidence across studies or tests. For example, three 3-values of 4, 5, and 6 yield 7, 8, and 9, so 0.
5. Semi-tail Units and Bahadur Efficiency
Semi-tail units naturally encode efficiency in terms of Bahadur slopes. The Bahadur exact slope for a sequence 1 under an alternative 2 can be written as
3
and the asymptotic Bahadur slope becomes 4 with 5. Thus, the 6-value per sample is an interpretable "bits" or "semi-tail units per observation" rate.
Semi-tail efficiency difference between two tests 7 is 8. A positive 9 means test 0 drives the sample 1 units deeper into the tail (i.e., 2 smaller 3-value) compared to test 4, an interpretable "distance" in evidence (Vos, 28 Jun 2025).
6. Unification and Advantages over Classical Scales
Semi-tail units provide a universal, interpretable, and additive scale for expressing the extremeness of results across all testing contexts. Key unifying features:
- All decision thresholds are linear and equally spaced (5 increases by 6 when 7 halves).
- No distribution-specific lookup tables are required.
- Additivity for combining independent evidence (direct sum of 8).
- Simple efficiency comparison: differences in 9 per sample correspond to multiplicative changes in tail probability/explanatory power.
- Applicable to any ordered distribution, not just those with tabulated quantiles (e.g., permutation tests, empirical distributions, and discrete combinatorial ranks).
These properties, as demonstrated in applications from normal and 0-tests to poker hand rankings, distinguish the semi-tail unit from p-values, z-scores, and conventional quantile measures (Vos, 28 Jun 2025).
7. Worked Examples and Applications
Percentile to 1-value mapping:
- 2
- 3
- 4
- 5
Discrete ranking (combinatorial, e.g., poker hands):
Let 6 be total outcomes, 7 the rank or better, 8. Royal flush: 9.
Combining 00-values:
Three experiments give 01-values 02, 03, 04; sum is 05, corresponding to 06.
Efficiency:
The difference in semi-tail slope per sample (07) quantifies how much more rapidly one test accumulates evidence, providing a bits-scale improvement directly interpretable (e.g., 08 is a 09 reduction in 10 per observation in the asymptote).
Semi-tail units constitute a universal and principled system for standardization and comparison of test statistics, evidence accumulation, and efficiency reporting in modern statistical research (Vos, 28 Jun 2025).