Skip to content
Preprint

Gaffke's confidence interval for the mean of bounded data is inadmissible but asymptotically efficient

Jul 2026 · 4 citations · ⚡ 1 influential · 15 references
Mathematics Computer Science Engineering

Abstract

Given observations $\mathbf x=(x_1,\dots,x_n)$, Gaffke (2005) defined \[ K_n(\mathbf x)=\mathbb{P}_{\mathbf D}\!\left\{\sum_{i=1}^n x_iD_i\le 1\right\}, \qquad (D_0,D_1,\ldots,D_n)\sim\mathrm{Dirichlet}(1,\ldots,1), \] and conjectured that it is a $p$-value whenever the inputs are independent e-values. Recently, Vlassis and Thomas (2026) proved this conjecture. Inverting the tests for observations in $[0,1]$ gives the confidence interval studied by Learned-Miller and Thomas (2020), which reduces to Clopper--Pearson for Bernoulli data. We give a finite- and large-sample account of Gaffke's test and interval. First, for every $\mathbf x\in[0,\infty)^n$ and every elementary symmetric polynomial $e_k$, \( K_n(\mathbf x)e_k(\mathbf x)\le {n\choose k}, \) so the Gaffke $p$-value never larger than the SymPol $p$-value of Ming et al. (2026). However, Gaffke's p-value is inadmissible. For $n=2$, we construct a valid rule that is strictly smaller on mixed configurations and is the unique admissible rule that dominates $K_2$. A neutral-face extension proves inadmissibility of $K_n$ for every $n\ge2$. If one independent uniform random variable is allowed, there is an even simpler full-dimensional improvement: on the upper orthant, where $K_n(\mathbf x)=1/\prod_i x_i$, replace it by $U/\prod_i x_i$. The equal-tail Gaffke confidence interval $I_n$ is nevertheless first-order asymptotically efficient: for iid observations on $[0,1]$ with unknown variance $\sigma^2>0$, \[ \sqrt n\,\operatorname{Width}(I_n)\longrightarrow 2\sigma z_{1-\alpha/2}\qquad\text{almost surely}. \] Our simulations also find that, among a variety of bounded-mean intervals considered, the Gaffke interval is the shortest, including comparisons with a recent empirical Berry--Esseen procedure having the same first-order Gaussian target.

View source

Similar papers

Preprint Jul 2026

An Exact Distribution-Free Test for Means of Nonnegative Random Variables

Let $X=(X_1,\ldots,X_n)$ be independent nonnegative random variables, not necessarily identically distributed. Let $D=(D_0,D_1,\ldots,D_n)\sim\operatorname{Dir}(1,\ldots,1)$ be independent of $X$, and define $K(x)=\mathbb{P}\{\sum_{i=1}^n x_iD_i\le1\}$. We prove that, for every $n\ge1$, whenever $\mathbb{E} X_i\le1$ for every $i$, $\mathbb{P}\{K(X)\le\alpha\}\le\alpha$ for all $0\le\alpha\le1$. Thus $K(X)$ is a finite-sample, distribution-free $p$-value for testing the null hypothesis $\mathbb{E}X_i \le 1$ for all $i$. This proves a conjecture of Gaffke (2005).

Nikos Vlassis, P. Thomas · 6 citations · ⚡3
Preprint Jul 2026

On a conjecture of Lamkin and Tkocz: Log-convexity of moments of Bernoulli sample means

Let $X_1,X_2,\ldots$ be independent $\mathrm{Bernoulli}(\theta)$ random variables, and let $\bar X_n = n^{-1}(X_1 + \cdots + X_n)$. We prove that, for every real $p \geq 1$, the sequence $\{\mathsf{E}(\bar X_n^p)\}_{n \geq 1}$ is log-convex. This proves the Bernoulli case of a conjecture of Lamkin and Tkocz [Canad. Math. Bull., 65(2):271-278, 2022]. The proof conditions on the total number of successes among $2n$ trials and reduces the desired inequality to a convex-order comparison between two normalized quadratic functions of hypergeometric random variables. All the log-convexity inequalities are strict for $p>1$ and $0<\theta<1$.

Frédéric Ouimet · 1 citation
Preprint Aug 2026

Sharp Tail Bounds Beyond Twice the Mean

Consider $n$ independent, non-negative, mean at most one random variables, $X_1,X_2,\ldots$. We show the following bound on the probability of their sum exceeding a threshold $t$: \[ \mathbb{P}\left[\sum_{i=1}^n X_i\ge t\right] \leq 1-\left(1-\frac{1}{t}\right)^n \text{ for all } t\ge 2n+1 \,. \] To prove this, we consider a relaxed optimization problem over a set of sequences of ordered, but non-independent random variables. This allows us to reformulate it recursively as dynamic programming problem. The bound becomes an equality for the binary i.i.d.~random variables satisfying $\mathbb{P}\left[X_i=0\right]= 1-\frac{1}{t}$ and $\mathbb{P}\left[X_i=t\right]=\frac{1}{t}$, which remains the maximizer in the relaxed problem.

P. Strack, Jannik M. Westermann · 1 citation
Preprint Aug 2026

Empirical likelihood confidence regions for ordered bivariate means

Let $\boldsymbol{X}_i=(X_{1i},X_{2i})^\top$ be independent and identically distributed observations with mean $\boldsymbol{\mu}=(\mu_1,\mu_2)^\top$ constrained by $\mu_1\leq\mu_2$. We study empirical-likelihood inference for a fixed mean vector and distinguish it from the previously known test of equality against an ordered alternative. At a fixed interior point, the constrained empirical likelihood ratio has the usual $\chi^2_2$ limit. At a fixed boundary point $(m,m)^\top$, its limit is the chi-bar-square distribution $\tfrac12\chi^2_1+\tfrac12\chi^2_2$. By contrast, profiling the unknown common mean in the equality-versus-order test yields $\tfrac12\chi^2_0+\tfrac12\chi^2_1$, the $k=2$ ordered-mean case of El Barmi (1996). We give an exact reduction of the latter statistic to the empirical likelihood of the paired differences, establish the localization step needed for the fixed-boundary expansion, and derive a local-to-boundary limit showing that interior calibration is not uniform over $n^{-1/2}$-neighborhoods of the boundary. Monte Carlo experiments under Gaussian, Student $t_5$, and shifted log-normal sampling examine fixed, boundary, and local regimes with explicit numerical-failure accounting. Illustrative paired-data analyses show the practical distinction between fixed-candidate confidence regions, directional equality tests, and ordinary scalar empirical-likelihood intervals truncated to the nonnegative parameter space.

N. Garg · 0 citations
Preprint Aug 2026

Gaussian-efficient testing by betting on the mean of bounded data

Given $[0,1]$-valued random variables $X_1,\dots,X_n$ such that $\mathbb{E}[X_i | X_1,\dots,X_{i-1}]= \mu$ for all $i$, we propose a new nonasymptotic confidence interval for $\mu$ that is obtained by inverting terminal e-values generated by a novel betting strategy. When the data are iid, its limiting width matches that of the central limit theorem (``Gaussian-efficient''), finally surpassing the inefficient limits of previous betting intervals. Our main conceptual advance involves designing betting fractions that track the conditional rejection probability of the most powerful terminal test in a limiting Gaussian experiment. When one predictable variance estimator is shared across candidate means, the deterministic inversion is an interval for every data sequence and its two endpoints can be found easily. The width can be improved further with external randomization. In simulations, our method yields the tightest intervals to date; for every distribution tested and all sufficiently large $n$, our deterministic version beats STaR-Bets and is competitive with Gaffke, while the randomized improvement beats both. It thus combines finite-sample validity under martingale dependence, easy endpoint computation, Gaussian-efficient inference for iid data, and excellent empirical performance. We also extend the construction and its efficiency theory to sampling without replacement, where it again achieves state-of-the-art empirical performance.

Diego Martinez-Taboada, Aaditya Ramdas · 0 citations
Preprint Jul 2026

The $L_1$-Discrepancy with Nonnegative Weights Suffers from the Curse of Dimensionality

We prove that the $L_1$-discrepancy with arbitrary nonnegative weights suffers from the curse of dimensionality. More precisely, for every $\varepsilon \in (0,1)$ and $d \in \mathbb{N}$, the inverse of the $L_1$-discrepancy satisfies \[ N_{1,+}(\varepsilon, d) \ge \frac{(1-\varepsilon)^2}{1 + \varepsilon} \left( \frac{3+2 \sqrt{3}}{6}\right)^d, \] where $(3+2\sqrt{3})/6 = 1.07735\ldots$. The proof combines a change to a volume-biased probability measure with a fractional-moment estimate for the normalized discrepancy function. The lower bound applies, in particular, to equally weighted point sets. The argument uses the nonnegativity of the weights in an essential way and does not cover arbitrary signed weights.

Josef Dick · 1 citation