In this work, we study the problem of testing conditional independence between random variables $X$ and $Y$ given a confounder $Z$. The local permutation test (LPT) offers a principled approach to this problem by partitioning the $Z$-space into pre-specified bins, and permuting the $X$ and $Y$ data within each bin, to assess the significance of an observed test statistic. However, when the partitions are pre-fixed, the resulting partition can be poorly balanced, as some bins may contain most of the samples while others contain only a few. This motivates the use of data-adaptive binning strategies, such as equisized bins with a fixed (typically small) number of points. We study this natural and practically important extension of LPT, providing finite-sample bounds on the Type I error for an arbitrary test statistic, providing stronger validity results than previously known. We also show that LPT attains power comparable to the oracle likelihood ratio tests derived from the Neyman-Pearson lemma. Within a linear confounder model class, we further analyze the effect of bin size and demonstrate that constant bin sizes can match the performance of partitions with growing bin-size. These results, further supported by extensive numerical simulations, position the proposed data-adaptive strategy as both practically implementable and statistically efficient.
Given $[0,1]$-valued random variables $X_1,\dots,X_n$ such that $\mathbb{E}[X_i | X_1,\dots,X_{i-1}]= \mu$ for all $i$, we propose a new nonasymptotic confidence interval for $\mu$ that is obtained by inverting terminal e-values generated by a novel betting strategy. When the data are iid, its limiting width matches that of the central limit theorem (``Gaussian-efficient''), finally surpassing the inefficient limits of previous betting intervals. Our main conceptual advance involves designing betting fractions that track the conditional rejection probability of the most powerful terminal test in a limiting Gaussian experiment. When one predictable variance estimator is shared across candidate means, the deterministic inversion is an interval for every data sequence and its two endpoints can be found easily. The width can be improved further with external randomization. In simulations, our method yields the tightest intervals to date; for every distribution tested and all sufficiently large $n$, our deterministic version beats STaR-Bets and is competitive with Gaffke, while the randomized improvement beats both. It thus combines finite-sample validity under martingale dependence, easy endpoint computation, Gaussian-efficient inference for iid data, and excellent empirical performance. We also extend the construction and its efficiency theory to sampling without replacement, where it again achieves state-of-the-art empirical performance.
Diego Martinez-Taboada, Aaditya Ramdas· 0 citations
We introduce a distribution-free goodness-of-fit test, termed the omega-1 test, which naturally complements the Kolmogorov--Smirnov test and Cram\'{e}r--von Mises test and can be viewed as their (piecewise) linear analog. Defined as an $\mathrm{L}^{1}$-functional of the empirical process, the test statistic improves on balancing sensitivity to localized and diffuse alternatives and gives a robust and interpretable measure of distributional discrepancy, apart from close connections to the Wasserstein 1-distance. For finite samples, we derive a finite-dimensional computational form for the statistic under general conditions, which leads to various explicit formulas for its null distribution. Under mild continuity assumptions, the limiting statistic is distribution-free, with explicit distribution formulas. In composite settings, the statistic is also compatible with the Khmaladze transformation, enabling asymptotically distribution-free testing. The limiting transformed statistic also has an explicit distribution that escapes reliance on intractable compensator processes or purely numerical evaluation. Simulation results indicate rapid convergence of the finite-sample distributions to their limiting counterparts and support the practical applicability of the test.
Testing whether two independent samples arise from the same underlying distribution is a fundamental statistical problem. We propose a new class of two-sample distribution tests based on a family of probability metrics $\Delta_{n,p}$, constructed from repeatedly integrated quantile functions. On their respective domains, these metrics are proved to be genuine distributional distances. The case $n=1$ recovers the $p$-Wasserstein distance, which requires finite $p$-th moments; for $n\geq2$, the proposed metrics are well defined and require only finite first moments. The asymptotic properties of the plug-in statistic are established, including strong consistency and limiting distributions under the null and fixed alternatives. A permutation calibration for finite-sample inference is also proposed. We further derive an asymptotic power function under local alternatives. Finally, the finite-sample performance of the proposed tests is examined through simulation studies, and their reduced sensitivity to extreme upper-tail observations is illustrated through a real data application.
Marginal Mann-Whitney effects are widely used across various fields of research, and extensions of this estimand have been developed in many directions in statistical methodology. In this paper, we focus on an extensions for repeated measurements and factorial designs subject to randomly missing data. In a previous work by Rubarth et al. (2022a), asymptotically correct tests were developed under the assumption of deterministic missing indicators. In contrast, the approach in the present paper accounts for the stochastic nature of missing values under realistic mechanisms. Thus, the involved covariance matrix incorporates the true variability of missing data. The combination with a randomization procedure using random permutations within each data point yields asymptotically exact tests and a generally improved type-I error control. Additionally, the tests control the type-I error for finite sample sizes in the special case of exchangeable sampling distributions. Simulations across a wide range of settings demonstrate the benefits of the proposed method in small samples, also for different missingness mechanisms. A real data analysis about school children learning math illustrates several practical aspects of the tests'application.
Dennis Dobler, Jörg-Tobias Kuhn, L. Amro et al.· 0 citations
We introduce a flexible model for covariate-dependent multiple testing which can be encoded using a nonparametric Gaussian mixture model. Weight-localized predictive recursion (PRx), a new development in the methodology of Newton's predictive recursion algorithm, is then leveraged to estimate the components of this mixture model, allowing for recovery of the covariate-localized false discovery rate $\text{Pr}(H_i = 0|z_i,x_i)$ using a single, unified algorithm. This quantity represents the most direct extension of Efron's local false discovery rate to the covariate-dependent setting, and admits provable Bayesian FDR control properties under simple rejection rules. We introduce several procedures for estimating and thresholding the local false discovery rate, and show using various simulations and a real-data example that our procedures lead to increased power, tighter Bayesian FDR control, and more interpretable rejections. We furthermore show that this holds for fixed and randomized hypothesis labels, indicating that our proposed methods perform well under both frequentist and Bayesian interpretations of multiple testing.
The recent work of Sarkar and Zhang (2025) introduced Positive Tail Dependence Under the Null (PTDN) and developed Generalized Shifted Benjamini-Hochberg (BH) procedures for two-sided Gaussian $z$- and $t$-testing under known covariance structures. This paper develops further consequences of that framework. First, we derive explicit dependence-adaptive lower and upper bounds for the FDR of the original BH procedure in terms of the conditional variance parameters $\tau_i=1-R_i^2$, where $R_i^2$ is the squared multiple correlation between the $i$th statistic and the remaining coordinates. These bounds recover the exact BH FDR under independence and provide finite-sample, covariance-specific information complementary to generic bounds. We also identify conditions under which the coordinate-specific calibration of shifted BH can provide a rejection advantage over the original BH procedure. Second, we consider the practically important setting in which the covariance matrix is unknown but an independent Wishart estimator is available. Using simultaneous lower confidence bounds for the $\tau_i$'s, we construct a confidence-bound shifted BH procedure and establish finite-sample FDR control. To our knowledge, this is the first shifted-BH-type procedure with a finite-sample guarantee for two-sided Gaussian mean testing under a completely unknown covariance matrix estimated independently. Numerical studies illustrate the behavior of the covariance-adaptive bounds, the potential advantage of shifted BH over BH, and the performance of confidence-bound shifting under unknown covariance.
D. Ghosh, S. Sarkar· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.