Jackknife-based test statistics are inferential procedures that use leave-one-out replicates and pseudovalues to estimate bias, variance, and construct test statistics.
The classical approach computes delete-one replicates to form pseudovalues and yields Wald-type and chi-square calibrated tests for smooth, iid data.
Modern methods, including jackknife empirical likelihood and infinitesimal jackknife, extend these ideas to handle nonlinear U-statistics, model comparison, and complex data structures.
Jackknife-based test statistics are inferential procedures constructed from leave-one-out replicates, jackknife pseudo-values, or infinitesimal perturbations of data weights. In the classical formulation, the jackknife is used for bias estimation, standard error estimation, confidence limits, and test-statistic construction via pseudo-values; in later developments, jackknife empirical likelihood (JEL) converts nonlinear or U-statistic estimating problems into likelihood-ratio statistics with Wilks-type chi-square limits, while infinitesimal-jackknife methods extend the same perturbation logic to covariance estimation, model comparison, and local approximations to cross-validation and bootstrap re-fitting (McIntosh, 2016, N et al., 2017, Ghosal et al., 2022, Giordano et al., 2019).
1. Classical jackknife construction and the pseudovalue test
The delete-1 jackknife starts from data X1​,…,Xn​ and forms the jackknife samples
The same construction yields the jackknife bias estimate
biasjack​=(n−1)(θ^(⋅)​−θ^),
and the bias-corrected estimator
θ^jack​=nθ^−(n−1)θ^(⋅)​.
The central inferential object is the pseudovalue
X1​,…,Xn​0
The average pseudovalue equals the jackknife estimator, and the pseudovalues can be treated as if they were independent random variables. This leads to the classical asymptotic normal test statistic
X1​,…,Xn​1
where X1​,…,Xn​2 is the usual unbiased sample standard deviation computed from pseudovalues. In this form, jackknife-based testing is a Wald-type procedure centered on the average pseudovalue rather than on a direct plug-in asymptotic variance formula (McIntosh, 2016).
For two-sample statistics, the same logic is formalized through asymptotic linearity. If
X1​,…,Xn​3
with
X1​,…,Xn​4
then the jackknife variance estimator is consistent and asymptotically unbiased for the asymptotic variance
X1​,…,Xn​5
This framework is used for common mean estimators under ordered variances, where the estimators are piecewise-defined through random weights and the variance analysis is difficult because the estimator changes form depending on whether X1​,…,Xn​6 or X1​,…,Xn​7. Under the stated assumptions, these common-mean estimators satisfy a CLT and admit large-sample Wald intervals based on jackknife variance estimation (Steland et al., 2017).
A persistent limitation of the classical jackknife is already explicit in the early resampling literature: it is intended for iid data and smooth or sufficiently linear functionals, is not suitable for correlated data or time series, and may fail for non-smooth estimators such as the median. This limitation became one of the main motivations for pseudo-value likelihood methods and infinitesimal perturbation methods (McIntosh, 2016).
2. Jackknife empirical likelihood and Wilks-type ratio statistics
The modern theory of jackknife-based test statistics is dominated by JEL. Across S-Gini indices, probability weighted moments, and Gini correlations, the construction follows the same pattern: start from a X1​,…,Xn​8-statistic or a ratio of X1​,…,Xn​9-statistics, rewrite the parameter as the solution of a mean-zero estimating equation, construct jackknife pseudo-values
and apply ordinary empirical likelihood to the linear constraint induced by the pseudo-values. For the relative S-Gini index X[i]​={X1​,X2​,…,Xi−1​,Xi+1​,…,Xn​}.1, for example, the estimating equation is
The same paper emphasizes the computational rationale: the jackknife converts a nonlinear s(â‹…)0-statistic inference problem into a linear empirical likelihood problem (N et al., 2017).
For probability weighted moments
s(â‹…)1
the unbiased estimator is a s(â‹…)2-statistic based on
s(â‹…)3
and the JEL ratio
s(â‹…)4
satisfies
s(â‹…)5
The adjusted version AJEL appends the pseudo-value
s(â‹…)6
to avoid convex-hull problems; under s(â‹…)7,
s(â‹…)8
Both JEL and AJEL are then used for confidence intervals and for testing
For Gini correlations, the parameter is a ratio of two i0-statistics,
i1
and the estimating kernel is
i2
The resulting single-parameter JEL ratio satisfies
i3
The same framework also covers testing the equality of the two Gini correlations,
i4
and two-sample comparisons of the difference vector
i5
A recurring theme in these papers is that JEL avoids explicit asymptotic variance estimation while preserving a standard Wilks-type limit under finite-second-moment and non-degeneracy conditions (Sang et al., 2018).
θ^(i)​:=s(X[i]​),θ^(⋅)​=n1​i=1∑n​θ^(i)​.7 or θ^(i)​:=s(X[i]​),θ^(⋅)​=n1​i=1∑n​θ^(i)​.8
θ^(i)​:=s(X[i]​),θ^(⋅)​=n1​i=1∑n​θ^(i)​.9 or Varjack​(θ^)=nn−1​i=1∑n​(θ^(i)​−θ^(⋅)​)2,0
3. Symmetry, goodness-of-fit, and independence tests built from jackknifed Varjack​(θ^)=nn−1​i=1∑n​(θ^(i)​−θ^(⋅)​)2,1-statistics
Several one-sample goodness-of-fit and symmetry problems are now handled by first constructing a scalar departure measure and then applying JEL to the associated pseudo-values. For diagonal symmetry in multivariate data, the null
The natural sample statistic is a difference of two Varjack​(θ^)=nn−1​i=1∑n​(θ^(i)​−θ^(⋅)​)2,4-statistics, but under Varjack​(θ^)=nn−1​i=1∑n​(θ^(i)​−θ^(⋅)​)2,5 it is degenerate, which prevents direct application of standard JEL theory. The paper resolves this by splitting the sample, constructing two pseudo-value arrays Varjack​(θ^)=nn−1​i=1∑n​(θ^(i)​−θ^(⋅)​)2,6 and Varjack​(θ^)=nn−1​i=1∑n​(θ^(i)​−θ^(⋅)​)2,7, enforcing a common-mean constraint, and proving the Wilks-type limit
In practice the null is SE(θ^)jack​={nn−1​i=1∑n​(θ^(i)​−θ^(⋅)​)2}1/2.7, so independence is tested by SE(θ^)jack​={nn−1​i=1∑n​(θ^(i)​−θ^(⋅)​)2}1/2.8 against a SE(θ^)jack​={nn−1​i=1∑n​(θ^(i)​−θ^(⋅)​)2}1/2.9 critical value (N. et al., 2021).
For log symmetry on the positive real line, the construction begins from the characterization
biasjack​=(n−1)(θ^(⋅)​−θ^),0
for a positive, real-valued, strictly monotone continuous function biasjack​=(n−1)(θ^(⋅)​−θ^),1. Choosing
biasjack​=(n−1)(θ^(⋅)​−θ^),2
produces the PWM-based departure measure
biasjack​=(n−1)(θ^(⋅)​−θ^),3
which under algebraic rewriting becomes
biasjack​=(n−1)(θ^(⋅)​−θ^),4
The associated biasjack​=(n−1)(θ^(⋅)​−θ^),5-statistic leads to pseudo-values biasjack​=(n−1)(θ^(⋅)​−θ^),6, a JEL ratio
biasjack​=(n−1)(θ^(⋅)​−θ^),7
and the limit
biasjack​=(n−1)(θ^(⋅)​−θ^),8
The explicit motivation is that the normal-based test is difficult to implement because the asymptotic variance is difficult to estimate (S et al., 2024).
The same strategy appears in goodness-of-fit testing for the standard Cauchy law. Using the characterization of Arnold (1979), the paper defines the discrepancy
biasjack​=(n−1)(θ^(⋅)​−θ^),9
estimates it by an unbiased order-3 θ^jack​=nθ^−(n−1)θ^(⋅)​.0-statistic, forms jackknife pseudo-values
θ^jack​=nθ^−(n−1)θ^(⋅)​.1
and proves
θ^jack​=nθ^−(n−1)θ^(⋅)​.2
The AJEL variant appends
θ^jack​=nθ^−(n−1)θ^(⋅)​.3
to enforce feasibility when the convex hull of the jackknife pseudo-values fails to contain zero (Vishnu et al., 2024).
4. Two-sample, θ^jack​=nθ^−(n−1)θ^(⋅)​.4-sample, and multivariate jackknife likelihood tests
A major branch of the literature uses jackknife-based test statistics for equality of distributions or equality of functionals across several samples. In the upper-semivariance problem, two independent nonnegative populations with cdfs θ^jack​=nθ^−(n−1)θ^(⋅)​.5 and θ^jack​=nθ^−(n−1)θ^(⋅)​.6 are compared under
θ^jack​=nθ^−(n−1)θ^(⋅)​.7
where
θ^jack​=nθ^−(n−1)θ^(⋅)​.8
The paper constructs the scalar departure measure
θ^jack​=nθ^−(n−1)θ^(⋅)​.9
rewrites it as expectations involving pairwise comparisons, and estimates it by a symmetric-kernel X1​,…,Xn​00-statistic X1​,…,Xn​01. After pooling the sample and forming pseudo-values
X1​,…,Xn​02
the JEL ratio
X1​,…,Xn​03
yields the test statistic X1​,…,Xn​04, with null limit X1​,…,Xn​05. The practical appeal is explicit: JEL avoids the practical difficulty of variance estimation in the normal-based method (Suresh et al., 2024).
For the X1​,…,Xn​06-sample homogeneity problem, the paper reformulates
X1​,…,Xn​07
as testing independence between a continuous random vector X1​,…,Xn​08 and a categorical variable X1​,…,Xn​09. The categorical Gini correlation
X1​,…,Xn​10
characterizes equality of the X1​,…,Xn​11 distributions through X1​,…,Xn​12. The test is built from the Gini contrast
X1​,…,Xn​13
with pooled and groupwise pseudo-values X1​,…,Xn​14 and X1​,…,Xn​15. The resulting JEL log-likelihood ratio
X1​,…,Xn​16
satisfies
X1​,…,Xn​17
A distinctive computational claim is that no permutation procedure is required, unlike many energy-distance or distance-correlation tests (Sang et al., 2019).
The multivariate X1​,…,Xn​18-sample extension treats X1​,…,Xn​19-statistics based on three or more independent samples. For three samples,
X1​,…,Xn​20
the construction pools all observations into X1​,…,Xn​21, defines leave-one-out statistics X1​,…,Xn​22, and forms
X1​,…,Xn​23
Because the pseudo-values are not identically distributed across the sample blocks, the JEL constraint is written as
X1​,…,Xn​24
Under finite-second-moment, non-degeneracy, and sample-balance conditions,
X1​,…,Xn​25
The paper develops this for confidence intervals for differences in VUS measurements and for HUM-type functionals, and repeatedly contrasts JEL with normal approximation and kernel-smoothed bootstrap intervals (Garg et al., 2024).
These multi-sample results show a characteristic feature of jackknife-based testing: the underlying estimand can be a high-order, multi-sample, or multivariate X1​,…,Xn​26-statistic, but the final inferential object is still a low-dimensional empirical-likelihood ratio calibrated by a standard chi-square law.
5. Infinitesimal jackknife, higher-order expansions, and covariance-based tests
The infinitesimal jackknife (IJ) replaces delete-1 recomputation by local differentiation with respect to data weights. For weighted estimating equations
X1​,…,Xn​27
the higher-order infinitesimal jackknife (HOIJ) is the Taylor expansion of X1​,…,Xn​28 in the weights X1​,…,Xn​29 around the all-ones vector X1​,…,Xn​30. The first derivative recovers the ordinary IJ,
X1​,…,Xn​31
and higher-order derivatives are computed recursively from lower-order derivatives and higher-order derivatives of X1​,…,Xn​32. The core recursion uses only one matrix inverse X1​,…,Xn​33 for all orders, and the paper proves finite-sample error bounds of the form
X1​,…,Xn​34
This work does not develop an explicit hypothesis test statistic or confidence interval, but it makes an important inferential point: for bootstrap weights, the first-order approximation reproduces the usual sandwich covariance estimator, so covariance estimates based on the linear approximation do not improve on asymptotic normal theory; higher-order terms are needed for higher-order bootstrap accuracy (Giordano et al., 2019).
A direct testing application of the IJ appears in model comparison. Extending the IJ from variance estimation to covariance estimation between two models fitted on the same data, the paper defines directional derivatives X1​,…,Xn​35 and estimates the covariance block by
X1​,…,Xn​36
For two predictors evaluated at X1​,…,Xn​37,
X1​,…,Xn​38
the covariance of the difference is estimated by
X1​,…,Xn​39
and the comparison test statistic is
X1​,…,Xn​40
The same machinery is used for testing whether a boosting stage made a statistically significant change, and for uncertainty quantification of sums, differences, and more general linear combinations of models (Ghosal et al., 2022).
Predictive inference provides a different jackknife-based branch. The jackknife+ replaces the classical jackknife interval centered at X1​,…,Xn​41 by an interval based on the leave-one-out fitted values at the test point,
The same paper is explicit that the original jackknife has no universal guarantee and can have coverage equal to zero. This is not a likelihood-ratio test, but it is an important corrective to the misconception that all leave-one-out jackknife inference is automatically valid in unstable learning problems (Barber et al., 2019).
6. High-dimensional IV inference, Bayesian survey pseudo-likelihoods, and scope conditions
In linear IV regression with endogeneity, heteroskedastic disturbances, and many potentially weak instrumental variables, jackknife ideas have been used to construct a full testing trinity. The model is
X1​,…,Xn​44
and the paper introduces jackknife objective functions based on quadratic forms and ratio-of-quadratic-forms. For the simple null
X1​,…,Xn​45
the proposed statistics are
X1​,…,Xn​46
For general linear restrictions
X1​,…,Xn​47
the paper defines restricted estimators and corresponding statistics X1​,…,Xn​48, X1​,…,Xn​49, X1​,…,Xn​50, and X1​,…,Xn​51. Under the null, the natural asymptotic limit is a weighted sum of chi-squares,
X1​,…,Xn​52
where X1​,…,Xn​53. A central contribution is that modified objective functions produce ordinary chi-square limits,
X1​,…,Xn​54
This is an explicitly jackknife-based testing framework rather than a variance-estimation device, and it is designed for settings where many potentially weak instruments make standard IV inference fragile (Crudu et al., 16 Apr 2026).
In complex survey sampling, jackknife pseudo-values also support a Bayesian pseudo-likelihood route. For a X1​,…,Xn​55-statistic parameter X1​,…,Xn​56, the pseudo-values are
X1​,…,Xn​57
and the survey-weighted jackknife empirical likelihood imposes
X1​,…,Xn​58
optionally together with the auxiliary-information constraint
X1​,…,Xn​59
Under a non-informative prior X1​,…,Xn​60, the Bayesian jackknife pseudo-empirical likelihood pseudo-posterior is asymptotically normal, with local expansion around the JEL estimator or its weighted analogue. Inference is then obtained through posterior quantiles, and testing proceeds by inversion of credible sets rather than by a frequentist likelihood-ratio cutoff (Shang et al., 2023).
Taken together, these developments delineate the scope conditions of jackknife-based testing. Recurrent assumptions are finite second moments of the kernel, non-degenerate first-order projections, and asymptotic balance of sample fractions. Where those conditions fail, several papers introduce explicit repairs: AJEL adds an artificial pseudo-value to fix convex-hull failures, HOIJ uses higher-order derivatives because first-order bootstrap linearization is equivalent to asymptotic normal theory, and jackknife+ modifies the classical jackknife because leave-one-out residual calibration alone does not control instability (Bhati et al., 2018, Giordano et al., 2019, Barber et al., 2019).
The resulting picture is broad but coherent. Jackknife-based test statistics include classical pseudovalue X1​,…,Xn​61-statistics, JEL likelihood ratios for scalar and vector X1​,…,Xn​62-statistic functionals, chi-square and chi-bar-square tests in many-instrument IV models, covariance-aware chi-square tests for model comparison, and Bayesian pseudo-likelihood procedures for complex surveys. What unifies them is not a single formula but a common reduction: a difficult estimator is recast into leave-one-out or infinitesimal perturbation objects that behave like approximately independent estimating values, after which standard normal, empirical-likelihood, or quadratic-form calibration becomes possible.
“Emergent Mind helps me see which AI papers have caught fire online.”
Philip
Creator, AI Explained on YouTube
Sign up for free to explore the frontiers of research
Discover trending papers, chat with arXiv, and track the latest research shaping the future of science and technology.Discover trending papers, chat with arXiv, and more.