asymptotic inference

Deriving large-sample (asymptotic) distributions and limits—e.g., via central limit theorems—to quantify uncertainty, produce valid confidence/credible intervals, and characterize null distributions and variance estimators under specified model assumptions.

asymptoticinference

Recent Skill Trend

Momentum and market value over time
Trending
Score
No comparison yet
-0.02
Aug 01, 2026Aug 01, 2026
Career
Value
No comparison yet
$200K/year
Aug 01, 2026Aug 01, 2026

Recommended Survey Paper

Quick overview of the field
View more

Must-Read Papers

Most classic and influential ideas
View more

Can we have it all? Non-asymptotically valid and asymptotically exact confidence intervals for expectations and linear regressions

Jul 22, 2025
AD
Alexis Derumigny
🏛️ Delft University of Technology | CREST-ENSAE | INRAE

This paper addresses the challenge of constructing confidence sets (CSs) for semiparametric models that simultaneously achieve finite-sample reliability and asymptotic efficiency. We introduce the novel concept of “non-asymptotically valid and asymptotically exact” (NAVAE) CSs and establish sufficient conditions for their existence. Under mild moment conditions—such as bounded kurtosis or weak exogeneity—we construct closed-form NAVAE confidence intervals for linear combinations of expectations and regression coefficients. These intervals moderately widen classical central-limit-theorem-based intervals to ensure non-asymptotic coverage validity, while embedding a uniform asymptotic exactness framework to robustly accommodate heteroskedasticity and weakly exogenous covariates. Simulation studies demonstrate accurate finite-sample coverage and optimal asymptotic convergence rates, substantially outperforming conventional asymptotic methods. Furthermore, we characterize the theoretical limits of the approach under highly skewed distributions, including the Bernoulli case.

Bridging gap between large- and finite-sample inferenceConstructing non-asymptotically valid confidence intervalsEnsuring asymptotic exactness under moment conditions

Adaptive A/B Tests and Simultaneous Treatment Parameter Optimization

Oct 13, 2022
YW
Yuhang Wu
🏛️ University of California, Berkeley | Amazon.com Inc

Classical algorithms for strongly convex stochastic optimization achieve fast convergence (O(1/√n)) but suffer from asymptotically non-negligible bias, violating the conditions required for a valid central limit theorem (CLT) and thus impeding asymptotically efficient statistical inference. Method: We propose the first dual-objective algorithm that simultaneously guarantees fast convergence and a provable CLT. Our approach integrates stochastic approximation, asymptotic statistical inference, and adaptive experimental design into a unified framework that ensures asymptotic normality of the estimator. Contribution/Results: We establish theoretical guarantees that the algorithm retains the O(1/√n) convergence rate while satisfying the CLT. Numerical experiments demonstrate substantial improvements over existing methods in estimation accuracy, confidence interval coverage, and identification of optimal treatment parameters. The method provides a new paradigm for continuous, parameterized A/B testing in online platforms—balancing optimization efficiency with statistical reliability.

Addresses non-vanishing bias in stochastic optimization statistical inferenceDevelops algorithm maintaining fast convergence with valid central limit theoremEnables reliable confidence intervals for optimal objective value estimation

This paper identifies the systematic failure of classical statistical methods—including mean-based inference, principal component analysis (PCA), and asymptotic normality assumptions—under heavy-tailed distributions, particularly in medium-sample-size (medium-*n*) real-world settings where they are routinely misapplied. Methodologically, it challenges the uncritical adoption of Gaussian and stable-distribution assumptions and introduces the “Median Law” theoretical framework, which formalizes fundamental limitations under heavy tails: unreliable sample means, distorted empirical distributions, and degenerate principal components. The approach integrates extreme value theory, generalized stable distribution modeling, robust parametric estimation, and pre-asymptotic analysis. Empirical validation draws on counterexamples from finance, economics, and psychology, supplemented by cross-disciplinary case studies. The core contribution is a foundational rethinking of uncertainty quantification and causal inference: it demonstrates that many canonical “cognitive biases” are, in fact, rational inferences under heavy-tailed probability structures—thereby advocating a paradigm shift in statistical practice from idealized asymptotics to empirically grounded probabilistic modeling.

Examining real-world statistical behavior between small and infinite sample sizesIdentifying failures in economic and psychological models from wrong distributionsInvestigating misapplication of statistical techniques to fat-tailed distributions

Generalized Universal Inference on Risk Minimizers

Jan 31, 2024
ND
N. Dey
🏛️ North Carolina State University

This paper addresses uncertainty quantification for risk-minimizing estimators in machine learning, overcoming limitations of classical approaches that rely on restrictive distributional assumptions and asymptotic theory. We propose the first general-purpose, finite-sample, distribution-free, and frequentist-valid inference framework applicable to *any* risk minimizer. Our method is grounded in the generalized likelihood ratio test, integrated with empirical process analysis and data-driven tuning, and inherently supports anytime-valid inference. Theoretically, it guarantees exact coverage of confidence sets for *all* finite sample sizes—without asymptotic approximations. Empirically, it consistently outperforms classical asymptotic methods across diverse tasks, demonstrating both high accuracy and strong robustness. This work establishes a new paradigm for model-agnostic statistical inference.

Achieve finite-sample validity in statistical learningEstimate unknowns with uncertainty quantificationGeneralize universal inference for risk minimizers

Distribution-Free Calibration of Statistical Confidence Sets

Nov 28, 2024
LM
Luben Miguel Cruz Cabezas
🏛️ Federal University of S~ao Carlos | University of S~ao Paulo

In statistical inference, confidence sets—especially under complex models or small sample sizes—often fail to achieve nominal coverage levels, particularly in likelihood-free inference (LFI) settings. To address this, we propose TRUST and TRUST++, two distribution-free, simulation-based calibration methods that adapt conformal prediction principles to confidence set construction with redundant parameters, thereby establishing the first distribution-agnostic calibration framework for statistical inference. Our methods guarantee finite-sample local coverage and asymptotic conditional coverage, while enabling self-assessment of simulation cost. Theoretically, we prove their robustness against model misspecification and simulation imperfection. Empirically, TRUST and TRUST++ significantly improve coverage accuracy across both tractable and intractable likelihood models, consistently outperforming existing approaches—especially in small-sample regimes.

Achieving distribution-free conditional coverage with simulationsCalibrating confidence sets for valid statistical inferenceHandling nuisance parameters in small-sample complex models

Latest Papers

What's happening recently
View more

This study addresses the limitations of conventional inference methods rooted in sampling variability when sample sizes approach the population size. By constructing finite populations with known parameters and leveraging CPU/GPU-accelerated repeated sampling experiments, the authors examine the evolution of the randomization distribution of the sample mean across varying sampling fractions. Integrating finite population theory with numerical precision analysis, they demonstrate that in high-coverage scenarios, estimation error predominantly stems from computational precision and architectural constraints rather than sampling randomness. The findings reveal that sampling variability becomes negligible well before exhaustive enumeration is reached, thereby challenging a foundational assumption of classical inferential statistics and offering a basis for rethinking statistical paradigms in the context of large-scale, near-complete data.

finite population samplinglarge-scale datasampling fraction

This work addresses the conservatism arising from the infinite-time validity assumption in sequential inference by introducing a “confidence horizon” framework that constructs anytime-valid confidence sequences within a finite time boundary. By integrating group sequential methods with adaptive Neyman allocation, the framework enables early stopping under budgetary or ethical constraints while preserving inferential accuracy. Key contributions include the first incorporation of a finite-time horizon into the anytime-valid inference paradigm, the establishment of explicit connections to classical group sequential boundaries (e.g., Pocock and O’Brien–Fleming), and the derivation of closed-form asymptotic quantiles that circumvent repeated integration. This analytical advance substantially improves the computational efficiency of critical value calculation and enhances the precision of treatment effect estimation.

anytime-valid inferenceconfidence sequencesfinite time horizon

This study addresses Bayesian inference for low-dimensional target parameters in semiparametric models, particularly under the presence of complex nuisance components that may compromise frequentist properties. To this end, we construct posterior distributions by integrating estimating function methods with nonparametric Bayesian techniques—such as Dirichlet processes and Bayesian bootstrap—under conditions weaker than the classical stochastic equicontinuity assumption. We establish asymptotic normality and consistency of the resulting posterior, rigorously identifying the key assumptions required to guarantee desirable frequentist behavior. The theoretical analysis systematically elucidates how relaxing these assumptions affects inferential performance. Extensive simulations corroborate the effectiveness of the proposed methodology, demonstrating its robustness and accuracy in practical settings.

asymptotic normalityBayesian semi-parametric modelsfrequentist properties

This work addresses the challenge of verifying mathematical proofs generated by large language models by formally encoding, for the first time, an entire advanced undergraduate probability textbook—including its measure-theoretic foundations—into Lean. To bridge the semantic gap between the textbook’s exposition and the abstract formalism of the Mathlib library, the authors introduce an “interface lemma” strategy. Combined with structured proof engineering and formalization techniques specific to measure theory, this approach yields a reusable, machine-verifiable infrastructure spanning fourteen textbook chapters. The resulting formalization not only provides rigorous verification of all stated theorems and explicit articulation of their assumptions but also establishes a robust foundation for reliable AI-assisted mathematics, educational applications, and future formalization efforts in probability theory.

formalizationLeanmathematical infrastructure

This study addresses the estimation of parameters of the form θ₀ = E[F_Y⁻¹∘F_Z(X)] in the “changes-in-changes” model, for which existing methods lack theoretical guarantees when variables are unbounded. The authors construct a plug-in estimator based on empirical quantiles and establish its √n-consistency and asymptotic normality under assumptions weaker than those in the current literature. They further propose a novel consistent estimator for the asymptotic variance. The theoretical analysis leverages empirical process theory and plug-in methods for quantile functions. Monte Carlo simulations demonstrate that the proposed variance estimator substantially outperforms existing alternatives, leading to markedly improved inference accuracy.

asymptotic normalitychanges-in-changesempirical quantile

Hot Scholars

AR

Aaditya Ramdas

Associate Professor (with tenure), Carnegie Mellon University
Machine LearningStatistics
SS

Sergey Samsonov

HSE university, Moscow
high-dimensional probabilityMarkov ChainsMCMC
AJ

Arnulf Jentzen

The Chinese University of Hong Kong, Shenzhen (CUHK-Shenzhen) & University of Münster
Stochastic AnalysisNumerical AnalysisApplied and Computational MathematicsPDEs
AN

Alexey Naumov

Professor, HSE University
probability theorystatisticsmachine learningrandom matrices