Searcharxiv⌕ Search

arXiv · 2609.37939

Semiparametric Bernstein-von Mises theorems from Stein's method

Abstract

We introduce a novel proof strategy for semiparametric Bernstein-von Mises theorems based on Stein's method combined with an information-geometric framework. Rather than controlling the effect of the prior on the asymptotic marginal posterior distribution of the functional of interest through the stability of an integrated likelihood under a suitable perturbation, we characterise its influence through the prior-weighted divergence of suitable vector fields over the statistical model. Applying Stein's method in this setting instead of the usual Laplace-transform approach yields explicit non-asymptotic upper bounds on the bounded-Lipschitz distance between marginal posterior distributions and the corresponding Gaussian limits predicted by semiparametric efficiency theory. These bounds consist entirely of local scalar differential quantities associated with the functional, the prior and a chosen vector field-typically related to the efficient influence function-evaluated over posterior contraction sets. We apply the theory to quadratic functionals in Gaussian white-noise models and to linear, quadratic and general integral functionals in histogram density models, with both conjugate and non-conjugate priors. For linear and quadratic functionals, the same theory identifies the differential term responsible for posterior bias and allows us to remove it through a natural functional correction, yielding Bernstein-von Mises-type theorems in regimes where they may fail for the original uncorrected functional.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Paul Rosa. 2026-09-29. Semiparametric Bernstein-von Mises theorems from Stein's method. https://arxiv.org/abs/2609.37939

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

On the Nonasymptotic Scaling Guarantee of Hyperparameter Estimation in Inhomogeneous, Weakly-Dependent Complex Network Dynamical Systems

Hierarchical Bayesian models are increasingly used in large, inhomogeneous complex network dynamical systems by modeling parameters as draws from a hyperparameter-governed distribution. However, theoretical guarantees for these estimates as the network population size grows have been lacking. A critical concern is that hyperparameter estimation may diverge for larger networks, undermining the model's reliability. Formulating the system's evolution in a measure transport perspective, we propose a theoretical framework for estimating hyperparameters with mean-type observations, which are prevalent in many scientific applications. Our primary contribution is a nonasymptotic bound for the deviation of estimate of hyperparameters in inhomogeneous complex network dynamical systems with respect to network population size, which is established for a general family of optimization algorithms within a fixed observation duration. For systems with independent nodes, we first establish a fixed-accuracy probabilistic hyperparameter bound and then use it to derive an explicit nonasymptotic hyperparameter bound. We subsequently extend both bounds to the more challenging and practically relevant setting of systems with weakly-dependent nodes. We validate our theoretical findings with numerical experiments on two representative models: a Susceptible-Infected-Susceptible model and a Spiking Neuronal Network model. In both cases, the results confirm that the estimation error decreases as the network population size increases, aligning with our theoretical guarantees. This research proposes the foundational theory to ensure that hierarchical Bayesian methods are statistically consistent for large-scale inhomogeneous systems, filling a gap in this area of theoretical research and justifying their application in practice.

math.ST↗

Intrinsic-dimension empirical Bernstein inequalities for bounded self-adjoint operators

Operator-valued concentration inequalities are foundational to the analysis of modern high-dimensional statistics and randomized algorithms. However, standard oracle bounds are frequently limited in practice: they require explicit a priori knowledge of the true variance, and often explicitly scale with the ambient dimension, rendering them vacuous for infinite-dimensional or heavily structured operators. Motivated by these challenges, we establish the first empirical Bennett and Bernstein inequalities for sums of independent, bounded, self-adjoint Hilbert-Schmidt operators. Our fully data-driven bounds replace the unknown variance with an empirical estimate and rely strictly on the intrinsic dimension rather than the ambient dimension. This structural shift yields computable, dimension-free guarantees with a sharper first-order asymptotic radius for non-isotropic random matrices and seamlessly extends to infinite-dimensional Hilbert spaces. We demonstrate that our empirical bounds achieve asymptotic sharpness with the best known oracle rates. Finally, as an independent byproduct, we derive novel empirical concentration guarantees for the intrinsic dimension itself.

math.ST↗

Bentkus-type asymptotic e-values

Asymptotic e-values are emerging as a powerful alternative to asymptotic p-values, particularly in post-hoc inference and multiple testing, where significance levels may be data-dependent. Existing asymptotic e-values, however, suffer from the ``missing factor,'' a scaling inefficiency resulting in overly conservative inference. Drawing on the framework of near-optimal concentration inequalities developed by Bentkus in the 2000s, we introduce Bentkus-type asymptotic e-values and prove that they successfully eliminate the missing factor. We also demonstrate both theoretically and empirically that Bentkus-type e-values consistently deliver sharper inference than existing alternatives, leading to tighter post-hoc confidence intervals and higher rejection rates in multiple testing procedures.

math.ST↗