https://arxiv.org/api/YntILsauGi6pZqH4vviyRiHa7oo2026-09-11T20:01:07Z105723015http://arxiv.org/abs/2502.16015v2On the computation of the cumulative distribution function of the Normal Inverse Gaussian distribution2026-09-06T20:21:30ZIn this paper, we obtain various series and asymptotic expansions involving the modified Bessel function of the second kind for the normal inverse Gaussian cumulative distribution function. The new expansions accelerate computations, complementing the numerical integration methods implemented in statistical software packages. We also provide a detailed description of the algorithm and its corresponding implementation in C++. The performance and accuracy of the algorithm are extensively tested and benchmarked with open-source implementations, offering superior accuracy and speed-ups of a factor from 5 to 60.2025-02-22T00:12:53ZGuillermo Navas-Palenciahttp://arxiv.org/abs/2609.01937v2nethist: An R package for Nonparametric Graphon Estimation via Network Histograms2026-09-06T18:22:54ZUnderstanding the generative mechanism of real-world networks is crucial for analyzing connection patterns and making inference from network data. Graphons are widely used to model such mechanisms. Network histogram methods are nonparametric approaches based on blockmodel approximations that provide an intuitive view of network connection structures. However, there is a lack of software packages that construct network histograms. We introduce the R package nethist, which implements network histogram-type graphon estimators within a unified interface. This package is applicable to both single-layer and multilayer networks, and includes graphical summaries for examining both local and global network structures. By providing comprehensive network analysis tools, nethist facilitates understanding of complex systems of interrelated vertices.2026-09-01T23:03:26ZYoungseok SongSofia C. Olhedehttp://arxiv.org/abs/2609.06727v1Phase transitions and approximations of mean squared error for state-space models with fractional differencing2026-09-06T16:57:33ZWe study trend estimation in state-space models in which the trend has a fractional stochastic difference of order $d>0$ and the observation errors form a short-range-dependent stationary process. Using finite-sequence fractional summation and differencing operators, we analyze the penalized least-squares estimator obtained by shrinking the fractional differences of the trend. We derive asymptotic mean squared error (MSE) approximations for all $d>0$ and identify a sharp phase transition at $d=1/2$. When $d>1/2$, the estimator is consistent and its optimally balanced MSE has order $n^{-(2d-1)/(2d)}$. At the boundary $d=1/2$, we obtain a refined finite-sample approximation and show that the MSE decreases at the slower order $\log\log n/\log n$. When $0<d<1/2$, the MSE converges to an explicit positive limit, so consistent recovery of the trend is impossible under the considered scaling. We also describe a practical criterion for choosing the penalty parameter and differencing order, and numerical experiments illustrate the MSE approximations and the behavior of the selection procedure.2026-09-06T16:57:33Z25 Page2, 2 FiguresPrabir BurmanXiucai DingRobert H. Shumwayhttp://arxiv.org/abs/2609.06413v1Shrinkage invalidates the Hosmer-Lemeshow test: goodness of fit for penalized logistic regression, with an application to glaucoma diagnosis2026-09-06T06:13:48ZClinical prediction models are increasingly fitted by penalized logistic regression, because collinearity or many candidate predictors makes maximum likelihood unstable or impossible. Calibration is then almost always assessed by a grouped goodness-of-fit test such as the Hosmer-Lemeshow test. We show that this combination is invalid. Under ridge regression the grouped standardized residuals acquire a non-centrality induced by shrinkage, so the reference distribution used in practice is wrong, and at the penalty that most improves the fitted probabilities the test rejects correctly specified models between 92 and 100 per cent of the time. We derive the corrected law and define the shrinkage-corrected Hosmer-Lemeshow test, which subtracts an estimate of that non-centrality, restoring the maximum likelihood reference exactly to first order, and is made valid by prepivoting at a power cost we measure. We also give the attenuation law governing what any such test can detect once the linear predictor must be estimated. In glaucoma diagnosis by confocal laser tomography, where the maximum likelihood estimate does not exist, the corrected test finds the evidence for misfit weaker by more than three orders of magnitude: the fitted risks are too flat rather than mis-ordered, so the model needs recalibration rather than rebuilding.2026-09-06T06:13:48Z18 pages, 4 figures, 5 tables; supplementary material (13 pages) included as an ancillary file. Submitted to the Journal of the Royal Statistical Society Series C. Reproduction archive: https://doi.org/10.5281/zenodo.21903202. Methods implemented in the R package ebrahim.gof (CRAN), function shrink.gof()Ebrahim Khaled Ebrahimhttp://arxiv.org/abs/2602.06301v2Design-Conditional Prior Elicitation for Dirichlet Process Mixtures: A Unified Framework for Cluster Counts and Weight Control2026-09-05T20:17:47ZDirichlet process mixtures describe variation in effects or latent traits across sites, studies, or examinees. The concentration parameter controls the number and relative sizes of the clusters, but its hyperprior is often chosen by default. We describe a method for choosing this hyperprior when the number of units is fixed by design. Analysts specify an expected number of clusters and their uncertainty about that count. We translate these judgments into a hyperprior and examine what it implies about the sizes of the clusters. When the count and size judgments conflict, the Dual-Anchor procedure chooses and reports the trade-off. Simulations show that the default hyperprior can favor a single cluster when estimates for individual units are noisy. Applications to a multisite trial, a meta-analysis, and a vocabulary test show sensitivity in heterogeneity summaries despite relatively small changes in individual estimates. The DPprior R package implements the calibration and diagnostics.2026-02-06T01:49:16ZJoonHo Leehttp://arxiv.org/abs/2512.12749v3Residual-augmented flow matching operators for probabilistic partial differential equations2026-09-05T20:12:47ZLearning surrogate models for physical systems with latent uncertainty remains challenging in data-scarce regimes: deterministic neural operators fail to characterize uncertainty, while generative approaches require large ensembles of high-fidelity solution operator simulations and often sacrifice resolution generalizability. In this work, we propose a residual-augmented probabilistic operator learning framework that casts flow-matching-based generative modeling in infinite-dimensional function spaces while leveraging inexpensive low-fidelity solution operators as an inductive bias. Rather than learning the full high-fidelity stochastic solution operator directly, the proposed framework learns probabilistic residual operators that characterize the discrepancy between low- and high-fidelity solutions. By parameterizing the vector field in flow matching using neural operators conditioned on both the known system input and low-fidelity solution, the framework amortizes probabilistic inference across input conditions while enabling uncertainty-aware and resolution-generalizable predictions across spatial discretizations. Numerical experiments on stochastic advection, Burgers', and Darcy flow systems demonstrate that the residual-augmented formulation improves predictive accuracy under the same high-fidelity data budget, while the probabilistic operator learning formulation enables accurate characterization of uncertainty in low-data regimes compared to learning high-fidelity stochastic operators directly from data.2025-12-14T16:06:10ZSahil BholaKarthik Duraisamyhttp://arxiv.org/abs/2404.18556v3Doubly Adaptive Importance Sampling2026-09-05T07:37:50ZWe propose an adaptive importance sampling scheme for Gaussian approximations of intractable posteriors. Optimization-based approximations like variational inference can be too inaccurate while existing Monte Carlo methods can be too slow. Therefore, we propose a hybrid where, at each iteration, the Monte Carlo effective sample size can be guaranteed at a fixed computational cost by interpolating between natural-gradient variational inference and importance sampling. The amount of damping in the updates adapts to the posterior and guarantees the effective sample size. Gaussianity enables the use of Stein's lemma to obtain gradient-based optimization in the highly damped variational inference regime and a reduction of Monte Carlo error for undamped adaptive importance sampling. The result is a generic, embarrassingly parallel and adaptive posterior approximation method. Numerical studies on simulated and real data show its competitiveness with other, less general methods.2024-04-29T09:56:32Z37 pages, 13 figuresJournal of Computational and Graphical Statistics 35 (2026) 330-338Willem van den BoomAndrea CremaschiAlexandre H. Thiery10.1080/10618600.2025.2530048http://arxiv.org/abs/2501.00837v2A Generalized Fiducial Framework for Partially Identified Treatment Effects2026-09-05T02:25:03ZOver the past two decades, there has been renewed interest in fiducial inference, a statistical framework originally proposed by R. A. Fisher in the 1930s. Existing contributions, however, have largely focused on point-identified models, where the parameter vector is uniquely determined by the joint distribution of the observables and the model assumptions. This paper develops a unified fiducial inference procedure for partially identified models, thereby extending the applicability of the framework to settings in which the identified set is non-degenerate. The acceptance rate of the proposed sampler naturally serves as a diagnostic: high values support the identifying assumptions, while near-zero values signal their potential violation. We establish a Bernstein-von Mises theorem for the fiducial distribution, which provides theoretical guarantees for the proposed sampler. As two leading examples, we provide uncertainty quantification for the instrumental variable model and the mediation model under a range of causal assumptions and target parameters. The proposed methodology is illustrated through extensive simulations and empirical applications.2025-01-01T13:38:42ZYifan CuiJan Hannighttp://arxiv.org/abs/2609.05772v1Flexible Transformations for Bayesian Score Calibration2026-09-04T23:26:29ZModern statistical models are growing increasingly complex in an effort to realistically capture system dynamics. Using standard simulation-based inference, these models may be computationally prohibitive, necessitating the use of model calibration methods. Bayesian score calibration is a computationally efficient framework for model calibration with strong theoretical guarantees. This framework learns an appropriate correction for an approximate model using a small number of simulations from the data-generating process. Currently, only a location-scale transformation has been explored, which may lack the flexibility to correct the complex error introduced by some approximate models. In this paper, we develop two flexible transformations for use in the Bayesian score calibration framework. The first is a polynomial extension, which can appropriately adjust approximate models with location-varying error. The second is a sequential application of Bayesian score calibration, which can accommodate approximate models with posteriors that have low support for the true parameter values. We also discuss an additional diagnostic for use with this framework. We demonstrate the increased flexibility these two approaches provide over Bayesian score calibration in two illustrative simulation studies.2026-09-04T23:26:29Z35 pages (main body 22, appendixes 10, reference 3), 16 figures, and 9 tablesAdam BrethertonJoshua J. BonDavid J. WarneDavid J. NottChristopher Drovandihttp://arxiv.org/abs/2609.05713v1Approximating Bayesian leave-one-group-out cross-validation2026-09-04T20:44:01ZWhen data are grouped, hierarchical or multilevel models are commonly used to account for group-level variation with group-specific parameters. Leave-one-group-out cross-validation (LOGO-CV) is a suitable tool for evaluating predictive performance for new groups, providing an estimator of the expected log predictive density (elpd). Brute-force LOGO-CV requires one model refit per held-out group, often using computationally expensive inference algorithms such as MCMC. This is costly, particularly for large numbers of groups or complex model structures. Commonly used importance sampling approximations, intended to reduce this cost, tend to fail because the group-specific parameters of the held-out group must be integrated out. We identify two key challenges in LOGO-CV elpd estimation: approximating the LOGO posterior and computing the grouped marginal likelihood. We compare 11 strategies, including 5 newly proposed, to address them. Among others, we combine Pareto-smoothed importance sampling or adaptive importance sampling with integration techniques such as Laplace approximation, adaptive Gauss-Hermite quadrature, and bridge sampling. We evaluate these strategies in both simulation experiments and real-world case studies, which show that marginalising over the group-specific parameters substantially improves the reliability of the importance sampling approaches.2026-09-04T20:44:01Z35 (+18 supplementary) pages (excluding references); 13 (+13 supplementary) figuresAnna Elisabeth RihaSvenja JedhoffPaul-Christian BürknerAki Vehtarihttp://arxiv.org/abs/2609.05683v1Nonparametric Hypothesis Testing of High-dimensional Clustering With Application to Single-cell RNA Data2026-09-04T19:36:26ZSingle-cell RNA sequencing studies routinely use clustering to define putative cell types and cell states, yet the observed separation may arise from sampling variability rather than genuine biological heterogeneity. This paper studies formal significance testing of such clustering structure in high-dimensional data. Existing SigClust methods assess clustering significance through Monte Carlo simulation under a Gaussian single-cluster null, but this assumption can be unreliable for normalized gene expression data and other non-Gaussian settings. We propose SigClust-LCP, a nonparametric extension that models a single cluster by a log-concave distribution. To make this approach computationally feasible in moderate to high dimensions, we develop a score-matching estimator for log-concave projection inspired by recent generative modeling ideas. We establish theoretical guarantees for the estimator and for its use in clustering significance testing. Simulations show that SigClust-LCP controls Type-I error more reliably than existing methods across a range of unimodal and mixture distributions while retaining competitive power. In a single-cell RNA sequencing analysis of Hydra cells, the method avoids spurious subclusters within annotated cell populations and supports biologically meaningful separation across lineages and body-axis regions.2026-09-04T19:36:26Z29 pages, 4 figures and 3 tablesYifan DaiDi WuYufeng Liuhttp://arxiv.org/abs/2406.06231v4Statistical Inference for Privatized Data with Unknown Sample Size2026-09-04T13:44:19ZWe develop both theory and algorithms to analyze privatized data in unbounded differential privacy (DP), where even the sample size is considered a sensitive quantity that requires privacy protection. We show that the distance between the sampling distributions under unbounded DP and bounded DP goes to zero as the sample size $n$ goes to infinity, provided that the noise used to privatize $n$ is at an appropriate rate; we also establish that Approximate Bayesian Computation (ABC)-type posterior distributions converge under similar assumptions. We further give asymptotic results in regimes where the privacy budgets vary, establishing similarity of sampling distributions as well as showing that the MLE in the unbounded setting converges to the bounded-DP MLE. To facilitate valid, finite-sample Bayesian inference on privatized data under unbounded DP, we propose a reversible jump MCMC algorithm which extends the data augmentation MCMC of Ju et al. (2022). We also propose a Monte Carlo EM algorithm to compute the MLE from privatized data in both bounded and unbounded DP. We apply our methodology to analyze a linear regression model as well as a 2019 American Time Use Survey Microdata File which we model using a Dirichlet distribution.2024-06-10T13:03:20Z19 pages before references, 46 pages in total, 4 figures, 5 tablesJordan AwanAndres Felipe BarrientosNianqiao Juhttp://arxiv.org/abs/2609.05096v1Post-Corrected Raw-Score Martingale Posterior Sampling for von Mises-Fisher Models2026-09-04T12:50:13ZWe develop a finite-horizon calibration method for raw-score martingale posteriors, with von Mises--Fisher models as the main worked example. Starting from the maximum likelihood estimator, predictive paths are generated by simulating future observations from the current fitted model and updating the natural parameter by unpreconditioned score increments. The main methodological step is to separate predictive simulation from covariance calibration. Raw-score increments have Fisher-information covariance, whereas Bernstein--von Mises calibration requires inverse-information covariance. We therefore apply a terminal linear correction based on a local information estimate. For more efficient implementation, we also introduce a hybrid version that replaces the omitted tail of the infinite predictive continuation by a Gaussian approximation with matching leading-order quadratic variation. We prove fixed-n convergence, a finite-horizon approximation bound for the Gaussian tail, and a Bernstein--von Mises limit for the hybrid post-corrected sampler under local regularity and consistent terminal calibration. Simulations show that tail correction reduces truncation-induced underdispersion, and an OSCAR ocean-current example illustrates local directional uncertainty summaries.2026-09-04T12:50:13ZYi XuXinye ChenSheng Jianghttp://arxiv.org/abs/2408.17346v4On Nonparanormal Likelihoods2026-09-04T09:40:50ZNonparanormal models describe the joint distribution of multivariate responses via latent Gaussian, and thus parametric, copulae while allowing flexible nonparametric marginals. Some aspects of such distributions, for example conditional independence, are formulated parametrically. Other features, such as marginal distributions, can be formulated non- or semiparametrically. Such models are attractive when multivariate normality is questionable but interpretability paramount.
Most estimation procedures perform two steps, first estimating the nonparametric part. The copula parameters come second, treating the marginal estimates as known. This is sufficient for some applications. For other applications, e.g. when a semiparametric margin features parameters of interest or when standard errors are important, a simultaneous estimation of all parameters might be more advantageous.
We present suitable parameterisations of nonparanormal models, possibly including semiparametric effects, and define four novel nonparanormal log-likelihood functions. In general, the corresponding one-step optimisation problems are shown to be non-convex. In some cases, however, biconvex problems emerge. Several convex approximations are discussed. From a low-level computational point of view, the core contribution is the score function for multivariate normal log-probabilities computed via Genz procedure.
As a demonstration for the versatility of the theoretical and computational framework, we present a series of nonparanormal models for transformation discriminant analysis when some biomarkers are subject to limit-of-detection problems. Possible empirical gains of full maximum likelihood estimation compared to two-step approaches are illustrated in a simulation study targeting semiparametric efficient polychoric correlation analysis where a theoretical benchmark is available.2024-08-30T15:17:53ZTorsten Hothornhttp://arxiv.org/abs/2607.11631v2Markov Chain Monte Carlo with Diffusion Paths2026-09-03T17:36:05ZSampling from multimodal distributions is a longstanding challenge for classical local Markov chain Monte Carlo (MCMC) methods. A popular remedy is to introduce a sequence of intermediate distributions that interpolate between the target and a simpler reference. The classical choice, tempering, raises the density to a power, but distorts the relative weights of asymmetric modes and can lead to poor mixing. We instead propose interpolating along the diffusion path, the marginals of a noising diffusion process that carries the target toward a Gaussian. This path preserves the relative weights of the modes and enjoys favorable mixing properties, which we make precise through a spectral-gap analysis of the corresponding ideal transition kernel. Sampling along the path requires its intermediate scores, which can be estimated from the unnormalized target through variational approaches, yielding only an approximate sampler. To remove the resulting bias, we introduce the Metropolis-adjusted diffusion path (MAD-Path) sampler, which corrects the diffusion-path proposal in an augmented path space and leaves the target invariant regardless of the accuracy of the learned score or the discretization error. We further quantify how these two errors affect the acceptance probability, providing guidance for practical tuning. Experiments on a range of Bayesian posteriors show that MAD-Path improves global exploration and mode-weight estimation relative to tempering-based MCMC methods and unadjusted diffusion samplers.2026-07-13T14:48:19ZUpdate Section 5Han ChenSifan LiuJun Yang