https://arxiv.org/api/sFqOKLqckth3Lh7e4j8MluoML2s 2026-10-02T19:34:09Z 10675 240 15 http://arxiv.org/abs/2503.24311v2 Selective Inference in Graphical Models via Maximum Likelihood 2026-08-25T13:29:09Z The graphical lasso is a widely used algorithm for fitting undirected Gaussian graphical models. However, for inference on functionals of edge values in the learned graph, standard tools lack formal statistical guarantees, such as control of the type I error rate. We introduce a selective inference method for asymptotically valid inference after graphical lasso selection with added randomization. We obtain a selective likelihood, conditional on the event of selection, through a change of variable on the known density of the randomization variables. Our method enables interval estimation and hypothesis testing for a wide range of functionals of edge values in the learned graph using the conditional maximum likelihood estimator. Our numerical studies show that introducing a small amount of randomization: (i) achieves better results than data splitting in terms of model selection quality; (ii) ensures intervals of bounded length also in high-dimensional settings where data splitting suffers from low power or does not reach asymptotic regime due to limited samples for inference; (iii) enables inference for a wide range of inferential targets in the learned graph, including partial correlations and measures of node influence and connectivity between nodes. 2025-03-31T16:57:04Z Sofia Guglielmini Gerda Claeskens Snigdha Panigrahi http://arxiv.org/abs/2603.04003v2 Efficient Bayesian Estimation of Dynamic Structural Equation Models via State Space Marginalization 2026-08-25T12:00:04Z Dynamic structural equation models (DSEMs) combine time-series modeling of within-person processes with hierarchical modeling of between-person differences and differences between timepoints, and have become very popular for the analysis of intensive longitudinal data in the social sciences. An important computational bottleneck has, however, still not been resolved: whenever the underlying process is assumed to be latent and measured by one or more indicators per timepoint, currently published algorithms rely on inefficient brute-force Markov chain Monte Carlo sampling which scales poorly as the number of timepoints and participants increases and results in highly correlated samples. The main result of this paper shows that the within-level part of any DSEM can be reformulated as a linear Gaussian state space model. Consequently, the latent states can be analytically marginalized using a Kalman filter, allowing for highly efficient estimation via Hamiltonian Monte Carlo. This makes estimation of DSEMs computationally tractable for much larger datasets -- both in terms of timepoints and participants -- than what has been previously possible. We demonstrate the proposed algorithm in several simulation experiments, showing it can be orders of magnitude more efficient than standard Metropolis-within-Gibbs approaches. 2026-03-04T12:47:12Z Øystein Sørensen http://arxiv.org/abs/2608.15013v2 Structure-Preserving Visualization of Complex Systems through Discrete Approximation: An Application to Argo Data 2026-08-25T10:21:33Z This paper presents a framework for constructing structure-preserving representations of complex systems through discrete approximation, and demonstrates its use in studying the vertical temperature and salinity structures in the mesopelagic zone across the global ocean using the ARGO dataset. Clustering serves as a means of organizing complexity into a finite set of structures that approximate the overall oceanic conditions, and a color encoding design then integrates these structures into a coherent map, with the three color components derived from interpretable geometric features of a profile: its initial level, its magnitude of variation, and its shape. Instead of focusing on specific depth levels or computing zonal averages within selected regions, our approach preserves the full vertical structure of individual profiles and incorporates each profile in the global ocean, capturing both fine-scale profile detail and large-scale spatial variability. By clustering over one million profiles collected over a decade, we identify and characterize representative profile shapes, which form the basis for a visualization strategy that provides an integrated, comprehensive, and interpretable presentation of the large-scale spatial distributions of these oceanic vertical patterns. 2026-08-15T03:50:19Z Shang-Ying Shiu Fushing Hsieh Ting-Li Chen http://arxiv.org/abs/1007.0842v5 Erratum to "Higher order scrambled digital nets achieve the optimal rate of the root mean square error for smooth integrands" 2026-08-25T08:34:01Z This note corrects several statements and proof steps in J.~Dick, \emph{Annals of Statistics} \textbf{39} (2011), 1372--1398, DOI: 10.1214/11-AOS880, arXiv:1007.0842 v4. The principal smooth-function result remains valid: an order-$d$ nested-uniformly scrambled digital $(t,m,ds)$-net has root mean square integration error $O(N^{-\min(α,d)-1/2+\varepsilon})$ for every $\varepsilon>0$ for functions in the unanchored Sobolev space with square-integrable mixed partial derivatives up to order $α$ in each variable. The finite-difference variation introduced in the paper neither coincides with the displayed Sobolev norm nor directly controls the finite differences used in Appendix~B; the claimed theorem for that variation is therefore withdrawn here. In addition, the variance bound in Theorem~10 is missing a square, and the proof for $d>α$ does not prove the logarithmic power printed in the theorem. We give a direct Sobolev proof, and a corrected logarithmic factor. 2010-07-06T09:30:11Z Josef Dick http://arxiv.org/abs/2501.05458v3 Generative Modeling: A Review 2026-08-24T23:40:33Z We organize the generative-modeling literature around three classes of generators, corresponding to three distinct inferential tasks: estimating counterfactual outcome distributions in causal inference, recovering posteriors from simulated parameter--outcome pairs, and forming predictive outcome distributions. The unifying representation relies on the noise outsourcing theorem of Kallenberg, which expresses a conditional distribution as a deterministic function of its inputs and an independent noise variable. Within this organization we develop generative Bayesian computation, a method in the parameter--outcome class: a quantile neural network, trained on simulated pairs under the pinball loss, that targets the posterior of the parameter directly, without invertible architectures or density evaluation, and that serves equally as a predictive generator once the roles of parameter and outcome are exchanged. We illustrate the framework on an agent-based Ebola transmission application, where generative Bayesian computation recovers accurate posteriors at substantially lower cost than rejection-based simulation inference, while avoiding the density-evaluation and invertibility constraints of competing generators. 2024-12-24T21:27:48Z arXiv admin note: substantial text overlap with arXiv:2305.14972 Maria Nareklishvili Nick Polson Vadim Sokolov http://arxiv.org/abs/2606.15360v4 Matched generating elements in maximum entropy density reconstruction 2026-08-24T15:15:49Z Moment-constrained maximum entropy (MaxEnt) reconstructs a density as an exponential family whose sufficient statistics are chosen constraint functions. We study this choice separately from the numerical coordinates used to solve the dual problem. An odd-only family on a symmetric support forces f(x)f(-x) to be constant, so parity matching is necessary. A logarithmic-rational constraint produces the Student/Cauchy family and remains population-defined when required integer moments do not exist; a parity-matched one-parameter path provides a reproducible fractional-exponent alternative containing the monomials. The numerical study uses a common damped-Newton solver and audits the six-distribution benchmark of Rajan et al. Stable parity-block divided differences, residual replay in the original constraint coordinates, and independent adaptive integration identify a raw-coordinate false convergence on D5. Under the corrected computation, the selected PATP arm is within 1.22 times the oracle on D5 but retains selected-to-oracle error ratios of 30.7, 8.88, and 3.33 on D2, D4, and D6 at n=12. These checks preserve the negative selector conclusion without treating a numerical artifact as statistical evidence. The Pearson row is reproduced, whereas the published GOPoly rows are not under the printed protocol. Matched LogRat results show family capacity; neither tested selector supports automatic choice among constraint families. 2026-06-13T15:45:18Z Numerical correction to the PATP arm: stable parity-block divided-difference coordinates with original-constraint replay replace a raw-power quadrature false convergence. D5 is resolved, while selected-to-oracle gaps remain on D2, D4, and D6, so automatic selection remains unresolved. 26 pages, three figures, eight tables. Code: https://github.com/SZabolotnii/Ku-MaxEnt-code-supplement Serhii Zabolotnii http://arxiv.org/abs/2410.01223v27 Statistical Taylor Expansion: A New and Path-Independent Method for Uncertainty Analysis 2026-08-24T13:54:42Z Statistical Taylor expansion is a rigorous extension of conventional Taylor expansion that replaces each precise input variable with a random variable of known distribution and sample count, then computes the mean, deviation, and a bounding reliability of every result. By tracking the propagation of input uncertainties through all intermediate steps, it renders the final result path-independent, with precise quantification of the tracking quality. This path-independence sets it fundamentally apart from conventional numerical approaches, which are path-dependent. This study presents an implementation called variance arithmetic and demonstrates its performance across diverse mathematical applications. This study also reveals the potentially substantial impact of numerical errors in library functions, the defect of applying input uncertainties as weights in conventional regression, and the modeling error of the discrete Fourier transformation. The concept of statistical algebra is also introduced. 2024-10-02T04:02:21Z 69 pages, 67 figures Chengpu Wang http://arxiv.org/abs/2606.22622v2 A New Algorithm for Totally Positive Approximations 2026-08-24T13:03:53Z We revisit the problem of approximating a bivariate distribution with finite support by another such distribution which is totally positive of order two (TP2). Approximation is meant in a maximum likelihood sense. 2026-06-21T18:07:19Z Lutz Duembgen Philip Stange http://arxiv.org/abs/2608.22921v1 fdWasserstein: Optimal Transport Methods for Covariance Operators of Functional Data 2026-08-24T07:57:00Z Data increasingly arrive as collections of curves - a voice recording, a growth trajectory, a day of sensor readings - where each observation is a whole function rather than a single number. The usual question asked of such data is how the average curve differs from one group to the next. But the average is only half the picture: two populations of curves can share almost the same mean and still differ profoundly in how they fluctuate around it, and it is often this variability - the pattern of covariation within a curve - that carries the scientific signal. Comparing populations at this level means comparing their covariance operators, and statistics on covariances is impaired by their non-linearity. The fdWasserstein package equips R users with functions to make such comparisons. It is centered on the geometry of optimal transport, under which covariance operators can be meaningfully averaged, contrasted, and interpolated. It provides the Procrustes-Wasserstein distance between covariance operators, their Frechet mean (barycenter), an ANOVA-type permutation test for the equality of several covariances, principal component analysis of covariance variation, and an entropy-regularized soft clustering of curves by their covariance structure. We outline the underlying ideas, discuss the implementation and demonstrate the complete workflow on the phoneme data shipped with the package. 2026-08-24T07:57:00Z V. Masarotto http://arxiv.org/abs/2608.22843v1 Favourable Missingness in Semi-Supervised Classification for Exponential Mixture Models 2026-08-24T06:21:13Z Semi-supervised classifiers are commonly trained from samples in which all features are observed but some class labels are missing. When label missingness is independent of the observed data, unavailable class memberships reduce Fisher information relative to a completely classified sample. We study a different regime in which the probability of label missingness depends on posterior classification uncertainty, so that the observed missing-label indicators can themselves carry information about the Bayes decision boundary. Building on the conditionally weighted information decomposition of Ahfock and McLachlan, we develop this phenomenon for a two-component exponential mixture. Although the exponential model is non-Gaussian, asymmetric, and supported on the positive half-line, its log-posterior odds remain linear in the feature. We derive Bayes' rule and its exact error rate, formulate entropy-logistic and squared-discriminant missingness mechanisms, and obtain the full partially classified likelihood. We then derive a decomposition of the Fisher information into the complete-data information, the conditionally weighted loss due to missing labels, and the information contributed by the missing labels. Numerical quadrature identifies regions in which the full likelihood classifier has asymptotic relative efficiency above or below one. Monte Carlo experiments with finite training samples broadly support the population calculations, with the largest departures from the asymptotic predictions occurring near the transition at which the relative efficiency crosses one. 2026-08-24T06:21:13Z 26, 3 figures Huanchao Zhou Jinran Wu Fariborz Setoudehtazang Geoffrey J. McLachlan http://arxiv.org/abs/2609.27969v1 voigtinference: Exact likelihood calculus and conditional attribution for the Voigt profile 2026-08-24T01:27:38Z Fast, accurate algorithms for the Faddeeva function w(z), and hence the Voigt profile K(x,a), have existed for four decades, and analytic first derivatives are available in some implementations. What the established libraries reviewed here have not provided is the full likelihood calculus of the normalized Voigt distribution: applications still commonly resort to pseudo-Voigt approximations, finite-difference derivatives, or numerical convolution for parameter inference. Because w'(z) = -2z w(z) + 2i/sqrt(pi), every derivative of the Voigt log-likelihood is an algebraic function of K and the dispersion part L(x,a) = Im w(z), from the single complex evaluation that delivers the profile. This yields the score and Hessian in closed form, and the expected Fisher information by one-dimensional quadrature of an analytic integrand; for fixed interior widths sigma, gamma > 0, the MLE of the center and both widths is consistent and asymptotically normal at rate sqrt(n), despite the distribution having no finite mean or variance, so conventional likelihood-based standard errors apply. The conditional mean of the Gaussian component given an observation is (y - mu) - gamma L/K: a redescending function that attributes moderate deviations to the Gaussian (Doppler/resolution) component and extreme ones to the Lorentzian tail. The package voigtinference (Python, NumPy/SciPy, with a cross-validated Julia companion) supplies the toolkit: score, full parameter Hessian, expected information, Newton-based unbinned maximum likelihood with boundary diagnostics, conditional component moments, and evaluation validated against high-precision references at extreme width ratios. It applies directly to unbinned non-relativistic, constant-width Breit-Wigner x Gaussian resonance fits and supplies analytic Jacobians for line-shape refinement. Companion paper: arXiv:2605.01665. 2026-08-24T01:27:38Z 17 pages, 3 figures Peter Reinhard Hansen Chen Tong http://arxiv.org/abs/2608.22553v1 Non-asymptotic Analysis of Matérn Regression: The Roles of Target and Kernel Lengthscales 2026-08-23T18:57:49Z Theoretical guarantees for kernel regression are typically formulated in terms of smoothness, but practical accuracy depends critically on how the design resolution compares with the target and kernel lengthscales. We develop a finite-sample theory for Matérn regression on periodic domains with quasi-uniform designs, covering noiseless interpolation and noisy kernel ridge regression. Our minimax result shows that accurate recovery requires a design dense enough to resolve the target lengthscale and sufficient information at that scale to overcome noise. We prove that Matérn interpolation additionally requires the design to resolve the kernel lengthscale: if the kernel is too short relative to the point spacing, an additional error remains even when the target is well resolved. For noisy Matérn regression, we derive a three-term fixed-ridge risk characterization consisting of target-scale bias, kernel-scale bias, and variance, and show that optimizing over the ridge parameter yields four distinct contributions. When the target is no more than twice as smooth as the kernel, choosing a kernel lengthscale longer than the target does not worsen the oracle risk, whereas choosing one too short can. Thus these resolution conditions determine when accurate recovery becomes possible, while smoothness determines how rapidly the error decreases thereafter. 2026-08-23T18:57:49Z Daniel Sanz-Alonso http://arxiv.org/abs/2608.16573v2 Bessel-Debiased Pseudo-Marginal MCMC for Generalised Bayesian Inference 2026-08-23T16:05:59Z Generalised Bayesian inference uses weights of the form $\exp\{-β_n\ell_n(θ)\}$ even when the loss is only estimated. Exponentiating an unbiased loss estimate changes the target, and when $β_n\asymp n$ an ordinary Monte Carlo loss estimate with variance of order $M^{-1}$ requires a per-proposal budget $M$ of order $n^2$ to keep the leading log-weight variance bounded. We introduce Sign-Corrected Bessel Debiasing (SCBD), a signed pseudo-marginal method based on independent block estimates of the loss, and study its ordinary-MC and independently randomised quasi-Monte Carlo (RQMC) implementations. Under an i.i.d. Gaussian block model, a Bessel factor constructed from the block sample variance exactly removes the Gaussian exponential inflation despite the variance being unknown. For general non-Gaussian finite blocks, the method targets a posterior differing from the intended posterior by a parameter-dependent multiplicative factor. Under regularity conditions, the uncorrected and corrected ordinary-MC targets have total-variation errors of orders $β_n^2/M_n$ and $β_n^3/M_n^2$. If an RQMC block estimator has variance $\mathcal O\{B^{-α}(\log B)^{d-1}\}$, the corresponding errors are of orders $δ^{\mathrm{RQ}}_{n,M_n}$ and $(δ^{\mathrm{RQ}}_{n,M_n})^{3/2}$, where $δ^{\mathrm{RQ}}_{n,M_n}=β_n^2M_n^{-α}{\log(2+M_n)}^{d-1}$. The same variance rate gives a sufficient budget of order $n^{2/α}$, up to logarithmic factors, for bounded leading log-weight variance when $β_n\asymp n$. The numerical examples show that favourable RQMC representations can inherit this budget scaling and that variance reduction and Bessel correction are complementary. Compared to existing exact corrections, Bessel debiasing is essentially "for free". It is generic, easy to code and supported by theory. 2026-08-17T13:36:50Z 55 pages, 11 figures Yingkai Lu Jeong Eun Lee Geoff K. Nicholls http://arxiv.org/abs/2608.22254v1 Analysis of Nonnegative Observations using Gamma Model with 2 Factors (ANOGaM-2): Theory, Method and Applications with Real-life Data (including R code) 2026-08-23T07:32:39Z Two-factor ANOVA is widely used in experimental studies but relies on additivity, normality, independence, and homoscedasticity. These assumptions are often violated for nonnegative, positively skewed observations. Although Box--Cox-type transformations are commonly used, they may reduce interpretability and require a subjective choice of transformation. We propose an alternative framework in which nonnegative observations affected by two factors are modeled by gamma distributions with unknown shape and scale parameters that may depend on factor levels. We develop likelihood ratio tests (LRTs) for main and interaction effects. The asymptotic LRT (ALRT) uses the asymptotic chi-square distribution, which may be inaccurate for small to moderate samples. We therefore propose a parametric bootstrap LRT (PBLRT) that determines critical values by simulation. Extensive simulations show that the PBLRT maintains the nominal significance level well. Real-data examples demonstrate its applicability and show that its inferences can differ from those of traditional ANOVA. 2026-08-23T07:32:39Z 58 pages Buu-Chau Truong Nuong Thi Thuy Tran Nabendu Pal http://arxiv.org/abs/2206.08401v5 Is Decentralized Finance Actually Decentralized? An Interdisciplinary Framework Integrating Network Theory, Agent-Based Simulation, and Longitudinal Evidence from Aave, GHO Issuance, and Cross-Chain Expansion 2026-08-23T07:30:39Z Decentralized finance (DeFi) can broaden access while leaving activity, network position, and infrastructure concentrated. We develop a four-dimensional framework for participation, activity distribution, structural position, and infrastructure dependence, integrating network theory, theorem-consistent agent-based simulation, and longitudinal analysis of 1,956,216 Aave V3 Pool events. We study GHO issuance on Ethereum (15 July 2023) and its first cross-chain expansion to Aave's existing Arbitrum market (2 July 2024). Excluding each activation week, mean weekly active position-holder addresses increased by 91.0% around Ethereum issuance and 1.7% around Arbitrum expansion, while activity concentration fell by 31.5% on Ethereum but rose by 58.1% on Arbitrum. On a common 2024 calendar, the Arbitrum--Gnosis DiD-style change is +1.9833 for log participation and -0.01842 for position-holder-event HHI. Rule-based simulations recover the analytical equilibrium and show why aggregate growth can coexist with lower, unchanged, or higher concentration, while chain dispersion alone cannot establish route or shared-component resilience. Role-aware analysis further shows that network-structure conclusions vary by protocol action and scale. Intellectually, the framework explains why four dimensions of decentralization can diverge. Practically, it helps researchers, protocol designers, governance communities, and policymakers assess stablecoin growth without equating adoption with decentralization. 2022-06-16T18:26:38Z Ziqiao Ao Lin William Cong Gergely Horvath Luyao Zhang