https://arxiv.org/api/VrGl02KuGRYV07rv+ig1VbA6fYU2026-09-11T20:45:12Z105724515http://arxiv.org/abs/2510.19722v2Co-SIVI: A Correlated Semi-Implicit Variational Approach for Spatial Models2026-09-03T15:04:15ZWe propose correlated semi-implicit variational inference (Co-SIVI), a scalable approach for full posterior approximation in large spatial models with exponential-family likelihoods. Co-SIVI incorporates dependence directly into the conditional variational distribution of spatial random effects through an iterative weighted least squares algorithm that accommodates both Gaussian process and nearest-neighbor Gaussian process (NNGP) priors. For large samples, we further propose reparameterizing the variational family for the covariance parameters to better capture posterior dependence. Co-SIVI addresses an important limitation of semi-implicit variational inference (SIVI), for which dependence induced through the mixing distribution may be insufficient when spatial random effects are explicitly included in the variational family. In simulations with Gaussian, Poisson, Gamma, and Bernoulli outcomes, Co-SIVI closely reproduces Hamiltonian Monte Carlo (HMC) results at substantially lower computational cost, while SIVI performs similarly for marginalized Gaussian models. We apply SIVI and Co-SIVI, both with NNGP priors, to two large-scale datasets: temperature data modeled with a Gaussian likelihood and housing price data modeled with a Gamma likelihood. Overall, Co-SIVI provides a scalable and flexible alternative to HMC for full posterior approximation in large non-Gaussian spatial models.2025-10-22T16:07:57Z31 pages without the appendix and 60 with the appendix, 5 figures without appendix and 21 figures with appendix, 2 tables, 2 algorithmsSébastien GarneauDepartment of Epidemiology, Biostatistics and Occupational Health, McGill UniversityCarlos T. P. ZaniniDepartment of Statistical Methods, Federal University of Rio de JaneiroAlexandra M. SchmidtDepartment of Epidemiology, Biostatistics and Occupational Health, McGill Universityhttp://arxiv.org/abs/2506.19554v5Modeling uncertainty in the covariance matrix for probabilistic forecast reconciliation2026-09-03T11:54:38ZIn minimum trace (MinT) forecast reconciliation, the covariance matrix of the base forecast errors plays a crucial role. Typically, this matrix is estimated and then treated as known. This can lead to an underestimation of the variance of the predictive distribution. To address the problem, we propose a Bayesian reconciliation model that accounts for uncertainty in the estimation of the covariance matrix. By adopting an Inverse-Wishart prior and assuming Gaussian residuals, the reconciled predictive distribution follows a multivariate t-distribution, obtained in closed form, rather than a multivariate Gaussian distribution. We evaluate our method on three tourism-related datasets, including a new publicly available dataset. Empirical results show that our approach consistently improves prediction intervals compared to MinT reconciliation.2025-06-24T12:05:24ZChiara CarraraDario AzzimontiGiorgio CoraniLorenzo Zambon10.1016/j.ijforecast.2026.07.003http://arxiv.org/abs/2602.19590v2Metaorder modelling and identification from public data2026-09-03T09:05:52ZMarket-order flow in financial markets exhibits long-range correlations. This is a widely known stylised fact of financial markets. A popular hypothesis for this stylised fact comes from the Lillo-Mike-Farmer (LMF) order-splitting theory. However, quantitative tests of this theory have historically relied on proprietary datasets with trader identifiers, limiting reproducibility and cross-market validation. We investigate whether it can be recovered from anonymous public data using synthetic metaorder reconstruction. Using transaction and quote data for the largest 239 stocks by market capitalisation on the JSE as of 13 March 2026 with the data range being 1 January 2023 until 31 December 2025, we conduct a grid search over reconstruction parameters and evaluate each configuration against established metaorder stylised facts and the LMF relation. Configurations selected to minimise errors across the metaorder impact stylised facts reproduce the targeted aggregate properties but yield a poor LMF relation. Configurations selected to minimise the LMF discrepancy recover the relation by construction while retaining several broad impact features, although some stock-level execution and decay fits are weaker. These asymmetric results show that recovering aggregate impact stylised facts alone is insufficient to identify LMF-consistent order splitting, while the LMF-targeted result establishes compatibility within the reconstruction class rather than an independent test. The findings support consistency with, rather than direct validation of, the LMF theory using anonymous market data.2026-02-23T08:28:46Z23 pages, 15 figures. Updated to refine the test results. Previously this version appeared as arXiv:2608.30999 which was submitted as a new work by accidentEzra GoliathTim Gebbiehttp://arxiv.org/abs/2603.00277v2CliPS -- How to identify cluster distributions in Bayesian mixture models2026-09-02T17:54:40ZWe propose the CliPS procedure when fitting Bayesian mixture models in the context of model-based clustering to identify the cluster distributions while simultaneously assessing the suitability of a cluster solution and validating the cluster structure. The procedure relies on the point process representation of a mixture model and is based on the assumption that a suitable cluster solution requires the clusters to be distinguishable with respect to a low-dimensional functional of the component-specific parameters of the mixture. CliPS maps the component-specific MCMC draws to the point process representation and identifies clusters there, exploiting that, while data distributions usually overlap, the posterior of these functionals are more and more separated for increasing sample size. We outline the procedure and illustrate its use on several model-based clustering examples.2026-02-27T19:50:45ZGertraud Malsiner-WalliSylvia Frühwirth-SchnatterBettina Grünhttp://arxiv.org/abs/2405.05073v2gasmodel: An R Package for Generalized Autoregressive Score Models2026-09-02T14:24:34ZGeneralized autoregressive score (GAS) models are a class of observation-driven time series models that employ the score to dynamically update time-varying parameters of the underlying probability distribution. GAS models have been extensively studied and numerous variants have been proposed in the literature to accommodate diverse data types and probability distributions. This paper introduces the gasmodel package, which has been designed to facilitate the estimation, forecasting, and simulation of a wide range of GAS models. The package provides a rich selection of distributions, offers flexible options for specifying dynamics, and allows to incorporate exogenous variables. Model estimation utilizes the maximum likelihood method.2024-05-08T14:14:59ZHolý, V. (2026). gasmodel: An R Package for Generalized Autoregressive Score Models. The R Journal, 18(1), 23-38Vladimír Holý10.32614/RJ-2026-002http://arxiv.org/abs/2609.02138v1HyperMC: Multi-Fidelity Hyperparameter Tuning for Stochastic Gradient MCMC2026-09-02T05:45:16ZStochastic gradient Markov chain Monte Carlo (SGMCMC) methods enable scalable Bayesian inference, but their performance depends strongly on hyperparameters such as the step size, mini-batch size, and number of leapfrog steps. Since most SGMCMC algorithms lack a Metropolis-Hastings acceptance rate, standard acceptance-based tuning methods are not directly applicable. We propose HyperMC, a multi-fidelity tuning framework that combines Hyperband-style resource allocation with kernel Stein discrepancy (KSD) evaluation. By running multiple successive-halving brackets, HyperMC balances broad exploration of a continuous hyperparameter space with increasingly accurate evaluation of promising configurations under a fixed computational budget. We further introduce Robust HyperMC, which uses global grid initialization followed by elite-guided local refinement to reduce sensitivity to random candidate generation and noisy finite-budget evaluations. Under suitable approximation and concentration conditions for the estimated KSD, we establish that the successive-halving component selects a near-optimal configuration among the sampled candidates with high probability and derive a sufficient computational budget for successful selection. Experiments on logistic regression, probabilistic matrix factorization, and Bayesian neural networks show that HyperMC improves posterior approximation or predictive calibration relative to MAMBA, grid search, and heuristic baselines, while Robust HyperMC yields more stable and reproducible tuning results.2026-09-02T05:45:16Z56 pages, 14 figures, 9 tablesMing TanXiyun Jiaohttp://arxiv.org/abs/2607.28413v2Windowed thinning and query complexity for the bouncy particle and Zigzag samplers2026-09-01T19:57:52ZLet $μ(d x)\propto e^{-U(x)} d x$ on $\R^d$, where $U$ is $m$-strongly convex and $L$-smooth, and denote by $κ=L/m$ the condition number. We consider windowed thinning, an exact simulation method for the bouncy particle sampler and the coordinate Zigzag process. The method divides a trajectory into deterministic windows and uses a gradient evaluation at the beginning of each window to construct a tractable local envelope for the event rate. Combining this construction with quantitative mixing estimates and finite-time bounds on the expected numbers of bounces and flips yields query complexity guarantees from a Gaussian cold start. For total-variation error $\varepsilon$, the expected query counts are $O(κ^{1/2}d\,(d\logκ+\log\frac1\varepsilon))$ gradient queries for the bouncy particle sampler and $O(κd^{1/4}(d\logκ+\log\frac1\varepsilon))$ full-gradient equivalents for Zigzag, where $d$ coordinate-partial queries count as one equivalent.2026-07-30T15:59:25Zv2: corrected a few typos and updated referencesJianfeng LuYinchen Luohttp://arxiv.org/abs/2504.02518v4Online Multivariate Regularized Distributional Regression for High-dimensional Probabilistic Electricity Price Forecasting2026-09-01T19:40:55ZProbabilistic electricity price forecasting (PEPF) is vital for short-term electricity markets, yet the multivariate nature of day-ahead prices - spanning 24 consecutive hours - remains underexplored. At the same time, real-time decision-making requires methods that are both accurate and fast. We introduce an online algorithm for multivariate distributional regression models, allowing efficient modeling of the conditional means, variances, and dependence structures of electricity prices. The approach combines multivariate distributional regression with online coordinate descent and LASSO-type regularization (absolute shrinkage and selection operator), enabling scalable estimation in high-dimensional covariate spaces. Additionally, we propose a regularized estimation path over increasingly complex dependence structures, allowing for early stopping and avoiding overfitting. In a case study using historical data from the German day-ahead market, the proposed method yields interpretable and well-calibrated joint prediction intervals for the 24-dimensional price distribution and provides robust performance across a range of proper scoring rules. The results underscore the importance of modeling the dependence structure of electricity prices. Furthermore, we analyze the trade-off between predictive accuracy and computational costs for batch and online estimation and provide a high-performing open-source Python implementation in the ondil package.2025-04-03T12:08:51ZRevised Version August 2026. 34 pages incl. appendix, 10 figures, 6 tables, 14Simon Hirschhttp://arxiv.org/abs/2507.04553v2AL-SPCE - Reliability analysis for nondeterministic models using stochastic polynomial chaos expansions and active learning2026-09-01T16:10:25ZReliability analysis traditionally relies on deterministic simulators, where identical inputs yield identical outputs. However, many real-world systems exhibit stochastic behavior, producing non-repeatable outcomes even under identical conditions. Stochastic simulators account for this behavior by representing the response as a random variable, whose intrinsic variability must be considered in reliability analysis. While Monte Carlo simulation can address this problem, its computational cost is often prohibitive. Stochastic emulators have therefore been introduced as surrogate models capable of reproducing the random simulator response at reduced cost. Recent studies have shown their potential for reliability analysis, but accurate estimates may still require relatively large training sets, which can be impractical for expensive models. In this work, we propose an active learning framework to further reduce the computational effort. Focusing on stochastic polynomial chaos expansions (SPCE), we introduce a learning function that identifies regions relevant to reliability estimation where the emulator exhibits high predictive uncertainty. We further exploit the asymptotic normality of the maximum likelihood estimator to quantify local prediction uncertainty. The resulting methodology, termed active learning stochastic polynomial chaos expansions (AL-SPCE), is validated on three types of problems. In all cases, AL-SPCE significantly improves computational efficiency compared with previous surrogate-based approaches and direct Monte Carlo simulation, while maintaining accurate reliability estimates.2025-07-06T22:07:57ZStructural Safety, Vol. 123, no. 102747, 2026A. PiresM. MoustaphaS. MarelliB. Sudret10.1016/j.strusafe.2026.102747http://arxiv.org/abs/2609.01133v1Scalable Inversion of Contests with Correlated Performances, Including Softmax and Multinomial Probit2026-09-01T12:09:42ZMultinomial probit choice probabilities over n alternatives are Gaussian orthant integrals, computed by simulation for thirty years, one expensive integral per alternative. Inversion, which is to say determining item attractiveness consistent with a prescribed choice probability vector, is even more difficult and has been considered impractical for correlated contests when n is large. Yet here, for families lying within a grammar including factor, block and hierarchical covariance structures, we exhibit a calibration tested at n = 1,000,000 reproducing probabilities to very high accuracy, even in the extreme tail. We must return to much smaller problems for any performance comparison to be possible due to limitations of the prior art. The Geweke-Hajivassiliou-Keane simulator is the standard (and still appropriate for high rank) but is two hundred times slower already at n = 200, and its measured cost grows roughly as n^2.8 while ours is linear. Furthermore our approach applies to any continuous performance distributions within reason: the Thurstone-Mosteller model families thereby become a practical alternative to logit at modern scale.2026-09-01T12:09:42ZPeter Cottonhttp://arxiv.org/abs/2510.01941v2A debiased Bernoulli factory and unbiased estimation of a probability2026-09-01T09:49:31ZGiven a known function $f : [0, 1] \to (0, 1)$ and a random but almost surely finite number of independent, Ber$(x)$-distributed random variables with unknown $x \in [0, 1]$, we prove the existence of an unbiased, $[0, 1]$-valued estimator of the probability $f(x) \in (0, 1)$. Our estimator is based on so-called debiasing, or randomly truncating a telescopic series of consistent estimators. Debiased estimators of a probability are not typically constrained to $[0, 1]$, or even bounded, even when all consistent estimators used as inputs are. We show that constructing the series of consistent estimators from the coefficients of a particular Bernoulli factory yields provable boundedness provided $f \in C^ρ[0, 1]$ for $ρ> 3$. Our result can be thought of as a novel Bernoulli factory with the appealing property that the required number of Ber$(x)$-distributed random variates is independent of their outcomes.2025-10-02T12:04:19Z10 pagesJere KoskelaToni KarvonenKrzysztof ŁatuszyńskiDario Spanòhttp://arxiv.org/abs/2609.00905v1When Metropolis and Hastings Meet Bradley and Terry: Exact MCMC From Preference Voting2026-09-01T08:36:43ZSampling from distributions conditioned on desired semantic properties is an emerging challenge in modern generative modeling. Metropolis-Hastings (MH) provides a principled route to conditional sampling, but requires access to exact pointwise target-density evaluations, which are not available in generative settings. Meanwhile, pairwise comparisons by humans or model "judge" are highly accessible and have proved valuable across diverse applications. We introduce Pref-MH, a general exact MH sampler for judge-induced conditional distributions using only stochastic binary pairwise comparisons. Our key observation is that the MH unnormalized density ratio matches the preference odds of the Bradley-Terry (BT) choice model. The central challenge is that while MH requires precise ratio computation, BT judges provide only sampled binary feedback. To this end, we develop a valid accept/reject rule whose resulting Markov chain provably converges to the target distribution. We further show that, for a fixed proposal kernel and budget, Pref-MH is optimal in the Peskun-Tierney sense among this class of exact reversible acceptance rules. Experiments on text generation and molecular design with LLM judges, as well as image generation with VLM judges, demonstrate that Pref-MH provides a practical and flexible approach to conditional sampling when comparative feedback is relatively easy to obtain.2026-09-01T08:36:43ZAriel SmogorghevskiNir RosenfeldYaniv Romanohttp://arxiv.org/abs/2607.25250v2A Copula-Based Regression Framework for Enhanced Prediction under Heteroscedasticity2026-09-01T08:22:58ZClassical regression approaches, including ordinary least squares, rely on strong assumptions such as constant variance and normality of residuals, which are often violated in real-world data. Although log-transformation is commonly used to stabilise variance, it may introduce re-transformation bias and fail to address heteroscedasticity and asymmetric dependence structures adequately. To overcome these limitations, this study proposes a copula-based regression framework for modelling data in the presence of heteroscedastic error structures. The proposed copula-based regression framework separates marginal distributions of the response and explanatory variables from their dependence structure, allowing flexible modelling of different tail-dependent relationships. The proposed approach explicitly accounts for heteroscedasticity without requiring restrictive distributional assumptions. A comprehensive simulation study and two real-world applications were considered under heteroscedastic scenarios to compare the performance of the proposed method with existing methods. The simulation results demonstrated that the proposed copula-based model consistently outperformed conventional approaches, achieving an average mean absolute percentage error of 0.21, compared with 0.27 and 0.36 for the linear and log-linear models, respectively. In the first application, which exhibited clear heteroscedasticity, the copula-based model achieved the lowest MAPE, although the overall differences between the copula-based and GAMLSS models were not substantial. In Application 2, all models achieved low predictive performance, as there was a moderate linear relationship between the response and predictor variables. Overall, the findings indicate that no single model consistently dominates across all settings, while copula-based regression provides a flexible and competitive alternative for heteroscedastic data.2026-07-28T03:52:19ZDeepani HemachandraJagath SenarathneMahasen Dehideniyahttp://arxiv.org/abs/2609.00773v1Deep Skew-t Mixture Models2026-09-01T06:02:11ZHigh-dimensional clustering is challenging when component distributions are both heavy-tailed and directionally asymmetric. We propose a deep skew-$t$ mixture model (DStMM), a hierarchical factor-analytic mixture based on the generalised-hyperbolic skew-$t$ normal mean--variance representation. A shared inverse-gamma mixing variable is propagated along each complete latent pathway, allowing heavy tails and directional asymmetry to be modelled jointly while preserving conditional Gaussianity. Each complete pathway therefore admits an exact GHST marginal representation. We formalise the reductions to symmetric deep $t$, Gaussian deep-mixture, and single-layer GHST factor-analytic models, discuss local non-identifiability and the implementation-level parameter-counting convention, and derive the conditional generalised inverse Gaussian law used for estimation. Estimation is carried out by a stochastic/Monte Carlo EM algorithm, with an explicit implementation-based parameter count for BIC architecture comparison. Simulation studies show that DStMM performs similarly to the symmetric robust model when skewness is absent but provides increasing gains as directional asymmetry becomes stronger, particularly under heavier tails; the same qualitative behaviour persists under smaller samples and unequal mixture proportions. Two real-data applications provide complementary evidence. On the UCI handwritten-digit benchmark, DStMM gives the strongest clustering performance under a common deep architecture, while on the Gas Sensor Array Drift data, DStMM improves on both deep Gaussian and deep $t$ alternatives and, under the implemented BIC criterion, selects a non-trivial second mixture layer. Together, these results support the value of propagating skewness and heavy-tail variation through a deep latent mixture while retaining an exact pathway-level likelihood.2026-09-01T06:02:11ZJinran WuYou-Gan WangGeoffrey J. McLachlanhttp://arxiv.org/abs/2603.07701v3Fractional Topological Phases, Flat Bands, and Protected Zero-Energy Edge States in Small Cyclic Quantum Walks2026-09-01T04:15:34ZWe report the first realization of a fractional topological phase in a fully unitary, noninteracting discrete-time quantum walk implemented on finite cyclic graphs. Using a single-coin split-step cyclic quantum walk (SCSS-CQW), we uncover topological phenomena that are inaccessible within conventional cyclic quantum-walk dynamics. The protocol enables controlled engineering of quasienergy spectra, flat bands, and topological phase transitions through the step-dependency parameter and coin-rotation angle. We show that cyclic graphs with even and odd numbers of sites exhibit qualitatively different band structures, while rotational flat bands arise exclusively in $4n$-site cycles; a general analytic condition for their emergence is derived. The SCSS-CQW produces fractional winding numbers $\pm \frac{1}{2}$ (Zak phases $\pm \fracπ{2}$), in sharp contrast with the integer invariants in standard cyclic quantum walks. These fractional invariants lead to an unconventional bulk-boundary correspondence and support protected {zero energy} edge states beyond the usual integer topological classification. In the step-dependent protocol, transitions between distinct fractional winding sectors generate robust edge modes pinned at zero quasienergy. Numerical simulations show that these states remain stable in the presence of both dynamic and static coin disorder as well as phase-preserving perturbations, while survival-probability analysis demonstrates their long-time persistence. Requiring only a constant number of detectors independent of the evolution time and fewer operators, the proposed scheme offers a minimal-resource and experimentally accessible platform for realizing fractional topology, flat bands, and protected zero energy edge states in small-scale synthetic quantum systems.2026-03-08T15:59:21Z30 pages, 27 figures, 3 tablesDinesh Kumar PandaAditi RathColin Benjamin