https://arxiv.org/api/kOB0jP+qJ62QnhBAdz97uZE0beo2026-10-02T18:05:06Z1067521015http://arxiv.org/abs/2505.08144v2Structural Packing and Dyadic Factorization of Sparse Positive Definite Matrices2026-08-30T22:00:24ZEfficient inversion of large sparse positive definite matrices requires exploiting sparsity patterns beyond those captured by conventional bandwidth reduction. In this work, we recast nested dissection, a prominent alternative, as a two-stage framework. The matrix was first packed into block-tridiagonal or dyadic form, followed by sparse Gram-Schmidt orthogonalization. This decomposition provided a unified perspective on sparse matrix factorization and inversion and identified dyadic structure as a fundamental component of sparse Cholesky factorization. For the first stage, we introduced a packing algorithm that recovered block-tridiagonal and dyadic patterns using a novel $\ell_1$ criterion. Using approximate distances obtained through classical multidimensional scaling, the method was effective when the target structure was sufficiently represented among the nonzero entries. Iterative application could also remove structural noise and reveal hidden dyadic organization, corresponding to separator identification in nested dissection. For the second stage, we developed the theory of dyadically structured matrices. We derived sparse factorization and inversion procedures, analyzed their computational complexity, and obtained an efficient inversion algorithm. A modified version reduced the cost of inverting block-tridiagonal matrices, demonstrating the benefit of exploiting their structure directly rather than treating them as generic band matrices.2025-05-13T00:33:44ZMichał KosKrzysztof PodgórskiHanqing Wuhttp://arxiv.org/abs/2512.09060v3All Emulators are Wrong, Many are Useful, and Some are More Useful Than Others: A Reproducible Comparison of Computer Model Surrogates2026-08-30T21:40:00ZAccurate and efficient surrogate modeling is essential for modern computational science, and there are a staggering number of emulation methods to choose from. With new methods being developed all the time, comparing the relative strengths and weaknesses of different methods remains a challenge due to inconsistent benchmarking practices and (sometimes) limited reproducibility and transparency. In this work, we present a large-scale, fully reproducible comparison of $29$ distinct emulators across $60$ canonical test functions and $40$ real emulation datasets. To facilitate rigorous, apples-to-apples comparisons, we introduce the R package \texttt{duqling}, which streamlines reproducible simulation studies using a consistent, simple syntax, and automatic internal scaling of inputs. This framework allows researchers to compare emulators in a unified environment and makes it possible to replicate or extend previous studies with minimal effort, even across different publications. Our results provide detailed empirical insight into the strengths and weaknesses of state-of-the-art emulators and offer guidance for both method developers and practitioners selecting a surrogate for new data. We discuss best practices for emulator comparison and highlight how \texttt{duqling} can accelerate research in emulator design and application.2025-12-09T19:25:50ZKellin N. RumseyGraham C. GibsonDevin FrancomReid Morrishttp://arxiv.org/abs/2604.21994v2Contrast-Space Projection for Network Meta-Analysis: An Exact and Invariant Study-Based Decomposition of Direct and Indirect Contributions2026-08-30T12:18:11ZNetwork meta-analysis (NMA) combines direct and indirect comparisons across a treatment network, but exact contribution decompositions that reproduce NMA estimates are lacking, especially for multi-arm trials with correlated contrasts. We develop a contrast-space projection formulation of NMA that expresses the estimator as a linear mapping of observed pairwise contrasts onto the consistency-constrained contrast space. Building on this representation, we define direct and indirect evidence through a canonical within-study reduction that removes algebraic redundancy and yields a unique, invariant study-level decomposition. The resulting covariance-aware weights exactly reconstruct the NMA estimator and can be further resolved into indirect path-level components. Under fixed effects, the same projection also represents the generalized Cochran Q decomposition into within-design heterogeneity and between-design inconsistency. The framework yields diagnostic and graphical tools, including forest plots, tension plots, and path-based visualizations. Applications to empirical networks illustrate how the approach provides a reproducible and interpretable account of evidence contributions in NMA.2026-04-23T18:23:32Z43 pages, 6 figures. Revised version with expanded theoretical development, clarification of the representation-invariant study-based direct and indirect decomposition, additional results on canonical reduction and path construction, revised generalized Cochran Q diagnostics, and updated empirical illustrations and expositionChong WangYanqi ZhangZhezhen JinAnnette O'Connorhttp://arxiv.org/abs/2608.29265v1Signed random Fourier features for fast density estimation with indefinite kernels2026-08-29T13:38:45ZKernel density estimation (KDE) is one of the most fundamental statistical estimators of density functions. Its direct implementation on a dataset of $N$ points incurs an $\mathcal{O}(N^{2})$ computational cost, which is prohibitive for large-scale datasets. Kernel approximation techniques can be applied to bring the computational cost down to $\mathcal{O}(N)$. The random Fourier features (RFF) technique, based on sampling from the spectral density of the kernel function, has become popular to speed up kernel estimators for machine learning applications. Unfortunately, it is restricted to positive definite kernels, while the majority of kernel functions popular in KDE, such as the parabolic kernel, do not satisfy this property. To overcome this limitation, this article introduces the signed random Fourier features (SRFF) technique. It is a generalization of RFF compatible with indefinite kernels whose inverse Fourier transform is absolutely integrable. The motivation for introducing this method is to speed up KDE in the case of multivariate compact kernels, which are generally not positive definite. We detail how to implement SRFF for both product kernels and isotropic kernels. For the class of Kuttner-Golubov kernels $K(\boldsymbol{x}_{i},\boldsymbol{x}_{j})=(1-\left\Vert \boldsymbol{x}_{i}-\boldsymbol{x}_{j}\right\Vert ^α)^β\mathbf{1}_{\{\left\Vert \boldsymbol{x}_{i}-\boldsymbol{x}_{j}\right\Vert \leq1\}}$ where $\boldsymbol{x}_{i}\in\mathbb{R}^{d}$, $\boldsymbol{x}_{j}\in\mathbb{R}^{d}$, $α>0$, $β>0$, which includes the triangular, parabolic, biweight, triweight, and other kernel functions of interest for KDE as particular examples, we provide an explicit acceptance-rejection algorithm to sample from its signed spectral density. Our numerical tests on a dataset of one million points confirm the computational efficiency and accuracy of SRFF for large-scale KDE.2026-08-29T13:38:45Z25 pages, 12 figuresXie WangNicolas LangrenéWen Chenhttp://arxiv.org/abs/2505.05438v3Scalable Bernoulli Factory MCMC for Intractable Marginalised Posteriors2026-08-29T12:39:07ZBernoulli factory MCMC algorithms implement accept-reject Markov chains without explicit computation of acceptance probabilities, and are used to target posterior distributions associated with intractable likelihood models. Intractable likelihoods naturally arise in continuous-time models and mixture distributions, or from the marginalisation of a tractable augmented model. Bernoulli factory MCMC algorithms often mix better than alternatives that target a tractable augmented posterior. However, we show that their computational performance typically deteriorates exponentially with data size, as happens in naively implemented pseudo-marginal algorithms. To address this, we propose a simple divide-and-conquer Bernoulli factory MCMC algorithm and prove that it has polynomial complexity for factorizing models, with the exact degree depending on the existence of efficient unbiased estimators of the intractable likelihood ratio. We demonstrate the effectiveness of our approach with applications to Bayesian inference in two intractable likelihood models, and observe respective polynomial cost of degree 1.2 and 1 in the data size.2025-05-08T17:27:44ZTimothée Stumpf-FétizonFlávio B. Gonçalveshttp://arxiv.org/abs/2608.29199v1Burn-in-Free Simulation of VARMA Time Series2026-08-29T11:20:34ZVarmapack is a software package for efficient, exact simulation of VARMA time series without a burn-in period. For stationary models, Varmapack can generate initial states and innovations from their joint stationary distribution. Alternatively, the user can supply initial states, in which case innovations are generated from their conditional distribution. The core C library has interfaces to R, Python, and Matlab. It also offers VARMAX simulation and computation of autocovariances and correlations, spectral radii, and impulse responses. Varmapack automatically selects between vector-Yule-Walker and state-space methods for covariance computation and uses level-3 BLAS operations for efficient generation of multiple replicates. Benchmarks on several platforms show substantial performance gains over existing simulation software, ranging from several-fold to more than three orders of magnitude for the packages and models considered. Varmapack is open source and publicly available through GitHub.2026-08-29T11:20:34ZKristján Jónassonhttp://arxiv.org/abs/2608.28566v1On two proofs of $d^2$ mixing of weighted Dikin walks2026-08-28T17:43:58ZWe study the mixing time of weighted Dikin walks for sampling from exponential distributions on polytopes and truncated positive-semidefinite (PSD) cones. Our first result gives a general total-variation mixing bound under strong self-concordance, $\barν$-symmetry, and mixed-trace regularity on the local metric. The key idea is to control the Metropolis--Hastings acceptance probability on a high-probability region rather than at every point. Applying this framework to the Lee--Sidford, Lewis-weight, and John metrics yields an $\widetilde O(d^2)$ mixing bound for sampling from polytopes, while applying it to a hybrid barrier yields an $\widetilde O(d^4)$ mixing bound for sampling from truncated PSD cones. Our second result establishes stronger $χ^2$-divergence guarantees and pointwise acceptance control using a new fourth-order bootstrap condition. For a suitably scaled Lee--Sidford metric, this yields an $\widetilde O(d^2)$ mixing bound in $χ^2$-divergence, improving on the previous $\widetilde O(d^{9/4})$ bound.2026-08-28T17:43:58Z36 pages. AI disclosure includedYuansi ChenYunbum Kookhttp://arxiv.org/abs/2608.28419v1Refining Relational Event Models: Bayesian Penalization and Variable Selection in REMs2026-08-28T15:05:40ZRelational Event Models (REMs) provide valuable insights into the dynamics of longitudinal social networks. Yet, the vast availability of potential predictors for a dyad's event rate poses the risk of selecting irrelevant variables and specifying an overfitted model that does not generalize to new data. Despite the recent popularity of Bayesian regularization methods to address this, there has not been a systematic evaluation of the performance of Exact Bayesian Regularization (EBR) and Approximate Bayesian Regularization (ABR) against standard Maximum Likelihood (ML) procedures commonly used for REMs. To address this, we conduct a simulation study in which directed Relational Event History (REH) data is generated with endogenous and exogenous effects of varying strength and compare the performance of (a) unregularized ML estimation, (b) ABR through normal approximations of the likelihood with Ridge and Horseshoe priors, and (c) EBR with Horseshoe priors. We find that threshold-based criteria applied to ABR and EBR with Horseshoe priors outperform univariate selection using p-values from ML estimation in variable selection accuracy. Regarding predictive performance, ABR with Ridge priors and EBR with Horseshoe priors outperform ML REMs, particularly for small sample sizes. Given only small gains with large computational burdens when using EBR, we therefore advise researchers to use ABR with a suitable prior.2026-08-28T15:05:40ZJonathan KoopSara van ErpMahdi Shafiee Kamalabadhttp://arxiv.org/abs/2512.00220v3Iterated sampling importance resampling with adaptive number of proposals2026-08-28T13:07:51ZIterated sampling importance resampling (i-SIR) is a Markov chain Monte Carlo (MCMC) algorithm which is based on $N$ independent proposals. As $N$ grows, its samples become nearly independent, but with an increased computational cost. We discuss a method which finds an approximately optimal number of proposals $N$ in terms of the asymptotic efficiency. The optimal $N$ depends on both the mixing properties of the i-SIR chain and the (parallel) computing costs. Our method for finding an appropriate $N$ is based on an approximate asymptotic variance of the i-SIR, which has similar properties as the i-SIR asymptotic variance, and a generalised i-SIR transition having fractional `number of proposals.' These lead to an adaptive i-SIR algorithm, which tunes the number of proposals automatically during sampling. Our experiments demonstrate that our approximate efficiency and the adaptive i-SIR algorithm have promising empirical behaviour. We also present new theoretical results regarding the i-SIR, such as the convexity of asymptotic variance in the number of proposals, which can be of independent interest.2025-11-28T21:40:46ZPietari LaitinenMatti Viholahttp://arxiv.org/abs/2608.01625v2Target Trial Emulation with the R Package TTE: A Tutorial and Methodological Guide2026-08-28T09:41:23ZTarget trial emulation structures observational causal analyses around the protocol of an ideal randomized trial. By aligning eligibility, treatment assignment, time zero, and follow-up, it can reduce avoidable biases, but implementation still requires coordinated decisions about data construction, inverse probability weighting, diagnostics, outcome models, standardization, competing risks, and uncertainty estimation. This article provides a self-contained methodological guide and practical tutorial for TTE, an R package for target trial emulations with longitudinal observational data. We describe target trial protocols, intention-to-treat and per-protocol estimands, identification assumptions, baseline and person-period data structures, temporal ordering for longitudinal weights, stabilized treatment and censoring weights, weight truncation, balance and effective-sample-size diagnostics, weighted pooled discrete-time survival models, model-based standardization, competing-risk analysis, weighted Kaplan-Meier and Aalen-Johansen estimation, and cluster bootstrap at the original-individual level. Two fully synthetic examples illustrate end-to-end workflows: sodium-glucose cotransporter 2 inhibitor versus dipeptidyl peptidase-4 inhibitor initiation with all-cause death, and sequentially nested angiotensin receptor blocker versus calcium channel blocker trials with heart-failure hospitalization and competing death. The examples show how to obtain and diagnose estimates in R and interpret relative hazards, absolute risks, cumulative incidence, and differences between intention-to-treat and per-protocol effects.2026-08-03T02:57:36Z24 pages, 3 figures. The accompanying R package TTE is available from CRAN at https://doi.org/10.32614/CRAN.package.TTE. Worked-example scripts and supplementary materials are available at https://github.com/nomahi/TTEHisashi Nomahttp://arxiv.org/abs/2608.28110v1SCAN: Sequentially Detecting Change-points via Adaptive Nonparametric Inference2026-08-28T09:15:18ZModern time series are often long, serially dependent, and non-stationary. Existing change-point methods either target specific changes or become computationally intensive when using nonparametric costs on long series. Many also require thresholds to be carefully calibrated under serial dependence. We introduce SCAN, an offline method for detecting multiple distributional change-points in long, serially dependent univariate time series. SCAN compares adjacent windows using an integral probability metric, calibrates local discrepancies with a dependence-aware bootstrap, and refines candidate locations using a scaled 1-Wasserstein criterion, enabling detection of changes in mean, variance, and broader distributional structure within a unified framework. An ensemble over multiple window sizes reduces sensitivity to window size and threshold specification. We establish consistency of the estimated number and locations of change-points under exponential alpha-mixing dependence, and show that the localization statistic reduces to a CUSUM-type statistic under pure mean shifts. In simulations with up to one million observations, SCAN generally achieves higher covering and F1-scores than competing methods across mean and joint mean-variance shifts, particularly under serial dependence. On real data, SCAN identifies labeled activity transitions in sensor data and interpretable structural changes in hourly Bitcoin prices. Implementations are available in the Python package scan-cpd and R package scanr.2026-08-28T09:15:18ZAshoka PrabashwaraPatricia MenéndezLiam HodgkinsonStuart Leehttp://arxiv.org/abs/2608.28087v1A Design Concept of Forecasting Software for Normalized Vector Autoregressions with Fat Tails and Stochastic Volatility2026-08-28T08:55:22ZWe present a suite of R packages for macroeconomic forecasting that leverages advanced Bayesian, structural, multivariate, dynamic, hierarchical, non-linear, and non-Gaussian models. The suite enables both structural and predictive analyses, and is adapted to time series data across various types, dimensions, and sampling frequencies. Each additional feature increases computational complexity. To address this challenge, our software design incorporates a carefully curated selection of models, efficient algorithms implemented in C++, advanced econometric and numerical methods, robust handling of complex input and output objects, and standardised workflows. This approach combines the computational efficiency of C++ with the convenience of working with data in R. We demonstrate that our packages facilitate original research contributions in forecasting, as illustrated by our example in which vector autoregressions with non-centred stochastic volatility enhance density and point predictions relative to models with centred stochastic volatility.2026-08-28T08:55:22ZFei ShangXiaolei WangTomasz Woźniakhttp://arxiv.org/abs/2608.27884v1Exploiting Exact Conditionals Improves Conditioning: Provably Fast Mixing Time Bounds By Sampling from the Marginal2026-08-28T03:43:22ZThe problem of sampling from a probability distribution arises in many applications such as posterior sampling in hierarchical Bayesian inverse problems and Gaussian processes for machine learning. Markov chain Monte Carlo (MCMC) algorithms are often used for sampling from a target probability distribution, but implementations can be computationally expensive, especially for large-scale problems. In certain applications, the target distribution naturally factorizes into a lower dimensional marginal distribution and a conditional distribution that allows exact sampling. We describe an MCMC algorithm called MarCo that exploits such a structure and generates a Markov chain via Metropolis-Hastings sampling from the marginal distribution, followed by sampling from the exact conditional distribution. By design, MarCo constructs a Markov chain on the joint space that inherits the convergence behavior of the marginal MCMC algorithm. This provides multiple theoretical and computational advantages. We prove that MarCo can achieve improved mixing time upper bounds compared to direct sampling from the joint distribution. Moreover, compared to one-block methods that also exploit marginal-conditional structure, we use the framework of Peskun-Tierney ordering to show that MarCo has a larger right spectral gap and smaller asymptotic variance, thus leading to superior convergence properties. Numerical results illustrate the performance benefits of MarCo and are provided for various problems, including a semi-blind image deblurring example.2026-08-28T03:43:22Z26 pages, 3 figuresAbhijit ChowdharyFederica MilinanniJulianne ChungElizabeth Newmanhttp://arxiv.org/abs/2412.10038v3Stochastic Variational Inference for Structured Additive Distributional Regression2026-08-27T21:19:39ZStructured additive distributional regression extends generalized additive models by allowing all parameters of a response distribution to depend on structured additive predictors. Bayesian formulations regularize these models through prior distributions that enforce smoothness or shrinkage, but Markov chain Monte Carlo methods, the standard computational tool, scale poorly with sample size and require substantial runtime. We develop a scalable stochastic variational inference framework for approximate Bayesian inference in structured additive distributional regression. Our key contribution is a variational family for the regression coefficients that is amortized over covariates, responses, and smoothing parameters. This construction allows the variational location and precision to adapt locally to the data, which leads to substantially higher ELBO values in markedly less computational time compared to existing approaches. To balance accuracy and scalability, we consider a block-structured variant that constrains the precision matrix but preserves essential dependence across smooth components. Both approaches are compared with a state-of-the-art dense variational approximation and with the widely used Integrated Nested Laplace Approximation (INLA). We establish theoretical results on posterior propriety and local asymptotic Gaussianity that motivate the proposed Gaussian approximations, assess their performance in simulation studies involving logistic and gamma distributional regression, and demonstrate their scalability in an application to modeling injury counts in motor vehicle crashes in New York City with more than 200,000 observations. Across all settings, incorporating covariates, responses, and smoothing parameters into the construction of the variational distribution yields higher ELBO values in substantially less computational time compared with existing variational approaches.2024-12-13T10:59:28ZGianmarco CallegherThomas KneibJohannes SödingPaul Wiemannhttp://arxiv.org/abs/2608.27369v1Deep-Control BSDE: Layerwise Brownian-Weighted Regression for High-Dimensional Semilinear PDEs2026-08-27T17:05:43ZHigh-dimensional semilinear parabolic partial differential equations arise in stochastic control, financial engineering, and uncertainty quantification, but classical spatial discretizations suffer from the curse of dimensionality. Motivated by Gaussian perturbation and conditional regression in denoising score matching, we propose Deep-Control BSDE (\DCBSDE{}), a layerwise control-regression method for Markovian backward stochastic differential equations. From an implicit Euler scheme, we identify the discrete ideal control $z_n^π=\mathcal M_nu_{n+1}^π$ as the conditional projection coefficient of the successor value response onto the one-step Brownian increment. At each time level, the method freezes the successor value function, approximates the resulting Brownian conditional-moment target using finitely many branches, regresses the control, and then fits the value through the implicit BSDE relation. A frozen linear-response baseline and antithetic pairing reduce finite-branch fluctuations while preserving the conditional target, and a within-layer correction coordinates the value and control approximations. We establish backward stability through a contraction property of the Brownian projection and derive conditional consistency when discretization, local learning, numerical, and finite-branch errors vanish jointly. Experiments across six benchmarks demonstrate that \DCBSDE{} achieves a favorable overall balance among value accuracy, control accuracy, and dynamic consistency while exhibiting generally stable performance across random seeds. These results support the use of \DCBSDE{} in applications requiring reliable joint approximation of the value and control processes.2026-08-27T17:05:43ZMingcan WangXiangjun Wang