https://arxiv.org/api/ZPD+H0kedGXAy0sKfz7yeDq6vOM2026-10-02T21:51:59Z1067528515http://arxiv.org/abs/2608.19377v1Heteroscedastic Neural Surrogate Modeling for Robust and Rapid Bayesian Inference in Fusion Plasma Diagnostics2026-08-19T18:45:56ZBayesian inference via Markov Chain Monte Carlo (MCMC) provides effective parameter estimation, but its real-time application in complex physical systems is hindered by heavy computational bottlenecks and extreme sensitivity to statistical noise. We address this by proposing a neural-network-based probabilistic surrogate framework for rapid and robust MCMC inference. Using fusion plasma Thomson scattering diagnostics as a challenging, noise-dominated testbed, our approach employs a dual-head architecture to simultaneously estimate the expected physical emission spectrum and the channel-wise intrinsic measurement noise variance. By optimizing a Gaussian Negative Log-Likelihood (GNLL) objective, the learned aleatoric uncertainty dynamically buffers the sampler against pathological shot noise. Evaluations demonstrate that this surrogate framework achieves > 1500x acceleration over exact physical forward models, while simultaneously reducing inference error (RMSE) by >20% compared to standard homoscedastic neural baselines, offering a highly promising paradigm for real-time physical analysis.2026-08-19T18:45:56Z8 pages, 3 figuresLiyun ZhangNaoya MamadaKentaro SakaiTakeo HoshiToru Aonishihttp://arxiv.org/abs/2608.19291v1Rural Commute Patterns2026-08-19T14:52:20ZTransportation provides access to employment opportunities and essential services such as healthcare services, while urban areas have various transportation options, the situation differs in rural areas. Rural residents often have longer commute distances, limited access to public transit, and extended waiting times for public transportation if they exist, which can significantly impact their access to vital services and job opportunities. This study used data from the 2017 NHTS survey to examine the commuting patterns in rural areas by utilizing multinomial logistic regression to determine how various factors impact the choice of mode of transport in rural areas. Findings from this study revealed a higher dependency, 92.1%, on using personal vehicles when making trips in rural areas. Multinomial logistic regression results showed that socio-demographics, household, and trip characteristics affect the mode of transport used in a trip. Older adults, females, and individuals with higher education levels than high school graduates are less likely to use public transit when making trips. For household characteristics, the availability of vehicles in a household and households with higher income levels have lower probabilities of making a trip using public transit. Longer trip distances reduce the likelihood of a trip using active commuting modes such as walking and biking. These findings provide insights into understanding the transportation behaviors in rural areas and provide knowledge to be used in the planning and developing of transportation projects to promote equitable and accessible transportation in rural areas.2026-08-19T14:52:20ZThis a brief submitted as a response to COMMUTING IN AMERICA 2022Mohamed KhalafallaDoreen KobeloThobias SandoJaneroza Matyenihttp://arxiv.org/abs/2509.12173v2Extrapolation of Tempered Posteriors2026-08-19T14:33:41ZTempering is a popular tool in Bayesian computation, being used to transform a posterior distribution $p_1$ into a reference distribution $p_0$ that is more easily approximated. Several algorithms exist that start by approximating $p_0$ and proceed through a sequence of intermediate distributions $p_t$ until an approximation to $p_1$ is obtained. Our contribution reveals that high-quality approximation of terms up to $p_1$ is not essential, as knowledge of the intermediate distributions enables posterior quantities of interest to be extrapolated. Specifically, we establish conditions under which posterior expectations are determined by their associated tempered expectations on any non-empty $t$ interval. Harnessing this result, we propose novel methodology for approximating posterior expectations based on extrapolation and smoothing of tempered expectations, which we implement as a post-processing variance-reduction tool for sequential Monte Carlo.2025-09-15T17:37:40Z56 pages, 18 figuresMengxin XiZheyang ShenMarina RiabizNicolas ChopinChris J. Oateshttp://arxiv.org/abs/2608.18914v1A Composite Divergence Approach to Robust Multivariate Estimation under Cellwise and Casewise Contamination2026-08-19T13:47:14ZComposite likelihood (CL) methods provide a computationally efficient alternative to full likelihood inference for complex multivariate models by replacing the joint likelihood with a product of lower-dimensional marginal or conditional components. Like the MLE, however, the maximum CL estimator (MCLE) is highly sensitive to data contamination. On the other hand, robust divergence-based procedures such as the minimum density power divergence (DPD) estimator require the full joint density and so scale poorly to complex multivariate models.
We introduce the composite DPD (CDPD), a genuine statistical divergence built entirely from the low-dimensional component densities defining a CL, combining the computational scalability of CL with the robustness of the DPD. The resulting minimum CDPD estimator (MCDPDE) robustifies the MCLE without requiring integration over the full multivariate sample space.
We establish consistency, asymptotic normality, and the influence function of the MCDPDE under regularity conditions on the component models alone, without requiring correct specification of the full joint distribution. We show that it is qualitatively robust for every positive value of its tuning parameter, unlike the MCLE recovered as the limit. Because its components can be chosen at the pairwise or cell level, the framework guards simultaneously against casewise and cellwise contamination. Operating directly on component densities rather than elliptical distance structures, it extends robust inference beyond the elliptical models to which most existing cellwise-robust procedures are confined.
We develop computational algorithms implemented in the accompanying R package mvdpd. Simulation studies and real-data applications show that the MCDPDE achieves substantial robustness gains over the MCLE while retaining competitive efficiency under the assumed model.2026-08-19T13:47:14ZAbhik GhoshClaudio AgostinelliAyanendranath Basuhttp://arxiv.org/abs/2412.08895v4Bayesian Wideband Signal Detection via Source Signal Marginalization and RJMCMC2026-08-19T13:40:28ZConsider an array receiving unknown wideband signals from an unknown number of sources $k$. Wideband signals can occupy arbitrarily wide bandwidths, rendering demodulation-based approaches inapplicable, a common situation in settings involving acoustic signals. Here, we aim to determine $k$ given $N$ noisy array-valued measurements, a task known as the "detection problem," for which Bayesian model comparison is a common approach. To render Bayesian inference tractable, it is typically necessary to marginalize the source signals. Unfortunately, for wideband signals, naive marginalization has an unaffordable time complexity of $\mathcal{O}(N^3 k^3)$. As a result, fully Bayesian signal detection has yet to be demonstrated in wideband settings. In this work, we propose a wideband signal model that allows for computationally tractable marginalization of the source signals. We begin from the canonical model of linear time-invariant (LTI) signal propagation, which is then augmented into a circular convolution, all without loss of generality. This allows for efficient computation in the frequency domain, where the resulting linear system admits a decomposition into a sparse matrix we refer to as a \textit{stripe matrix decomposition}. Exploiting this sparsity pattern reduces the time complexity of computing the marginal likelihood to $\mathcal{O}(N k^3)$. These computational improvements enable efficient posterior inference via reversible-jump Markov chain Monte Carlo (RJMCMC). In this work, we use the non-reversible extension of RJMCMC (NRJMCMC), which often achieves lower autocorrelation and faster convergence than RJMCMC. Detection of the latent source signals can then be performed in a fully Bayesian manner using samples drawn by NRJMCMC. We evaluate our procedure by comparing it against generalized likelihood ratio testing (GLRT) and information criteria.2024-12-12T03:15:08Zv4: accepted versionIEEE Transactions on Signal Processing, 2026Kyurae KimPhilip T. ClemsonJames P. ReillyJason F. RalphSimon Maskell10.1109/TSP.2026.3721253http://arxiv.org/abs/2411.07776v3On theoretical guarantees and a blessing of dimensionality for nonconvex sampling2026-08-19T11:45:58ZGuarantees for algorithms sampling from nonlogconcave target measures on $\mathbb{R}^d$ are studied. For the class of measures with logdensities that have bounded Hessians and are strongly concave outside a Euclidean ball of radius $R$, it is shown that complete polynomial complexity can in fact be achieved if $R\leq c\sqrt{d}$. On the other hand, an exponential number of point evaluations is shown to be generally necessary for any algorithm as soon as $R\geq C\sqrt{d}$ for constants $C>c>0$. Importance sampling with a tail-matching proposal achieves the former, owing to a blessing of dimensionality. It is also shown that if strong concavity outside a ball is replaced by a distant dissipativity condition, then sampling guarantees must generally scale exponentially with $d$ in essentially all parameter regimes.2024-11-12T13:19:23ZMartin Chakhttp://arxiv.org/abs/2608.01613v2Robust Deep Mixture Models2026-08-19T03:12:04ZWe propose a robust deep mixture model based on a pathway-wise shared scale-mixture construction. Layer-specific component indicators are independently distributed according to their corresponding mixing proportions and jointly define a complete pathway through the latent hierarchy. Conditional on the selected pathway, a single gamma-distributed latent precision variable is shared across the deepest latent distribution, every intermediate latent transition, and the observation model. Integrating out this shared precision yields an exact multivariate Student-$t$ distribution for each complete pathway, allowing robustness to propagate coherently throughout the entire latent hierarchy rather than being introduced separately within individual latent layers. Model parameters are estimated using a stochastic expectation--maximisation algorithm. Complete-pathway responsibilities are evaluated analytically, whereas the shared latent precision variables and latent Gaussian variables are generated from their conditional distributions before updating the model parameters. The pathway-specific degrees-of-freedom parameters are estimated by one-dimensional numerical optimisation. Simulation studies demonstrate accurate recovery of the pathway-specific degrees-of-freedom parameters together with consistently improved clustering performance relative to the deep Gaussian mixture model under heavy-tailed and contaminated settings. Real-data applications further illustrate the ability of the proposed model to identify heterogeneous latent structures while reducing the influence of atypical observations. The proposed framework retains the hierarchical representation and parsimonious parameter-sharing structure of the deep Gaussian mixture model while providing coherent pathway-wise robustness.2026-08-03T02:40:11Z20 pages, 2 figuresJinran WuGeoffrey J. McLachlanhttp://arxiv.org/abs/2608.16911v2r2py: AI-Assisted Conversion of R Statistical Packages to Python2026-08-19T02:30:45ZThousands of R packages hold statistical methods with no native Python equivalent. Runtime bridges require an R installation; hand-written ports do not scale. Translation fails silently where the languages diverge, as in transform normalization, integer width, and argument evaluation. We present r2py, a framework that converts an R package into a native Python library using orchestrated language-model agents under human supervision, with correctness established by numerical comparison against the original at declared tolerances. The compiled code is retained unmodified, so any divergence lies in the translation. Seven phases decompose the work for independent invocations: structural analysis fixes conversion order, every base-R construct's rendering is settled in reviewable guides before code generation, and four verification methods each expose defects their predecessors miss. Packages reaching compiled code through .Call() add a five-phase prologue reconstructing the R C API they use. Conversions of KernSmooth and rpart reproduce R across 518 and 846 tests.2026-07-12T03:34:51Zv2: substantially extended. Adds a second case study (rpart) reached through R's .Call() interface, and a five-phase prologue reconstructing the portion of R's C API the package uses so its original C compiles without R. v1 covered KernSmooth only. Title shortenedYufei CaiJun Lihttp://arxiv.org/abs/2608.18441v1Convex Reparameterization and Self-Concordant Algorithms for Multivariate Regression with Covariance Estimation2026-08-19T02:13:09ZBuilding on a reparameterization for multivariate linear regression that yields a jointly convex penalized likelihood in the reparameterized regression coefficient matrix and the precision matrix, we show that the resulting scaled Gaussian loss is standard self-concordant. This places the joint estimation problem within composite self-concordant optimization and leads to two algorithms: a proximal gradient method and a damped proximal Newton method. In simulations, we evaluate algorithmic robustness, iterations to convergence, and elapsed time. In a protein expression application, compared with the classical-parameterization formulation, the proposed convex formulation attains similar mean squared prediction error and can be substantially faster when the fitted precision matrix is dense.2026-08-19T02:13:09Z4 figures and 2 tablesHongru ZhaoHuiqian Fenghttp://arxiv.org/abs/2512.02478v2Simulation and inference methods for non-Markovian stochastic biochemical reaction networks2026-08-19T00:35:45ZStochastic models of reaction networks are widely used to capture intrinsic noise in complex systems in the life sciences. Typical formulations of these models are based on Markov processes for which there is extensive research on efficient simulation and inference. However, there are complex processes in biology, such as gene transcription and translation, that introduce history dependent dynamics requiring non-Markovian processes to accurately capture the stochastic dynamics of the system. This greater realism comes with additional computational challenges for simulation and parameter inference. We develop efficient stochastic simulation algorithms for well-mixed non-Markovian stochastic reaction networks with stochastic delays that depend on system state and time. Our methods generalize the next reaction method and $τ$-leaping method to support arbitrary inter-event time distributions while preserving computational scalability. We also introduce a coupling scheme to generate exact non-Markovian sample paths that are positively correlated to an approximate non-Markovian $τ$-leaping sample path. This enables substantial computational gains for simulation and Bayesian inference through multilevel Monte Carlo and multifidelity schemes. We demonstrate the effectiveness of our approach using several non-Markovian examples, showing substantial gains in both simulation accuracy and inference efficiency. These results extend the practical applicability of non-Markovian models in systems biology and beyond.2025-12-02T07:15:17ZThomas P. SteeleDavid J. Warnehttp://arxiv.org/abs/2608.17697v1A multi-level preprocessing and modelling framework for spectral imaging of microplastics2026-08-18T12:16:04ZSpectral imaging provides chemically specific and spatially resolved analysis of microplastics, but its routine application is hindered by large data volumes, acquisition artefacts, spectral variability, and misidentification of polymers due to alike spectra. This study proposes a multi-level preprocessing and modelling framework for FT-IR spectral imaging of microplastics that integrates image-level, tile-level, and spectral-level corrections with scalable identification strategies.
Image-level variation associated with changing acquisition conditions was done with latent variable selection, while a background-based tile correction reduced illumination-related artefacts. Spectral preprocessing combined baseline correction, smoothing, derivative calculation, normalization, and wavelength selection, and only particle spectra were retained for further analysis to improve computational efficiency. For scalable identification, clustering was applied to particle spectra and spectral library matching was performed on cluster centroids instead of individual pixels. Among twelve evaluated matching strategies, a sign-invariant derivative-based cosine similarity method achieved perfect classification accuracy for polystyrene (PS), polyethylene terephthalate (PET), polyethylene (PE), and polypropylene (PP). The clustering-based workflow also produced more spatially coherent particle maps than direct software-based matching while substantially reducing processing time. The framework was evaluated for supervised classification-based MP indentification. These results show that multi-level correction combined with cluster-centroid spectral matching improves the robustness, efficiency, and interpretability of spectral-imaging-based microplastic identification.2026-08-18T12:16:04ZZina-Sabrina DumaTenzin TseringSara HeikkinenTuomo SoininenTuomas SihvonenArto KoistinenSatu-Pia Reinikainenhttp://arxiv.org/abs/2402.15292v2adjustedCurves: Estimating Confounder-Adjusted Survival Curves in R2026-08-18T10:39:31ZKaplan-Meier curves stratified by treatment allocation are the most popular way to depict causal effects in studies with right-censored time-to-event endpoints. If the treatment is randomly assigned and the sample size of the study is adequate, this method produces unbiased estimates of the population-averaged counterfactual survival curves. However, in observational studies, this is no longer the case. Instead, specific methods that allow adjustment for confounding must be used. We present the \texttt{adjustedCurves} \textbf{R} package, which can be used to estimate and plot these confounder-adjusted survival curves using a variety of methods from the literature. It provides a convenient wrapper around existing \textbf{R} packages on the topic and adds additional methods and functionality on top of it, uniting the sometimes vastly different methods under one consistent framework. Among the additional features are the estimation of confidence intervals, confounder-adjusted restricted mean survival times and confounder-adjusted survival time quantiles. After giving a brief overview of the implemented methods, we illustrate the package using publicly available data from an observational study including 2982 breast cancer.2024-02-23T12:38:27ZUnder Review in "Observational Studies"Robin DenzNina Timmesfeldhttp://arxiv.org/abs/2508.14487v3Bridge Sampling Diagnostics2026-08-18T09:34:19ZIn Bayesian statistics, the marginal likelihood is used for model selection and averaging, yet it is often challenging to compute accurately for complex models. Approaches such as bridge sampling, while effective, suffer from high variance when the proposal distribution overlaps poorly with the target posterior. To quantify this variance, we present a closed-form Monte Carlo standard error (MCSE) estimator for bridge sampling, extending classical variance approximations with a multi-chain effective-sample-size correction for autocorrelated MCMC draws and an exact log-scale variance. We show that the MCSE estimate itself is structurally capped at about 1.05, so values near this cap signal saturation rather than precision, and our calibration experiments show that the MCSE can be trusted when it is below 0.3. Furthermore, we introduce a hybrid score-matching proposal that regularizes the sample covariance using the local posterior geometry, significantly improving the stability of the estimator, and we demonstrate the efficacy of these methods using increasingly difficult simulated posteriors and real posteriors from the posteriordb database.2025-08-20T07:23:45ZRevised diagnostic recommendation, new hybrid score-matching proposal, extended experimentsGiorgio MicalettoAki Vehtarihttp://arxiv.org/abs/2608.17531v1The Snake Algorithm: A Rejection-Free Sampler for Binary Matrices with Fixed Margins2026-08-18T08:52:20ZWe study uniform sampling of binary matrices with fixed row and column sums, a recurring problem in ecological null models, Rasch-model testing, network analysis, and combinatorics. We propose the Snake algorithm, a rejection-free Markov chain Monte Carlo sampler that grows an alternating path until its first self-intersection and flips the resulting loop. The chain is reversible and irreducible on the fixed-margin state space, hence has the uniform stationary distribution. We prove that one step flips on the order of $\sqrt{n}$ entries in sparse and balanced square regimes, give upper bounds on the per-step path length, and show that the resulting work per flipped entry is rate optimal in sparse and balanced regimes and near-optimal up to a polylogarithmic factor under a one-sided half-balanced condition. A Markov-chain comparison, combined with the recently established universal spectral-gap bound for the swap chain, proves that the lazy Snake chain is rapidly mixing for every feasible pair of margins; in the permutation-matrix case, the raw chain has the sharp total-variation mixing time $Θ(n \log n)$. We also describe a directed-graph extension and an equal-margin label-shuffling variant. Numerical experiments against Swap, Rectangle Loop, Curveball, sequential importance sampling, and a directed edge-swap algorithm show consistent gains in move size, wall-clock convergence, and sampling efficiency.2026-08-18T08:52:20ZZipei NieGuanyang WangPeng Zhanghttp://arxiv.org/abs/2608.17246v1Physics-Informed and Hybrid Machine Learning in Additive Manufacturing: Application to Fused Filament Fabrication2026-08-18T01:02:59ZThis article investigates several physics-informed and hybrid machine learning strategies that incorporate physics knowledge in experimental data-driven deep-learning models for predicting the bond quality and porosity of fused filament fabrication (FFF) parts. Three types of strategies are explored to incorporate physics constraints and multi-physics FFF simulation results into a deep neural network (DNN), thus ensuring consistency with physical laws: (1) incorporate physics constraints within the loss function of the DNN, (2) use physics model outputs as additional inputs to the DNN model, and (3) pre-train a DNN model with physics model input-output and then update it with experimental data. These strategies help to enforce a physically consistent relationship between bond quality and tensile strength, thus making porosity predictions physically meaningful. Eight different combinations of the above strategies are investigated. The results show how the combination of multiple strategies produces accurate machine learning models even with limited experimental data.2026-08-18T01:02:59Z11 pages, JOM (Journal of The Minerals, Metals & Materials Society)JOM 72, 4695--4705 (2020)Berkcan KapusuzogluSankaran Mahadevan10.1007/s11837-020-04438-4