https://arxiv.org/api/B8sNcrOZtE4RK0/hSYTeBDcLKuM2026-09-11T20:13:58Z377403015http://arxiv.org/abs/2606.18197v2Sensitivity Bounds and Conservative Inference for Contagion under Latent Homophily2026-09-09T22:03:50ZWhether connected units are similar because influence spreads across ties or because similar units form ties is a long standing problem. We study a fixed network with two waves of nodal outcomes. Rather than positing a parametric model for network formation, we consider identification of contagion under latent homophily as a selection bias problem. We define a focal component controlled direct effect (CDE) that holds a tie present, changes one alter's lagged outcome, and permits dependence on the remaining fixed network background. We show that the gap between the CDE and the observed connected dyad risk ratio is governed by how strongly a latent variable shifts the composition of connected dyads. Under stated mean exchangeability and sensitivity restrictions, we develop interpretable nonparametric bounds for the primary naturally connected target. For inference conditional on the observed network and baseline outcomes, heterogeneous dyad specific means can invalidate the usual inclusion exclusion variance estimator; we establish asymptotically conservative one sided limits under conditional actor dissociation and stated regularity conditions. A simulation study characterizes the bounds' error control and power. We apply the framework to the 2008 U.S. House votes on the Troubled Asset Relief Program. The connected contrast among legislators suggests vote contagion and survives mild latent homophily and outcome susceptibility.2026-06-16T17:26:50ZDuncan A. Clarkhttp://arxiv.org/abs/2609.05567v2Random Jump Intensities and Bernstein Density Estimation in Ergodic Basic Affine Jump-Diffusion Processes2026-09-09T21:33:16ZThis paper develops a random-effects extension of the ergodic basic affine jump-diffusion (BAJD) model for a population of independent trajectories with individual unobserved jump intensities and common structural parameters. Under continuous-time observation, each intensity is estimated by the empirical jump frequency. We establish the strong consistency and stable mixed-normal limit of this estimator as the observation horizon tends to infinity. The estimated intensities are then used as pseudo-observations to construct a Bernstein-polynomial estimator of their common density. Its bias, variance, mean integrated squared error, and pointwise asymptotic normality are derived under a sequential asymptotic framework. The theory is developed for general positive jump-size distributions with finite activity and specialized to the Gamma$(k,λ)$ family, including the exponential and Erlang cases. For this family, we derive a pooled estimator of the common rate parameter and establish its consistency, asymptotic normality, and first-order conditional asymptotic independence from the individual intensity estimators. We also propose a consistent plug-in estimator of the population stationary mean. The methodology is illustrated through simulations and an empirical application to financial realized-volatility data.2026-09-03T22:32:31ZHamdi Fathallahhttp://arxiv.org/abs/2609.10826v1Processing and classifying bird songs using wavelet techniques and supervised learning2026-09-09T20:55:44ZThis study proposes an integrated framework for the processing and classification of invasive bird species vocalizations within natural soundscapes, characterized by high levels of environmental noise. We address the challenge of signal degradation by employing a Bayesian wavelet shrinkage methodology based on the Epanechnikov kernel prior, which offers a closed form decision rule and high computational efficiency for processing large bioacoustic datasets. The methodology was applied to recordings of three species obtained from the iNaturalist platform: \textit{Euphonia violacea}, \textit{Leiothrix lutea}, and \textit{Passer domesticus}. After signal denoising, we extracted a comprehensive set of features, including Mel-Frequency Cepstral Coefficients (MFCCs) and spectral indices such as entropy and zero-crossing rate. Several supervised learning models: Random Forest, Multinomial Logistic Regression and Support Vector Machine (SVM) were evaluated across different feature dimensionalities. Our results demonstrate that the proposed wavelet based preprocessing significantly enhances classification performance, with the SVM model achieving the highest accuracy (up to 0.9398) under a 10-dimensional MFCC configuration. This research provides a robust statistical tool for automated ecological monitoring and the management of biological invasions.2026-09-09T20:55:44ZLaura Lucia Dominguez BarriosFidel Aniano Causil BarriosAlex Rodrigo dos Santos SousaMariana Rodrigues Mottahttp://arxiv.org/abs/2301.05135v2On Existence Theorems for Conditional Inferential Models2026-09-09T20:21:56ZThe framework of Inferential Models (IMs) has recently been developed in search of what is referred to as the holy grail of statistical theory, that is, prior-free probabilistic inference. Its method of Conditional IMs (CIMs) is a critical component in that it serves as a desirable extension of the Bayes theorem for combining information when no prior distribution is available. The general form of CIMs is defined by a system of first-order homogeneous linear partial differential equations (PDEs). When admitting simple solutions, they are referred to as regular, whereas when no regular CIMs exist, they are used as the so-called local CIMs. This paper provides conditions for regular CIMs, which are shown to be equivalent to the existence of a group-theoretical representation of the underlying statistical model. It also establishes existence theorems for CIMs, which state that under mild conditions, local CIMs always exist. Finally, the paper concludes with a simple example and a few remarks on future developments of CIMs for applications to popular but inferentially nontrivial statistical models.2023-01-12T16:40:34ZThe main theorem is not correct. For a correct version, see the Frobenius theoremRongrong ZhangMichael Y. ZhuChuanhai Liuhttp://arxiv.org/abs/2605.20145v2Goal-Oriented Lower-Tail Calibration of Gaussian Processes for Bayesian Optimization2026-09-09T19:41:26ZGaussian process (GP) predictive distributions are commonly used in Bayesian optimization (BO) to guide the selection of evaluation points for expensive objective functions. The choice of kernel and hyperparameters has a strong influence on the exploration--exploitation trade-off. For minimization, sampling criteria such as expected improvement (EI) depend on both the probability mass below the current best value and the shape of the predictive distribution in this region. This article studies goal-oriented calibration of GP predictive distributions below a low threshold $t$ in the noiseless setting, for standard GP models with hyperparameters selected by maximum likelihood. We consider two complementary forms of calibration below $t$ for inputs distributed according to a reference measure $μ$: occurrence calibration over the design space and thresholded $μ$-calibration on sublevel sets of the form $\{x\in\mathbb{X}, f(x)\le t\}$. We propose tcGP, a post-hoc method that combines these two forms of calibration for GP predictive distributions below $t$. With fixed GP hyperparameters, the exact EI sampling criterion based on tcGP generates a sequence of evaluation points that is dense in the design space. Experiments on standard benchmarks show improved lower-tail calibration and BO performance relative to standard GP models and globally calibrated GP models.2026-05-19T17:32:25ZProceedings of the 43rd International Conference on Machine Learning (ICML), PMLR 306, 2026Aurélien PionEmmanuel Vazquezhttp://arxiv.org/abs/1904.00521v2Spatially Adaptive Ensemble Learning with Calibrated Predictive Uncertainty2026-09-09T18:25:47ZAir pollution exposure assessment often uses ensembles of spatio-temporal models, but a critical limitation is that conventional methods use deterministic weights and fail to quantify the uncertainty of their predictions. This yields suboptimal results and prevents the assessment of prediction reliability in health effects studies. We developed a new $\textbf{Bayesian nonparametric ensemble framework}$ that addresses these issues. Our method employs a dependent random measure to adaptively combine models based on their performance in specific feature sub-regions. It also nonparametrically models the ensemble's predictive cumulative density function (CDF), providing a data-consistent measure of uncertainty. We show that our method is asymptotically consistent and improves predictive accuracy for complex data distributions. The framework is demonstrated on simulated data and applied to generate a spatial prediction model for fine particle ($PM_{2.5}$) levels in Eastern New England, USA.2019-04-01T01:03:57ZYanran LiJeremiah Zhe LiuJohn PaisleyMarianthi-Anna KioumourtzoglouBrent A. Coullhttp://arxiv.org/abs/2609.10534v1Likelihood-free inference with nuisance parameters through normalizing flows2026-09-09T17:58:15ZWe present a simple decomposition of a neural-network-based normalizing flow that naturally uncovers a pivotal statistic (or something close) in the presence of nuisance parameters, based only on a sample generator from the distribution of interest. We show that the statistic is near-pivotal in the sense of minimum average KL-divergence of its $p$-values versus uniform and we argue that it can be expected to have good power when the dimension of the statistic equals the dimension of the parameter. It is able to incorporate prior knowledge about group invariances such as translation and scale. It can discover the one-sample $t$-test almost exactly, outperforms the Welch test in terms of worst-case size over a constrained variance-ratio range and achieves good calibration on partial biserial correlations, while showing higher power (and being much faster) on small-to-moderate samples than profile likelihood-ratio techniques.2026-09-09T17:58:15Z49 pages and 13 figures, including appendices. Code available at https://github.com/philassheton/NeuralCIsPhil Asshetonhttp://arxiv.org/abs/2602.20115v2Compound decisions and empirical Bayes via Bayesian nonparametrics2026-09-09T17:38:31ZWe study compound decision theory from a nonparametric Bayesian perspective, with particular emphasis on their relationship to empirical Bayes (EB) procedures. Motivated by the sharp risk guarantees available for EB procedures based on the nonparametric maximum likelihood estimator (NPMLE), we investigate whether analogous guarantees can be established for fully Bayesian decision rules. In a class of Gaussian compound decision problems, we show that the fully Bayesian posterior mean achieves near-optimal risk. Moreover, it is admissible as a genuine Bayes rule, whereas the corresponding NPMLE plug-in rule is inadmissible. Simulations illustrate the performance of nonparametric Bayes procedures relative to common alternatives. As an application, we apply our methodology to Census tract-level estimates of economic mobility from the Opportunity Atlas.2026-02-23T18:33:57Z69 pagesNikolaos IgnatiadisSid Kankanalahttp://arxiv.org/abs/2608.27599v2Activity-Conditioned Residual Association from Aggregated Relational Data2026-09-09T16:52:48ZAggregated relational data (ARD) record how many ties sampled respondents have to prespecified groups without revealing individual dyads. We develop a conditional test for prespecified cross-group concordance beyond an additive-activity network model. When the groups form an exhaustive partition, respondent degree is observed exactly. For arbitrary fixed activity values under an independent-Bernoulli additive-logit null, conditioning two respondents on equal degree yields a finite nonpositive sign restriction for their two disjoint group-count differences. Analyst-randomized pair thinning gives a conservative baseline. Under a stronger fixed-cell design with group-specific smooth activity profiles and positive cross-role overlap, an observable clipped full-pair statistic gives conservative one-network inference with graph-independent sampled respondents. Finite and growing matched-law constructions establish that the target contains bivariate information absent from a collapsed binary count. The full-pair procedure is uniformly consistent on a primitive relative-open neighborhood of a specified rank-one alternative. Fixed-dimensional lattice local limits make the conditioning and availability requirements explicit.2026-08-27T18:30:58Z40 pages, 2 tables. Expanded version with full-pair inference, complete proofs, and additional numerical evidenceYen-hsuan Tsenghttp://arxiv.org/abs/2501.10675v3Recovering Unobserved Network Links from Aggregated Relational Data: Bayesian Latent Surface Modeling and Penalized Regression2026-09-09T16:35:52ZAggregated relational data (ARD) record counts of ties to attribute-defined groups while leaving individual edges unobserved. We compare latent-geometry and regularized network estimators through a common observation map. The comparison distinguishes the realized adjacency matrix, conditional edge probabilities, and model parameters. We study roster-based ARD with known node-level group memberships, giving both estimators the same roster and aggregate counts.
We relate the aggregate means to a Poisson working likelihood and a Huber loss, and give their derivatives. Overlapping groups, shared edges, and reporting error affect the interpretation of these objectives. Geometry restricts the representation of edge probabilities, while regularization selects among candidate fits. Identification depends on the observation map and model restrictions rather than uniqueness of a numerical optimizer.
A reproducible synthetic experiment specifies the data-generating process, estimation algorithms, and evaluation targets under matched information. The matrix estimator gives better realized-edge rankings and aggregate fit, while the geometric estimator gives lower error for generating probabilities. The resulting framework organizes ARD reconstruction around the interaction of observation design, structural assumptions, and computation.2025-01-18T06:51:51Z14 pages, 2 figures. Substantially revised replacement of the withdrawn version. Clarified observation regime and targets; corrected likelihood and loss calculations; added a reproducible matched-input synthetic experiment. Code and saved results are included as ancillary filesYen-hsuan Tsenghttp://arxiv.org/abs/2609.10409v1Dynamic prediction intervals for survival times2026-09-09T16:30:09ZMost work on survival prediction focuses on estimating survival probabilities rather than predicting individual event times. Recent conformal methods have made it possible to construct prediction intervals for survival times with right-censored outcomes, but existing approaches are restricted to settings with covariates only measured at baseline and do not address dynamic prediction with longitudinal data. We study prediction intervals for survival times in a dynamic prediction framework with longitudinal covariates. Our approach uses Penalized Regression Calibration (PRC) as a working dynamic prediction model, combining linear mixed models for the longitudinal histories with a Cox model for post-landmark survival, and then applies a conformal calibration step to obtain prediction intervals. We compare naive intervals obtained by direct inversion of the survival function estimated by PRC to our dynamic conformal method. A Monte Carlo simulation study evaluates empirical coverage and interval length across sample sizes, censoring levels, landmark times, and non-proportional hazards (NPH) scenarios. We illustrate the proposed methodology by computing dynamic prediction intervals for the time until a dementia diagnosis in the ADNI dataset. The results show that naive inversion is often unreliable, whereas the proposed dynamic conformal method yields more stable predictive performance.2026-09-09T16:30:09Z20 pages. Supplementary material includedLorenzo CarvisigliaSaverio RanciatiMirko Signorellihttp://arxiv.org/abs/2506.09986v3Constrained Denoising, Empirical Bayes, and Optimal Transport2026-09-09T16:04:33ZIn latent variables models, two important goals are denoising and deconvolution: denoising aims to estimate the latent variables, whereas deconvolution aims to estimate the distribution of the latent variables. As has been recognized in the literature over the last century, these two goals are fundamentally in tension, since denoising yields a poor estimate of the distribution of the latent variables due to shrinkage, and deconvolution yields a distribution-valued estimate that carries no unit-specific information. In this paper, we provide a systematic study of denoisers, and empirical Bayes approximations thereof, which attain optimal denoising error subject to the constraint that the distribution of the denoised data matches, in some sense, the distribution of the latent variables. Our insight is that optimal transport allows practitioners to navigate the tension between denoising and deconvolution. More precisely, we propose a modular methodology that combines any suitable unconstrained empirical Bayes denoiser (arising, e.g., via $F$-modeling, $G$-modeling, conjugate-parametric models) with any suitable information about the distribution of the latent variables (e.g., its moments, support, or an approximation of the entire distribution via deconvolution) into a single denoised data set. We prove explicit rates of convergence for our proposed methodologies, and we apply the resulting methods in applications in astronomy, baseball analytics, and marketing.2025-06-11T17:57:17Z72 pages, 6 figures, 1 Table. Comments welcomeAdam Quinn JaffeNikolaos IgnatiadisBodhisattva Senhttp://arxiv.org/abs/2609.10313v1Semiparametric Inference for Conditional Shapley Feature Importance2026-09-09T15:19:24ZShapley values are widely used for post-hoc feature attribution, but most estimators return point quantities and do not quantify uncertainty, and popular implementations sample out-of-coalition features from their marginal distribution, which misattributes importance when features are dependent. This paper studies the conditional formulation, in which out-of-coalition features are integrated out under their true conditional distribution. The target is a global, loss-based importance that pairs a conditional value function with a SAGE-style loss aggregation. We propose a one-step estimator with K-fold cross-fitting and a U-statistic correction of the squared loss that removes the Monte Carlo bias of the naive plug-in; it is $\sqrt{n}$-consistent and asymptotically normal under double-robust rate conditions, and the resulting Wald interval attains nominal coverage. A Pinsker-type bound quantifies the bias from misspecifying the working copula class, while vine copulas keep conditional sampling tractable. In a Gaussian design study with n= 500, the empirical coverage of the 95% interval lies between 0.91 and 0.96 across all features, the test holds its Type-I rate at 0.05, and it reaches power one for moderate signals. Applied to the UCI Concrete and California Housing data, the method identifies the conditionally informative features with Bonferroni-controlled significance.2026-09-09T15:19:24ZAgostino Gnassohttp://arxiv.org/abs/2608.09218v3Online Learning of Scale Parameters in Score-Driven Filters2026-09-09T14:53:06ZA score-driven filter multiplies its scaled log-likelihood score by a scale parameter. We call this coefficient the gain and learn it online. Given the current state and realised scaled score, each admissible gain selects a reachable next state and predictive density. A scalar gain moves along a line; diagonal gains control coordinatewise transmission and may change direction. We evaluate gain selection using a one-step predictive Kullback--Leibler objective. In the scalar unscaled case, the negative consecutive-score product is a stochastic gradient; the positive product used in accelerated recursions is a descent direction. Positive scalar score scaling changes only the effective learning rate. Monotone differentiable gain links induce mirror-descent geometry, while persistence adds a Bregman pull towards a reference gain. Under convexity, compactness, integrability, and schedule conditions, projected and discounted mirror updates satisfy dynamic-regret bounds relative to time-varying, current-information comparators. Simulations isolate score scaling, link geometry, persistence, and coordinatewise gains. Across twelve equity indices, the bounded discounted-logistic gain records a lower out-of-sample mean negative log score than the constant gain in eleven markets, although market-level evidence is mixed. It also avoids the extreme transients of the numerically capped exponential-link benchmark. Improvements are largest in markets spanning multiple crises.2026-08-10T07:43:38Z62 pages, 10 figures, 13 tablesFabrizio LilloGiulia LivieriGianluca Palmarihttp://arxiv.org/abs/2609.10174v1A spatiotemporal negative binomial model with dynamic dispersion: An application to Tuberculosis infections2026-09-09T13:43:05ZTuberculosis (TB) remains a critical public health concern in Brazil, characterized by pronounced spatial heterogeneity and fluctuating temporal volatility. In this paper, we study monthly TB notifications across 61 microregions of Sao Paulo state from 2001 to 2024. To do this, we introduce a negative binomial spatial integer-valued generalized autoregressive conditional heteroskedastic (INGARCH) model featuring jointly dynamic conditional means and time-varying dispersion. To capture inter-regional spillovers, we incorporate both discrete adjacency structures and a novel continuous distance-based formulation leveraging the Matern correlation function. Parameter estimation via conditional maximum likelihood employs a two-step profile-likelihood iterative scheme, demonstrating solid finite-sample performance in simulation studies. Applied to the Sao Paulo TB surveillance data, the framework substantially outperforms standard Poisson and fixed-dispersion spatiotemporal baselines in empirical fit and uncertainty quantification, maintaining nominal 95% predictive coverage across both dense metropolitan centers and rural microregions. Our results reveal marked spatial heterogeneity in baseline incidence, dynamic overdispersion driven by localized outbreaks, and short-range spatial interaction decay. By accurately modeling spatiotemporal volatility, the proposed methodology provides a robust statistical tool to support public health surveillance, policy-making, and resource allocation.2026-09-09T13:43:05Z21 pages, 9 figures, 2 tablesRodrigo B. SilvaLuiza S. C. PiancastelliWagner Barreto-Souza