https://arxiv.org/api/PGvqVOfy8oE2A0o7DYTIQsgH7v02026-09-10T16:34:58Z37719015http://arxiv.org/abs/2609.10534v1Likelihood-free inference with nuisance parameters through normalizing flows2026-09-09T17:58:15ZWe present a simple decomposition of a neural-network-based normalizing flow that naturally uncovers a pivotal statistic (or something close) in the presence of nuisance parameters, based only on a sample generator from the distribution of interest. We show that the statistic is near-pivotal in the sense of minimum average KL-divergence of its $p$-values versus uniform and we argue that it can be expected to have good power when the dimension of the statistic equals the dimension of the parameter. It is able to incorporate prior knowledge about group invariances such as translation and scale. It can discover the one-sample $t$-test almost exactly, outperforms the Welch test in terms of worst-case size over a constrained variance-ratio range and achieves good calibration on partial biserial correlations, while showing higher power (and being much faster) on small-to-moderate samples than profile likelihood-ratio techniques.2026-09-09T17:58:15Z49 pages and 13 figures, including appendices. Code available at https://github.com/philassheton/NeuralCIsPhil Asshetonhttp://arxiv.org/abs/2602.20115v2Compound decisions and empirical Bayes via Bayesian nonparametrics2026-09-09T17:38:31ZWe study compound decision theory from a nonparametric Bayesian perspective, with particular emphasis on their relationship to empirical Bayes (EB) procedures. Motivated by the sharp risk guarantees available for EB procedures based on the nonparametric maximum likelihood estimator (NPMLE), we investigate whether analogous guarantees can be established for fully Bayesian decision rules. In a class of Gaussian compound decision problems, we show that the fully Bayesian posterior mean achieves near-optimal risk. Moreover, it is admissible as a genuine Bayes rule, whereas the corresponding NPMLE plug-in rule is inadmissible. Simulations illustrate the performance of nonparametric Bayes procedures relative to common alternatives. As an application, we apply our methodology to Census tract-level estimates of economic mobility from the Opportunity Atlas.2026-02-23T18:33:57Z69 pagesNikolaos IgnatiadisSid Kankanalahttp://arxiv.org/abs/2608.27599v2Activity-Conditioned Residual Association from Aggregated Relational Data2026-09-09T16:52:48ZAggregated relational data (ARD) record how many ties sampled respondents have to prespecified groups without revealing individual dyads. We develop a conditional test for prespecified cross-group concordance beyond an additive-activity network model. When the groups form an exhaustive partition, respondent degree is observed exactly. For arbitrary fixed activity values under an independent-Bernoulli additive-logit null, conditioning two respondents on equal degree yields a finite nonpositive sign restriction for their two disjoint group-count differences. Analyst-randomized pair thinning gives a conservative baseline. Under a stronger fixed-cell design with group-specific smooth activity profiles and positive cross-role overlap, an observable clipped full-pair statistic gives conservative one-network inference with graph-independent sampled respondents. Finite and growing matched-law constructions establish that the target contains bivariate information absent from a collapsed binary count. The full-pair procedure is uniformly consistent on a primitive relative-open neighborhood of a specified rank-one alternative. Fixed-dimensional lattice local limits make the conditioning and availability requirements explicit.2026-08-27T18:30:58Z40 pages, 2 tables. Expanded version with full-pair inference, complete proofs, and additional numerical evidenceYen-hsuan Tsenghttp://arxiv.org/abs/2501.10675v3Recovering Unobserved Network Links from Aggregated Relational Data: Bayesian Latent Surface Modeling and Penalized Regression2026-09-09T16:35:52ZAggregated relational data (ARD) record counts of ties to attribute-defined groups while leaving individual edges unobserved. We compare latent-geometry and regularized network estimators through a common observation map. The comparison distinguishes the realized adjacency matrix, conditional edge probabilities, and model parameters. We study roster-based ARD with known node-level group memberships, giving both estimators the same roster and aggregate counts.
We relate the aggregate means to a Poisson working likelihood and a Huber loss, and give their derivatives. Overlapping groups, shared edges, and reporting error affect the interpretation of these objectives. Geometry restricts the representation of edge probabilities, while regularization selects among candidate fits. Identification depends on the observation map and model restrictions rather than uniqueness of a numerical optimizer.
A reproducible synthetic experiment specifies the data-generating process, estimation algorithms, and evaluation targets under matched information. The matrix estimator gives better realized-edge rankings and aggregate fit, while the geometric estimator gives lower error for generating probabilities. The resulting framework organizes ARD reconstruction around the interaction of observation design, structural assumptions, and computation.2025-01-18T06:51:51Z14 pages, 2 figures. Substantially revised replacement of the withdrawn version. Clarified observation regime and targets; corrected likelihood and loss calculations; added a reproducible matched-input synthetic experiment. Code and saved results are included as ancillary filesYen-hsuan Tsenghttp://arxiv.org/abs/2609.10409v1Dynamic prediction intervals for survival times2026-09-09T16:30:09ZMost work on survival prediction focuses on estimating survival probabilities rather than predicting individual event times. Recent conformal methods have made it possible to construct prediction intervals for survival times with right-censored outcomes, but existing approaches are restricted to settings with covariates only measured at baseline and do not address dynamic prediction with longitudinal data. We study prediction intervals for survival times in a dynamic prediction framework with longitudinal covariates. Our approach uses Penalized Regression Calibration (PRC) as a working dynamic prediction model, combining linear mixed models for the longitudinal histories with a Cox model for post-landmark survival, and then applies a conformal calibration step to obtain prediction intervals. We compare naive intervals obtained by direct inversion of the survival function estimated by PRC to our dynamic conformal method. A Monte Carlo simulation study evaluates empirical coverage and interval length across sample sizes, censoring levels, landmark times, and non-proportional hazards (NPH) scenarios. We illustrate the proposed methodology by computing dynamic prediction intervals for the time until a dementia diagnosis in the ADNI dataset. The results show that naive inversion is often unreliable, whereas the proposed dynamic conformal method yields more stable predictive performance.2026-09-09T16:30:09Z20 pages. Supplementary material includedLorenzo CarvisigliaSaverio RanciatiMirko Signorellihttp://arxiv.org/abs/2506.09986v3Constrained Denoising, Empirical Bayes, and Optimal Transport2026-09-09T16:04:33ZIn latent variables models, two important goals are denoising and deconvolution: denoising aims to estimate the latent variables, whereas deconvolution aims to estimate the distribution of the latent variables. As has been recognized in the literature over the last century, these two goals are fundamentally in tension, since denoising yields a poor estimate of the distribution of the latent variables due to shrinkage, and deconvolution yields a distribution-valued estimate that carries no unit-specific information. In this paper, we provide a systematic study of denoisers, and empirical Bayes approximations thereof, which attain optimal denoising error subject to the constraint that the distribution of the denoised data matches, in some sense, the distribution of the latent variables. Our insight is that optimal transport allows practitioners to navigate the tension between denoising and deconvolution. More precisely, we propose a modular methodology that combines any suitable unconstrained empirical Bayes denoiser (arising, e.g., via $F$-modeling, $G$-modeling, conjugate-parametric models) with any suitable information about the distribution of the latent variables (e.g., its moments, support, or an approximation of the entire distribution via deconvolution) into a single denoised data set. We prove explicit rates of convergence for our proposed methodologies, and we apply the resulting methods in applications in astronomy, baseball analytics, and marketing.2025-06-11T17:57:17Z72 pages, 6 figures, 1 Table. Comments welcomeAdam Quinn JaffeNikolaos IgnatiadisBodhisattva Senhttp://arxiv.org/abs/2609.10313v1Semiparametric Inference for Conditional Shapley Feature Importance2026-09-09T15:19:24ZShapley values are widely used for post-hoc feature attribution, but most estimators return point quantities and do not quantify uncertainty, and popular implementations sample out-of-coalition features from their marginal distribution, which misattributes importance when features are dependent. This paper studies the conditional formulation, in which out-of-coalition features are integrated out under their true conditional distribution. The target is a global, loss-based importance that pairs a conditional value function with a SAGE-style loss aggregation. We propose a one-step estimator with K-fold cross-fitting and a U-statistic correction of the squared loss that removes the Monte Carlo bias of the naive plug-in; it is $\sqrt{n}$-consistent and asymptotically normal under double-robust rate conditions, and the resulting Wald interval attains nominal coverage. A Pinsker-type bound quantifies the bias from misspecifying the working copula class, while vine copulas keep conditional sampling tractable. In a Gaussian design study with n= 500, the empirical coverage of the 95% interval lies between 0.91 and 0.96 across all features, the test holds its Type-I rate at 0.05, and it reaches power one for moderate signals. Applied to the UCI Concrete and California Housing data, the method identifies the conditionally informative features with Bonferroni-controlled significance.2026-09-09T15:19:24ZAgostino Gnassohttp://arxiv.org/abs/2608.09218v3Online Learning of Scale Parameters in Score-Driven Filters2026-09-09T14:53:06ZA score-driven filter multiplies its scaled log-likelihood score by a scale parameter. We call this coefficient the gain and learn it online. Given the current state and realised scaled score, each admissible gain selects a reachable next state and predictive density. A scalar gain moves along a line; diagonal gains control coordinatewise transmission and may change direction. We evaluate gain selection using a one-step predictive Kullback--Leibler objective. In the scalar unscaled case, the negative consecutive-score product is a stochastic gradient; the positive product used in accelerated recursions is a descent direction. Positive scalar score scaling changes only the effective learning rate. Monotone differentiable gain links induce mirror-descent geometry, while persistence adds a Bregman pull towards a reference gain. Under convexity, compactness, integrability, and schedule conditions, projected and discounted mirror updates satisfy dynamic-regret bounds relative to time-varying, current-information comparators. Simulations isolate score scaling, link geometry, persistence, and coordinatewise gains. Across twelve equity indices, the bounded discounted-logistic gain records a lower out-of-sample mean negative log score than the constant gain in eleven markets, although market-level evidence is mixed. It also avoids the extreme transients of the numerically capped exponential-link benchmark. Improvements are largest in markets spanning multiple crises.2026-08-10T07:43:38Z62 pages, 10 figures, 13 tablesFabrizio LilloGiulia LivieriGianluca Palmarihttp://arxiv.org/abs/2609.10174v1A spatiotemporal negative binomial model with dynamic dispersion: An application to Tuberculosis infections2026-09-09T13:43:05ZTuberculosis (TB) remains a critical public health concern in Brazil, characterized by pronounced spatial heterogeneity and fluctuating temporal volatility. In this paper, we study monthly TB notifications across 61 microregions of Sao Paulo state from 2001 to 2024. To do this, we introduce a negative binomial spatial integer-valued generalized autoregressive conditional heteroskedastic (INGARCH) model featuring jointly dynamic conditional means and time-varying dispersion. To capture inter-regional spillovers, we incorporate both discrete adjacency structures and a novel continuous distance-based formulation leveraging the Matern correlation function. Parameter estimation via conditional maximum likelihood employs a two-step profile-likelihood iterative scheme, demonstrating solid finite-sample performance in simulation studies. Applied to the Sao Paulo TB surveillance data, the framework substantially outperforms standard Poisson and fixed-dispersion spatiotemporal baselines in empirical fit and uncertainty quantification, maintaining nominal 95% predictive coverage across both dense metropolitan centers and rural microregions. Our results reveal marked spatial heterogeneity in baseline incidence, dynamic overdispersion driven by localized outbreaks, and short-range spatial interaction decay. By accurately modeling spatiotemporal volatility, the proposed methodology provides a robust statistical tool to support public health surveillance, policy-making, and resource allocation.2026-09-09T13:43:05Z21 pages, 9 figures, 2 tablesRodrigo B. SilvaLuiza S. C. PiancastelliWagner Barreto-Souzahttp://arxiv.org/abs/2609.10028v1Dynamical Non-compensatory Multidimensional IRT Model Using Variational Approximation2026-09-09T11:01:34ZMultidimensional item response theory (MIRT) is a statistical test theory that precisely estimates multiple latent skills of learners from the responses in a test. Both compensatory and non-compensatory models have been proposed for MIRT: the former assumes that each skill can complement other skills, whereas the latter assumes they cannot. This non-compensatory assumption is convincing in many tests that measure multiple skills; therefore, applying non-compensatory models to such data is crucial for achieving unbiased and accurate estimation. In contrast to tests, latent skills will change over time in daily learning. To monitor the growth of skills, dynamical extensions of MIRT models have been investigated. However, most of them assumed compensatory models, and a model that can reproduce continuous latent states of skills under the non-compensatory assumption has not been proposed thus far. To enable accurate skill tracing under the non-compensatory assumption, we propose a dynamical extension of non-compensatory MIRT models by combining a linear dynamical system and a non-compensatory model. This results in a complicated posterior of skills, which we approximate with a Gaussian distribution by minimizing the Kullback-Leibler divergence between the approximated posterior and the true posterior. The learning algorithm for the model parameters is derived through Monte Carlo expectation maximization. Simulation studies verify that the proposed method is able to reproduce latent skills accurately, whereas the dynamical compensatory model suffers from significant underestimation errors. Furthermore, experiments on an actual data set demonstrate that our dynamical non-compensatory model can infer practical skill tracing and clarify differences in skill tracing between non-compensatory and compensatory models.2026-09-09T11:01:34ZHiroshi TamanoDaichi Mochihashihttp://arxiv.org/abs/2609.10020v1Cointegration by Parts: Locating Cointegration in Time2026-09-09T10:52:40ZTests for cointegration are typically applied to a single window spanning the entire sample, assuming that the long-run relationship holds throughout. When it holds over only a part of the sample, such tests lose power, because the stationary episode is diluted by periods without cointegration. We propose three statistics for testing whether two or more series cointegrate only over a part of the sample, each an infimum of the Engle-Granger statistic over recursive, backward-expanding, or doubly-flexible windows. We derive their limiting distributions and establish which alternatives each is consistent against. Only the doubly-flexible statistic has power against both break directions. Inspired by the seminal work of James G. MacKinnon, critical values are obtained by simulation and summarized through response surface regressions. We apply the tests to global mean sea level and global mean surface temperature anomalies. All three reject the null of no cointegration at the 5% level, locating it in a sub-period of the 1880--2019 record that coincides with documented discontinuities in sea surface temperature data collection.2026-09-09T10:52:40ZOlivia KvistJ. Eduardo Vera-Valdéshttp://arxiv.org/abs/2609.09981v1Optimal Value Inference for Reinforcement Learning2026-09-09T10:10:55ZWe study offline inference for the optimal value in reinforcement learning. Two new nuisances are derived as fixed points of a self-induced Bellman equation, in which we approximate the maximum Bellman operator by its softmax correspondence. We propose a debiased estimator through the Neyman orthogonality and establish its asymptotic normality under diverging horizons even when the behavior policy changes with time, as long as the nuisances have the statistical rates that can be achieved by many machine learning methods. We provide a concrete estimating procedure for these nuisances and show they can lead to valid inference. Synthetic experiments validate the numerical performance of our inference method, and we implement it in real-life decision-making problems, including bike repositioning and AI agentic tool use.2026-09-09T10:10:55ZNan LuEthan LeeJames M. RobinsDavid Simchi-LeviJunwei Luhttp://arxiv.org/abs/2601.21106v3Scalable Dirichlet Process Mixture Models with Unknown Concentration and Adaptive Covariance for High-Dimensional Clustering Applied to Leukemia Transcriptomics2026-09-09T09:33:09ZWe propose a novel method that performs adaptive clustering with DPMM using collapsed VI, while incorporating weakly-informative priors for DP concentration parameter alpha and base distribution G0. We illustrate the importance of G0 covariance structure and prior choice by considering different parameterisations of the data covariance matrix. On high-dimensional Gaussian simulations, our model demonstrates substantially faster convergence than a state-of-the-art MCMC splice sampler. We further evaluate performances on Negative Binomial simulations and conduct sensitivity analyses to assess robustness on realistic data conditions. Application to a publicly available leukemia transcriptomic data set comprising 72 samples and 2,194 gene expression successfully recovers every known sub-type, all while identifying additional gene expression-based sub-clusters with meaningful biological interpretation.2026-01-28T22:55:14Z22 pages with 5 figures and 1 tableAnnesh PalAguirre MimounRodolphe ThiébautBoris P. Hejblumhttp://arxiv.org/abs/1907.06994v2Regularized Estimation and Feature Selection in Mixtures of Generalized Linear Experts2026-09-09T08:53:00ZMixtures of experts (MoE) are conditional mixture models in which both the mixing proportions and the component densities depend on the predictors, and are widely used for regression, classification and model-based clustering of heterogeneous data. Fitting MoE by maximum likelihood becomes unstable, and sometimes infeasible, when the predictors are numerous or correlated. We propose a regularized maximum likelihood framework for simultaneous parameter estimation and feature selection in MoE whose experts belong to the generalized linear model family, covering Gaussian, Poisson and multinomial responses within a single formulation. Sparsity is induced in both the gating network and the experts through $\ell_1$ penalties, and the penalized log-likelihood is maximized by a proximal Newton-EM algorithm whose M-step reduces to weighted Lasso problems with closed-form coordinate-ascent updates. Unlike existing penalized MoE procedures, the algorithm requires neither a local quadratic approximation of the penalty nor any matrix inversion, it returns exactly sparse estimates without thresholding, and a proximal Newton-type variant guarantees a monotone increase of the penalized objective at every iteration. On simulated data and five real data sets, the method recovers the actual sparsity support and delivers prediction and clustering accuracy that is competitive with, and often better than, state-of-the-art regularized MoE. The source codes of our developed algorithms and their documentation are publicly available on Github at https://github.com/nv-thin/GLM-RMoE.2019-07-14T10:58:31ZThin Nguyen-VanFaicel ChamroukhiHa Hoang VanBao Tuyen Huynhhttp://arxiv.org/abs/2609.09831v1Proportional-limit asymptotics for Diaconis-Ylvisaker-penalised logistic regression with fitted intercept2026-09-09T07:39:36ZThis paper develops estimator-level asymptotic theory for maximum Diaconis-Ylvisaker prior penalised likelihood for logistic regression with a jointly fitted intercept and nonzero prior slope direction in the proportional-limit regime. For $\mathrm{N}(\mathbf{0}_p, p^{-1}\mathbf{I}_p)$ Gaussian covariates and $p/n\toκ\in(0,1)$, a conditional convex Gaussian min--max analysis yields almost-sure convergence of the fitted intercept and a pseudo-Lipschitz empirical law for the slope estimator. This estimator-level law gives asymptotic limits for out-of-sample scores, classification error, optimal thresholding and oracle calibration. It also yields oracle-adjusted fixed-block $Z$-statistics under isotropic Gaussian covariates, and yields the main ingredient in establishing the limiting distribution of the penalised likelihood-ratio test statistic and identifies the rescaling to recover the nominal chi-square distribution. We extend these results to Gaussian designs with arbitrary deterministic mean and positive-definite covariance via affine centering and whitening and discuss extensions to subgaussian covariates. Finally, we propose a consistent response-moment estimator of the oracle parameters entering the state equations that govern the slope limiting law and are required for feasible inference.2026-09-09T07:39:36ZPhilipp Sterzinger