https://arxiv.org/api/RHXfXBZGxXIXqQTFhgMkbnvBduM2026-09-11T20:47:15Z377404515http://arxiv.org/abs/2609.10646v1Mixture-based Nonparametric Estimation of Spatial Covariance Functions with Applications to HIV Key Population Size Estimation across Sub-Saharan Africa2026-09-09T13:14:17ZConsistent data on the sizes of key populations, such as female sex workers (FSWs), are often scarce, particularly at the sub-national level. Accurate size estimates are critical to effectively allocate resources and achieve HIV targets. Since FSW population sizes may be spatially correlated across areas, models that account for spatial dependence can improve estimation. An important component of such models is the covariance function, which characterizes the spatial dependence structure of the underlying process. In this work, we study spatial covariance functions to estimate FSW population sizes in Sub-Saharan Africa (SSA). Many spatial models rely on parametric covariance functions. However, parametric estimation can suffer from model mis-specification, potentially leading to inefficient or biased predictions. We therefore develop a robust non-parametric approach for estimating the covariance function of a stationary isotropic process in $\mathbb{R}^d$. We focus on a class of covariance functions that are valid in all dimensions, which includes popular kernels such as the exponential and Matérn kernels. Leveraging the fact that such covariance functions can be represented as infinite mixtures of scaled Gaussian kernels, we propose two estimation methods: weighted least squares and nonparametric maximum likelihood estimation to estimate the mixing measure of scaled Gaussian kernels. We also develop computationally efficient methods to solve these optimization problems using non-negative least squares and second-order descent updates. We evaluate the proposed methods through simulations and apply them to estimate the FSW population sizes at the sub-national level in SSA.2026-09-09T13:14:17Z45 pages, 21 tables, 11 figuresManushi SiriwardanaHyebin SongLe BaoStephen Berghttp://arxiv.org/abs/2609.10028v1Dynamical Non-compensatory Multidimensional IRT Model Using Variational Approximation2026-09-09T11:01:34ZMultidimensional item response theory (MIRT) is a statistical test theory that precisely estimates multiple latent skills of learners from the responses in a test. Both compensatory and non-compensatory models have been proposed for MIRT: the former assumes that each skill can complement other skills, whereas the latter assumes they cannot. This non-compensatory assumption is convincing in many tests that measure multiple skills; therefore, applying non-compensatory models to such data is crucial for achieving unbiased and accurate estimation. In contrast to tests, latent skills will change over time in daily learning. To monitor the growth of skills, dynamical extensions of MIRT models have been investigated. However, most of them assumed compensatory models, and a model that can reproduce continuous latent states of skills under the non-compensatory assumption has not been proposed thus far. To enable accurate skill tracing under the non-compensatory assumption, we propose a dynamical extension of non-compensatory MIRT models by combining a linear dynamical system and a non-compensatory model. This results in a complicated posterior of skills, which we approximate with a Gaussian distribution by minimizing the Kullback-Leibler divergence between the approximated posterior and the true posterior. The learning algorithm for the model parameters is derived through Monte Carlo expectation maximization. Simulation studies verify that the proposed method is able to reproduce latent skills accurately, whereas the dynamical compensatory model suffers from significant underestimation errors. Furthermore, experiments on an actual data set demonstrate that our dynamical non-compensatory model can infer practical skill tracing and clarify differences in skill tracing between non-compensatory and compensatory models.2026-09-09T11:01:34ZHiroshi TamanoDaichi Mochihashihttp://arxiv.org/abs/2609.10020v1Cointegration by Parts: Locating Cointegration in Time2026-09-09T10:52:40ZTests for cointegration are typically applied to a single window spanning the entire sample, assuming that the long-run relationship holds throughout. When it holds over only a part of the sample, such tests lose power, because the stationary episode is diluted by periods without cointegration. We propose three statistics for testing whether two or more series cointegrate only over a part of the sample, each an infimum of the Engle-Granger statistic over recursive, backward-expanding, or doubly-flexible windows. We derive their limiting distributions and establish which alternatives each is consistent against. Only the doubly-flexible statistic has power against both break directions. Inspired by the seminal work of James G. MacKinnon, critical values are obtained by simulation and summarized through response surface regressions. We apply the tests to global mean sea level and global mean surface temperature anomalies. All three reject the null of no cointegration at the 5% level, locating it in a sub-period of the 1880--2019 record that coincides with documented discontinuities in sea surface temperature data collection.2026-09-09T10:52:40ZOlivia KvistJ. Eduardo Vera-Valdéshttp://arxiv.org/abs/2609.09981v1Optimal Value Inference for Reinforcement Learning2026-09-09T10:10:55ZWe study offline inference for the optimal value in reinforcement learning. Two new nuisances are derived as fixed points of a self-induced Bellman equation, in which we approximate the maximum Bellman operator by its softmax correspondence. We propose a debiased estimator through the Neyman orthogonality and establish its asymptotic normality under diverging horizons even when the behavior policy changes with time, as long as the nuisances have the statistical rates that can be achieved by many machine learning methods. We provide a concrete estimating procedure for these nuisances and show they can lead to valid inference. Synthetic experiments validate the numerical performance of our inference method, and we implement it in real-life decision-making problems, including bike repositioning and AI agentic tool use.2026-09-09T10:10:55ZNan LuEthan LeeJames M. RobinsDavid Simchi-LeviJunwei Luhttp://arxiv.org/abs/2601.21106v3Scalable Dirichlet Process Mixture Models with Unknown Concentration and Adaptive Covariance for High-Dimensional Clustering Applied to Leukemia Transcriptomics2026-09-09T09:33:09ZWe propose a novel method that performs adaptive clustering with DPMM using collapsed VI, while incorporating weakly-informative priors for DP concentration parameter alpha and base distribution G0. We illustrate the importance of G0 covariance structure and prior choice by considering different parameterisations of the data covariance matrix. On high-dimensional Gaussian simulations, our model demonstrates substantially faster convergence than a state-of-the-art MCMC splice sampler. We further evaluate performances on Negative Binomial simulations and conduct sensitivity analyses to assess robustness on realistic data conditions. Application to a publicly available leukemia transcriptomic data set comprising 72 samples and 2,194 gene expression successfully recovers every known sub-type, all while identifying additional gene expression-based sub-clusters with meaningful biological interpretation.2026-01-28T22:55:14Z22 pages with 5 figures and 1 tableAnnesh PalAguirre MimounRodolphe ThiébautBoris P. Hejblumhttp://arxiv.org/abs/1907.06994v2Regularized Estimation and Feature Selection in Mixtures of Generalized Linear Experts2026-09-09T08:53:00ZMixtures of experts (MoE) are conditional mixture models in which both the mixing proportions and the component densities depend on the predictors, and are widely used for regression, classification and model-based clustering of heterogeneous data. Fitting MoE by maximum likelihood becomes unstable, and sometimes infeasible, when the predictors are numerous or correlated. We propose a regularized maximum likelihood framework for simultaneous parameter estimation and feature selection in MoE whose experts belong to the generalized linear model family, covering Gaussian, Poisson and multinomial responses within a single formulation. Sparsity is induced in both the gating network and the experts through $\ell_1$ penalties, and the penalized log-likelihood is maximized by a proximal Newton-EM algorithm whose M-step reduces to weighted Lasso problems with closed-form coordinate-ascent updates. Unlike existing penalized MoE procedures, the algorithm requires neither a local quadratic approximation of the penalty nor any matrix inversion, it returns exactly sparse estimates without thresholding, and a proximal Newton-type variant guarantees a monotone increase of the penalized objective at every iteration. On simulated data and five real data sets, the method recovers the actual sparsity support and delivers prediction and clustering accuracy that is competitive with, and often better than, state-of-the-art regularized MoE. The source codes of our developed algorithms and their documentation are publicly available on Github at https://github.com/nv-thin/GLM-RMoE.2019-07-14T10:58:31ZThin Nguyen-VanFaicel ChamroukhiHa Hoang VanBao Tuyen Huynhhttp://arxiv.org/abs/2609.09831v1Proportional-limit asymptotics for Diaconis-Ylvisaker-penalised logistic regression with fitted intercept2026-09-09T07:39:36ZThis paper develops estimator-level asymptotic theory for maximum Diaconis-Ylvisaker prior penalised likelihood for logistic regression with a jointly fitted intercept and nonzero prior slope direction in the proportional-limit regime. For $\mathrm{N}(\mathbf{0}_p, p^{-1}\mathbf{I}_p)$ Gaussian covariates and $p/n\toκ\in(0,1)$, a conditional convex Gaussian min--max analysis yields almost-sure convergence of the fitted intercept and a pseudo-Lipschitz empirical law for the slope estimator. This estimator-level law gives asymptotic limits for out-of-sample scores, classification error, optimal thresholding and oracle calibration. It also yields oracle-adjusted fixed-block $Z$-statistics under isotropic Gaussian covariates, and yields the main ingredient in establishing the limiting distribution of the penalised likelihood-ratio test statistic and identifies the rescaling to recover the nominal chi-square distribution. We extend these results to Gaussian designs with arbitrary deterministic mean and positive-definite covariance via affine centering and whitening and discuss extensions to subgaussian covariates. Finally, we propose a consistent response-moment estimator of the oracle parameters entering the state equations that govern the slope limiting law and are required for feasible inference.2026-09-09T07:39:36ZPhilipp Sterzingerhttp://arxiv.org/abs/2609.08097v2Nonparametric heterogeneous causal mediation with orthogonal machine learning2026-09-09T04:17:10ZCausal mediation analysis decomposes the total effect of an intervention on an outcome into a direct pathway and an indirect pathway transmitted through a mediator, but standard methods typically summarize these pathways using population average effects. In many applications, however, the indirect effect may vary substantially across individual profiles. We propose an orthogonal statistical learning framework for estimating heterogeneous causal mediation effects conditional on individual characteristics. The method constructs a class of weighted Neyman orthogonal losses motivated by influence function representations of weighted population average effects. These losses directly target conditional mediation estimands whose minimizers are locally insensitive to nuisance estimation errors. We implement the resulting learners under a two-stage meta-learning framework with regularized linear sieves as second-stage smoothers, and introduce a combination of targeted learning and orthogonal learning designed to improve stability when mediator density ratios are unstable. We establish $L^2$ and uniform limit theory and develop pointwise and uniform confidence bands. Simulation studies show that the proposed orthogonal learners reduce the mean integrated squared error by more than $50\%$ compared with existing model-based methods and provide computationally efficient inference in nonlinear settings. The CARDIA, PSACR, and STAR analyses reveal heterogeneous mediated effects across cardiometabolic, psychological, and educational settings.2026-09-08T01:15:31ZJiaqi TongYi ZhaoBhramar MukherjeeFan Lihttp://arxiv.org/abs/2609.09600v1Sliced $L^p$ Distributional Balancing2026-09-09T01:53:02ZA popular class of causal inference methods addresses confounding through weighting, which reweights treated and control groups to balance their covariate distributions without using outcome information, thereby preserving a design-based perspective. In this paper, we propose sliced $L^p$ distributional balancing (SLDB), a family indexed by $p\in[1,\infty)$ that measures imbalance by averaging squared $L^p$ distances between the cumulative distribution functions of one-dimensional linear projections. The Cramér--Wold device ensures that this criterion identifies equality of multivariate distributions, while projection reduces its computation to sorting-based one-dimensional operations. Because our method lies outside the maximum mean discrepancy (MMD) framework underlying many existing distributional balancing methods, their theoretical and computational tools do not directly apply. We therefore develop a computationally efficient projected subgradient descent algorithm for estimating balancing weights, offering improved computational complexity over MMD-based methods. Furthermore, we establish a novel theoretical framework for SLDB-based causal effect estimation and prove, under suitable conditions, $\sqrt{n}$-consistency and asymptotic normality, with the asymptotic variance attaining the semiparametric efficiency bound. Finally, we develop inferential procedures that do not require augmentation with an outcome model, thereby retaining the design-based principle. Simulation studies and a real-world application demonstrate that SLDB performs competitively with existing methods.2026-09-09T01:53:02Z34 pagesHaoran ZhangGuanhua ChenChan Parkhttp://arxiv.org/abs/2504.13124v2False Discovery Controlled Regions for Excursion Sets2026-09-09T01:47:27ZIdentifying areas where the signal is prominent is an important task in image analysis, with particular applications in brain mapping. In this work, we develop False Discovery Controlled Regions (FDCRs) for excursion sets, the set of locations where the values are above or below a given level. We achieve this by treating the confidence procedure as a testing problem at the given level, allowing control of the False Discovery Rate (FDR). Methods are developed to control the FDR, separately for positive and negative excursions, as well as jointly over both. Furthermore, power is increased by incorporating a two-stage adaptive procedure. Simulation results in a variety of different settings show that the FDCRs successfully control the FDR under the nominal alpha level. We showcase our methods with an application to functional magnetic resonance imaging (fMRI) data from the Human Connectome Project illustrating the improvement in statistical power over existing approaches.2025-04-17T17:41:05ZHowon RyuThomas Maullin-SapeyArmin SchwartzmanSamuel Davenporthttp://arxiv.org/abs/2609.09587v1A binary factor model2026-09-09T01:18:20ZThe orthogonal factor model has been a very useful tool in uncovering covariance structures in a set of variables through a smaller set of underlying factors. This old model is suitable for continuous variables with unbounded support, since the most common assumption for the observables and the factors is multivariate normality. In this work, we propose a factor model for binary data. Factors are negative dependent, so they avoid each other. We study the theoretical properties of the model and carry out a full Bayesian inference. We illustrate the performance of our proposal with simulated and real data sets and compare with the traditional benchmark.2026-09-09T01:18:20ZLuis E. Nieto-Barajashttp://arxiv.org/abs/2609.09544v1When is statistical evidence strong enough? Using hypothesis tests to value data collection2026-09-08T23:59:49ZWe recast statistical significance as a choice between making an immediate policy recommendation and deferring it until further evidence is collected. We show that the welfare-optimal decision corresponds, under minimax regret, to a statistical test whose level depends on the cost and precision of additional evidence. Inverting this rule, we introduce and recommend reporting the abstention-value (A-value) alongside traditional p-values to determine where additional data collection is most needed. The A-value defines the break-even welfare cost of abstaining and recommending further experimentation given the initial evidence. When experimentation capacity is limited, prioritizing additional data collection where A-values are the largest yields finite-sample welfare guarantees. We illustrate its implications for economic program evaluation.2026-09-08T23:59:49ZAristotelis EpanomeritakisDavide Vivianohttp://arxiv.org/abs/2510.17072v2DFNN: A Deep Fréchet Neural Network Framework for Learning Metric-Space-Valued Responses2026-09-08T23:58:03ZRegression with non-Euclidean responses---e.g., probability distributions, networks, symmetric positive-definite matrices, and compositions---has become increasingly important in modern applications. In this paper, we propose deep Fréchet neural networks (DFNNs), an end-to-end deep learning framework for predicting non-Euclidean responses---which are considered as random objects in a metric space---from Euclidean predictors. Our method utilizes the representation-learning power of deep neural networks (DNNs) to the task of approximating conditional Fréchet means of the response given the predictors, the metric-space analogue of conditional expectations, by minimizing a Fréchet risk. The framework is highly flexible, accommodating diverse metrics and high-dimensional predictors. We establish a universal approximation theorem for DFNNs, advancing the state-of-the-art of neural network approximation theory to general metric-space-valued responses, without making model assumptions or relying on local smoothing. We further establish rigorous generalization guarantees for DFNNs and derive corresponding risk bounds, providing, to the best of our knowledge, the first such theoretical results for deep learning regression with metric-space-valued responses. Empirical studies on synthetic distributional and network-valued responses, as well as real-world applications to predicting compositional responses in an Aitchison simplex and spherical responses, demonstrate that DFNNs consistently outperform all existing methods.2025-10-20T00:57:30ZKyum KimYaqing ChenParomita Dubeyhttp://arxiv.org/abs/2609.09536v1Differentially Private Average Treatment Effect Estimation by Propensity Score Blocking2026-09-08T23:30:41ZAverage treatment effect (ATE) estimation in observational studies is a fundamental statistical tool used frequently in social science, medicine, and other fields. These fields often work with sensitive data where privacy protections are important, so a differentially private mechanism for ATE estimation is highly desirable. Here we present two propensity score-based algorithms for ATE estimation on observational data, one improving the inverse probability weighting (IPW) method used in prior work, and the other using blocking on the propensity score (BPS). Both show lower error and less bias than prior work, with the BPS-based algorithm frequently reducing error by 75% or more compared to prior work.2026-09-08T23:30:41ZDuncan StewardsonGrayson W. WhiteAdam Grocehttp://arxiv.org/abs/2608.07979v2Sparse departures from independence in two-way tables: a heteroscedasticity profile and detection boundary, an adaptive higher-criticism gate, and an assumption-lean exact anchor2026-09-08T23:13:09ZTwo-way contingency tables are tested for independence throughout applied statistics (genomics, network and text co-occurrence, pharmacovigilance, ecology, survey cross-tabulation), and the routine test reads asymptotic Gaussian tail probabilities off the table cell by cell. On the tables people actually analyze this is badly miscalibrated: across 5,543 real public two-way tables, two-thirds have counts small enough or margins heterogeneous enough that the asymptotic maximum-cell independence scan false-positives at a mean 33% against a 0.05 target, while an exact margin-conditional anchor holds at about 1%. The failure is sharpest when the departure is sparse, the association concentrated in a few cells, where the omnibus chi-square is inefficient and the sharp object is the detection boundary. In three parts we answer where a sparse signal can be seen, which combiner attains that limit, and how to calibrate at small counts. Part I specializes the sparse-detection theory of Ingster (1997), Donoho and Jin (2004), and Chhor, Mukherjee and Sen (2024) to independence: the table is a heteroscedastic Gaussian sequence whose variance profile is the expected-count table, fixed by the margins via CVe^2 = (1 + CVr^2)(1 + CVc^2) - 1, giving an exact signal map and a closed-form separation radius valid under a moderate-deviation growing-count condition. Part II shows higher criticism, with its closed-form Jaeschke-Eicker null, attains that boundary adaptively over unknown sparsity. Part III measures the small-count calibration failure (size 0.3 to 0.6 even under uniform margins, so the driver is the per-cell tail), removes it with exact margin-conditional laws (size 0.003 to 0.009), and gives a routing rule sending large-count tables to the asymptotic gate and small-count tables to the exact anchor. Every claim is reproduced from openly deposited, deterministically seeded code.2026-08-08T07:28:47Z29 pages, 11 figures, 4 tables. Reproducibility package: doi:10.5281/zenodo.21844797William J. Dwyer