https://arxiv.org/api/6C+SFfJvCKaeN9Vm1yngTioZs402026-10-02T18:46:14Z1067522515http://arxiv.org/abs/2504.13291v3Estimating equations for causal survival analysis with pooled logistic regression2026-08-27T16:35:41ZBackground: Pooled logistic regression models are commonly applied in survival analysis. However, the standard implementation can be computationally demanding, which is further exacerbated when using the nonparametric bootstrap for inference. To ease these computational burdens, investigators often coarsen time intervals or assume a parametric functional form for time. These approaches impose restrictive assumptions, which may not always have a well-motivated substantive justification. Methods: Here, the pooled logistic regression model is re-framed using estimating equations to simplify computations and allow for inference via the empirical sandwich variance estimator, thus avoiding the more computationally demanding bootstrap. The proposed implementation is demonstrated using two examples with publicly available data. The performance of the empirical sandwich variance estimator is illustrated using a Monte Carlo simulation study. Results: As shown in the applied examples, the proposed implementation substantially reduced run-times and could be applied without needing to coarsen the data. In the simulation study, the empirical sandwich variance estimator results in nominal confidence interval coverage. Conclusions: The implementation proposed here offers a less computationally demanding alternative to the standard implementation of pooled logistic regression without needing to impose restrictive constraints on time.2025-04-17T19:04:07ZPaul N ZivichStephen R ColeBonnie E Shook-SaJustin B DeMonteJessie K Edwardshttp://arxiv.org/abs/2608.27063v1An Accurate and Single-Communication Federated Inference Algorithm2026-08-27T12:49:56ZJoint analyses across multiple institutions are increasingly important in biomedical and epidemiological research, particularly for rare diseases where datasets are typical small. However, privacy regulations and institutional policies often prevent the sharing of individual-level patient data. In this paper we present an accurate and single-communication federated inference algorithm. Single-communication federated inference enables statistical analyses through a single exchange of summary statistics between participating centers and a coordinating server, preserving privacy while reducing communication and computational costs compared with iterative federated learning. We extend a recently proposed single-communication federated inference strategy that is based on second-order Taylor expansions by using third-order expansions to better approximate local log-likelihood functions. The proposed method is evaluated through simulation studies based on real data and compared with existing federated inference strategies. The simulation studies assess the performance of the proposed method, with a particular focus on scenarios involving small local sample sizes, where quadratic approximations may fail to capture skewness and other higher-order characteristics of the log-likelihood function. They demonstrate that incorporating higher-order information of the log-likelihood function improves the accuracy while preserving the privacy, communication efficiency, and scalability required for collaborative biomedical and epidemiological research.2026-08-27T12:49:56ZLaura MontagnaniAnthony CC CoolenMarianne A Jonkerhttp://arxiv.org/abs/2608.18020v2Context-adjusted Player Evaluation for Twenty20 Cricket2026-08-27T08:54:18ZI develop a reproducible framework for evaluating individual batting and bowling performances in Twenty20 (T20) cricket on one interpretable scale of runs above expectation, built from two ball-level primitives. The first, Runs Above Expected (RAE), is the residual between the runs scored on a delivery and a contextual expectation of what an average performer would produce in the same situation. That expectation is a multiplicative log-linear model of the cohort scoring rate, whose per-cell estimator is shown to be a conditional Poisson maximum likelihood multiplier. It is fitted by iterated backfitting and stabilised by empirical Bayes shrinkage, so that thinly sampled contexts are pooled towards the population. A single opposition symmetry places run-scoring and run-prevention on the same footing. The second primitive, \emph{Dismissal Adjusted Runs} (DAR), prices a dismissal in that currency as the runs it forgoes, read off a batting side value function solved by dynamic programming through a Bellman expectation recursion, under observed play. Because dismissal is the expected end of every innings, the realised wicket cost (realDAR) is centered against its expectation (xDAR) under a league dismissal hazard rate. The centered quantity is a run-weighted mean zero martingale residual of the dismissal process, so a player is charged only for departing from average behaviour. The two primitives sum to a symmetric Impact, which makes batting and bowling comparable in centre as well as in unit. Estimated on over 2.7 million legal deliveries of men's T20 cricket, the framework recovers known contextual structure, agrees with the conventional rates it refines while correcting their context-blindness, and yields face-valid player, innings and season leaderboards for the Indian Premier League.2026-08-18T17:12:55ZMain paper: 23 pages, 13 tables, 5 figures; Supplementary materials: 9 pages, 18 tablesRhitankar Bandyopadhyayhttp://arxiv.org/abs/2603.02574v3An Augmented Rating System for Test cricket: adapting the Glicko rating system2026-08-27T08:03:20ZThe International Cricket Council's (ICC) Test cricket ratings use match and series results alone, stating no allowance about the precision of a rating, or about home advantage and toss. We fold both into the Glicko rating system, which gives a probabilistic expected score, and recalibrate its scale for Test cricket. The two enter as covariates whose significance and weights are estimated from match data. Home advantage is worth approximately 13 rating points and the toss roughly 8. They add, with no significant interaction. Over the two completed World Test Championship cycles (2021-23 and 2023-25) the model predicts 77.6% of decisive matches correctly in the first, and matches whichever of standard Elo and unmodified Glicko rating system is more accurate, at a lower Brier score and log loss in both. Permuting each cycle's match order 1000 times leaves every final rating unchanged. The model shows robustness to ordering, not fairness towards the fixture list, which the unbalanced calendar prevents judging. Our ordering follows the ICC's (Spearman 0.979 and 0.983). We add not a different ranking but a calibrated win probability per match, a deviation on every rating, and home and toss adjustments of estimated size, none of which the ICC supplies.2026-03-03T03:45:50Z26 pages, 14 tables, 0 figuresRhitankar BandyopadhyayDiganta Mukherjeehttp://arxiv.org/abs/2601.21217v2A Flexible Empirical Bayes Approach to Generalized Linear Models, with Applications to Sparse Logistic Regression2026-08-27T03:48:14ZWe introduce a flexible empirical Bayes approach for fitting Bayesian generalized linear models. Specifically, we adopt a novel mean-field variational inference (VI) method and the prior is estimated within the VI algorithm, making the method tuning-free. Unlike traditional VI methods that optimize the posterior density function, our approach directly optimizes the posterior mean and prior parameters. This formulation reduces the number of parameters to optimize and enables the use of scalable algorithms such as L-BFGS and stochastic gradient descent. Furthermore, our method automatically determines the optimal posterior based on the prior and likelihood, distinguishing it from existing VI methods that often assume a Gaussian variational. Our approach represents a unified framework applicable to a wide range of exponential family distributions, removing the need to develop unique VI methods for each combination of likelihood and prior distributions. We apply the framework to solve sparse logistic regression and demonstrate the superior predictive performance of our method in extensive numerical studies, by comparing it to prevalent sparse logistic regression approaches.2026-01-29T03:31:49ZDongyue XieMatthew Stephenshttp://arxiv.org/abs/2209.03957v2A Unified Statistical Procedure to Analyse Irreversible Thermal Curves2026-08-27T00:05:12ZDNA hybridisation experiments are crucial for studying the thermodynamic and kinetic profiles of various systems in nucleic acid chemistry. The phenomenon of hysteresis is commonly observed in many such UV thermal experiments involving unmodified or modified nucleic acids. In the presence of hysteresis, the thermal curves are irreversible and demand a significant effort to produce the reaction-specific kinetic and thermodynamic parameters. In this article, we describe a unified statistical procedure to analyse such thermal curves. More specifically, the proposed method allows one to handle the thermal curves for the formation of duplexes, triplexes, and various quadruplexes in exactly the same way. The proposed method uses a local polynomial regression to find the smoothed thermal curves and calculate their slopes. This method is more flexible and easier to implement than the least squares polynomial smoothing, which is currently almost universally used for such purposes. Full analyses of the curves, including computation of kinetic and thermodynamic parameters, can be done using freely available statistical software. The proposed procedure has been implemented in a web-based free software called anhysnuc, which can be found at https://sanjaychaudhuri.shinyapps.io/anhysnuc/. Finally, we illustrate our method by analysing irreversible curves encountered in the formation of a G-quadruplex and an LNA-modified parallel duplex.2022-09-02T21:19:39ZSanjay ChaudhuriDang Trung KienGopinatha Suresh KumarSouvik MaitiDaisuke MiyoshiJhimli Bhattacharyyahttp://arxiv.org/abs/2609.29575v1DeepGOF-1: A Pretrained Convolutional Goodness-of-Fit Test for Logistic Regression with a Computable Consistency Certificate2026-08-26T23:30:59ZGoodness-of-fit tests for logistic regression are least reliable where they are most needed: at small samples their levels drift from the nominal one, and combining them worsens the drift. We propose a test whose statistic is a convolutional network, trained once on simulated departures, that reads misfit as a picture: a grid of standardized residuals over covariate ranks. The analyst never trains. The network ships frozen, and the p-value is the rank of the observed score within the analyst's own bootstrap, so the level is a property of the calibration rather than of what the network learned. We prove exactness under pivotality, asymptotic exactness without it, and a consistency theorem whose key condition is computable from the frozen weights in one forward pass, giving a per-alternative certificate; we also measure the test's blind cone. On a pre-declared sixty-cell grid the deployed level stays in the nominal band in fifty-eight cells, while that grid's power criterion failed; on a published benchmark it is the most stable of thirteen levels across settings and sample sizes, and the test outpowers every partition test at every sample size while ranking seventh overall. A bank-failure application and a versioned release close the paper.2026-08-26T23:30:59Z18 pages, 4 figures, 3 tables. Submitted for publication. Research compendium (frozen weights, training corpus and generator, benchmark harness, and every per-replicate p-value) archived at doi:10.5281/zenodo.22113219. The deployable test ships as deepgof1() in the R package ebrahim.gofEbrahim Khaled Ebrahimhttp://arxiv.org/abs/2209.01269v2A Two-step Metropolis Hastings Method for Bayesian Empirical Likelihood Computation with Application to Quantile Regression and Bayesian Model Selection2026-08-26T23:21:28ZEmpirical likelihood-based methods have been used under the Bayesian framework (BayesEL) in recent times. For statistical inference, these methods require efficient Markov chain Monte Carlo (MCMC) samplers for drawing observations from the parameter posterior distributions. However, the complex, especially non-convex, nature of the empirical likelihood support makes such MCMC algorithms harder to design.
Such difficulties have restricted the use of BayesEL methods in many applications. In this article, we propose a two-step Metropolis-Hastings algorithm to sample from the BayesEL posteriors. Our proposal uses the current values of suitable subsets of the parameters and the estimating equations determining the underlying empirical likelihood to propose values of the remaining parameters.
The proposed method is thus suitable for sampling from BayesEL posteriors in many complex problems, especially those with discontinuous estimating equations, e.g., simultaneous quantile regression. Furthermore, the proposed method easily extends to BayesEL model selection through a reversible jump Markov chain Monte Carlo procedure. Several illustrative, real-life applications of our proposed methods are presented.2022-09-02T20:40:21ZSanjay ChaudhuriTeng YinSnehashis ChakrabortyRupsa Royhttp://arxiv.org/abs/2608.26450v1Kernelized Stein Discrepancy for Goodness-of-Fit Tests and Stein Sampling in R2026-08-26T22:56:34ZStein's method constructs computable discrepancies between a target distribution and a candidate distribution without requiring the target distribution's normalizing constant. These discrepancies support goodness-of-fit tests for model assessment as well as sampling tools for empirical approximation. The R package steinsampling provides the first unified R workflow for applying score-based Stein methods to kernel goodness-of-fit testing of independent or serially dependent observations, point transport, greedy point construction, and sample compression. High-level functions carry out each task in a single call, while the kernel, calibration, optimization, and transition components are provided separately so that users can replace any one of them. A single score and kernel setup can therefore be reused across sampling and testing, making these methods easier to reproduce, compare, and extend.2026-08-26T22:56:34ZJunhao GaoEry Arias-Castrohttp://arxiv.org/abs/2508.12627v3On computing and the complexity of computing higher-order $U$-statistics, exactly2026-08-26T16:02:17ZHigher-order $U$-statistics abound in fields such as statistics, machine learning, and computer science, but are known to be highly time-consuming to compute in practice. Despite their widespread appearance, a comprehensive study of their computational complexity is surprisingly lacking. This paper aims to fill this gap by presenting several results related to the computational aspect of $U$-statistics. First, we derive a useful decomposition from a $m$-th order $U$-statistic to a linear combination of $V$-statistics with orders not exceeding $m$, which are generally more feasible to compute. Second, we explore the connection between exactly computing $V$-statistics and Einstein summation, a tool often used in computational mathematics and quantum computing to accelerate tensor computations. Third, we provide an optimistic estimate of the time complexity for exactly computing $U$-statistics, based on the treewidth of a particular graph associated with the $U$-statistic kernel. The above ingredients lead to (1) a new, much more runtime-efficient algorithm to exactly compute general higher-order $U$-statistics, and (2) a more streamlined characterization of runtime complexity of computing $U$-statistics. We develop an accompanying open-source package called \texttt{u-stats} in both Python (https://github.com/zrq1706/U-Statistics-Python) and R (https://github.com/cxy0714/U-Statistics-R). We demonstrate through three examples in statistics that \texttt{u-stats} achieves impressive runtime performance compared to existing benchmarks. This paper also aspires to achieve two goals: (1) to capture the interest of researchers in both statistics and other related areas to further advance the algorithmic development of $U$-statistics and (2) to lift the burden of implementing higher-order $U$-statistics from practitioners.2025-08-18T05:01:10ZComments are welcome! 59 pages, 12 tables, 7 figures. An accompanying Python package is available at: https://libraries.io/pypi/u-stats or https://github.com/Amedar-Asterisk/U-Statistics-Python. Accepted by Statistics and ComputingXingyu ChenRuiqi ZhangLin Liuhttp://arxiv.org/abs/2605.28099v3A computationally-tractable measure of global sensitivity for sampling-based Bayesian inference2026-08-26T08:21:12ZBayesian inference can often be sensitive to the choice of hyperparameters of the prior or likelihood, yet defining and quantifying this sensitivity in a principled and computationally feasible way remains challenging in practice. Unfortunately, existing sensitivity methods are rarely applicable in modern Bayesian workflows due to their high computational cost and poor performance in moderate to high dimensions. To address these limitations, we introduce a new approach to global sensitivity analysis based on the Fisher divergence. Our method only requires a set of samples from a reference posterior and the ability to evaluate score functions, making it broadly computationally tractable. Under mild regularity conditions, it controls changes in the whole posterior, and provides a bound on the impact of perturbations on the first two moments. We demonstrate these strengths on challenging Bayesian inference problems which are practically out of reach of existing approaches, including generalised Bayesian inference for unnormalised models, inference in Bayesian models of time series, and neural simulation-based inference.2026-05-27T07:53:44ZArina OdnoblyudovaCharita DellaportaFrançois-Xavier Briolhttp://arxiv.org/abs/2608.25279v1Provable Non-Acceleration of Standard Strang Splittings of Kinetic Langevin Dynamics2026-08-26T01:27:39ZThe OBABO and BAOAB schemes and the other standard Strang splittings of kinetic (underdamped) Langevin dynamics are widely used Markov chain Monte Carlo algorithms. Under a suitable friction scaling, the underlying diffusion relaxes on a ballistic time scale, suggesting that these discretizations, suitably tuned, sample targets with condition number $κ$ in $O(\sqrtκ)$ iterations. We prove that no fixed choice of step size and friction, based only on the curvature bounds and the dimension, achieves this acceleration: total variation mixing time lower bounds for OBABO show that ballistic cold-start mixing fails uniformly over the smooth strongly convex class, and the lower bounds extend, with the same orders, to BAOAB and the other four Strang splittings. The proof transfers non-acceleration from optimization to sampling. Eliminating velocity gives an exact noisy heavy-ball recursion, and by the non-acceleration theorem of Goujaud, Taylor and Dieuleveut, for every tuning either some Gaussian target has a mode with relaxation time at least of order $κ$, or an attracting cycle exists on a smooth potential; dilating such a potential as $U_R(x)=R^2U(x/R)$ preserves its curvature bounds and produces metastability for a number of steps exponential in $R^2$, from an initial state at Wasserstein distance $O(R)$ from equilibrium. Using contraction estimates of Leimkuhler, Paulin and Whalley and a Wasserstein-to-total-variation regularization estimate, we prove a complementary upper bound of $O(κ)$ steps, up to logarithmic factors, for a fixed-parameter OBABO tuning. Hence, among fixed-parameter OBABO tunings, the optimal condition-number dependence of cold-start total variation mixing over this class is linear, up to logarithmic factors. A direct Gaussian calculation also rules out fixed-parameter acceleration for the left-endpoint exponential integrator.2026-08-26T01:27:39Z58 pages, 4 figures. Mixing time lower bounds ruling out ballistic (accelerated) cold-start mixing for all six standard Strang splittings of kinetic Langevin dynamics under fixed-parameter tunings, with a matching O(kappa) upper bound for OBABONawaf Bou-Rabeehttp://arxiv.org/abs/2606.05462v2A Two-Channel F-Transform Representation for Early Trajectory Characterization in Iterated Correlation Dynamics2026-08-25T22:45:43ZMany nonlinear iterative systems generate high-dimensional trajectories whose early behavior is informative but difficult to compare directly. This paper derives a fixed-dimensional F-transform coordinate representation for early trajectories of iterated Pearson correlation matrices. The construction is defined on the first five-point post-transient signal window, which is the shortest sampled window that simultaneously places the three symmetric fuzzy nodes at observed positions and supports a nondegenerate centered first-degree F-transform coefficient, thereby providing the earliest feasible local level--trend characterization within this sampled geometry. The representation combines two logarithmic observables of the post-transient dynamics: step size and contraction ratio. Applying this same four-coordinate construction to the step-size and contraction-ratio signals yields the eight-dimensional descriptor $Ψ=(v_1,v_2,v_3,s_2,u_1,u_2,u_3,r_2)$, with a common coordinate form across matrix sizes. For the fixed construction, $Ψ=M(q_2,\ldots,q_7)^{\top}$, $q_k=\logδ_k$, with $\operatorname{rank}M=6$. Thus, the descriptor is an injective, overcomplete representation of the six logged step sizes underlying the two channels. The representation is Lipschitz stable, and the centered first-degree coefficient recovers affine trends exactly. Convergence-length approximation is used as a downstream test of retained dynamical information. Across 22 matrix dimensions and 22,000 trajectories, repeated train--test evaluation shows predictive performance comparable to raw two-channel and statistical-summary representations. PCA shows that the first two principal components explain on average $84.47\%$ of the descriptor variance. Clustering reveals reproducible coarse organization, with the strongest mean silhouette at $k=2$ and high stability for smaller numbers of clusters.2026-06-03T21:39:55Z33 pages, 4 figures, 7 tablesIshrak Alhajj Hassanhttp://arxiv.org/abs/2608.24768v1Score-Based Ideal Observer Approximation via Denoising Score Matching for Signal-Known-Exactly Detection Tasks2026-08-25T16:05:55ZThe Bayesian Ideal Observer (IO) establishes the theoretical upper bound on task performance for binary detection tasks. However, analytical computation of the IO test statistic is generally intractable. Numerical approaches based on Markov-chain Monte Carlo (MCMC) methods, including their recent deep generative model-based extensions, typically require extensive posterior sampling for each test image. Supervised learning has also been investigated to approximate the IO performance. However, such methods are typically trained for a specific detection task and signal and may require retraining when the task or signal changes. The score function, defined as the gradient of the log probability density, encodes the local geometry of the data distribution and is a fundamental quantity in modern score-based generative modeling. This work reformulates the IO test statistic in terms of the score function and introduces a score-based ideal observer (SIO). The proposed SIO uses a denoising convolutional neural network trained exclusively on signal-absent images to estimate the signal-absent score function. Once trained, the resulting score model can be used to approximate the IO test statistic for detection tasks involving arbitrary additive signals, without per-image posterior sampling or signal-specific retraining. Numerical studies consider a signal-known-exactly (SKE) detection task with a stochastic lumpy-background model. The results demonstrate that the proposed SIO can closely approximate the IO performance.2026-08-25T16:05:55ZSubmitted to SPIE Medical Imaging 2027Weimin Zhouhttp://arxiv.org/abs/2602.10714v3A Non-asymptotic Analysis for Learning and Applying a Preconditioner in MCMC2026-08-25T14:29:20ZPreconditioning is a common method applied to modify Markov chain Monte Carlo algorithms with the goal of making them more efficient. In practice it is often extremely effective, even when the preconditioner is learned from the chain. We analyse and compare the finite-time computational costs of schemes which learn a preconditioner based on the target covariance or the expected Hessian of the target potential with that of a corresponding scheme that does not use preconditioning. We apply our results to various algorithms including the Unadjusted Langevin Algorithm (ULA) and the proximal sampler for an appropriately regular target, establishing non-asymptotic guarantees for versions of these algorithms that learn and use preconditioners. To do so, we establish non-asymptotic guarantees on the time taken to collect $N$ approximately independent samples from the target for schemes that learn their preconditioners under the assumption that the underlying Markov chain satisfies a contraction in the Wasserstein-2 distance. This approximate independence condition, that we formalize, allows us to bridge the non-asymptotic bounds of modern MCMC theory and classical heuristics of effective sample size and mixing time, and is needed to amortise the costs of learning a preconditioner across the samples it will help to produce.2026-02-11T10:19:56ZMax HirdFlorian MaireJeffrey Negrea