https://arxiv.org/api/UBcRTOy1ovMElE9ztyvZd6nUFPc2026-09-10T23:29:12Z121459015http://arxiv.org/abs/2602.13913v3Early warning indicators of spatial epidemics from gauge mediated contagion: a statistical mechanics approach applied to COVID19 in Germany2026-08-16T19:29:51ZWe develop a statistical mechanics model for spatially extended epidemic dynamics in which contagion is mediated by an explicit environmental field rather than by instantaneous local contact. Building on the Doi Peliti formalism, we construct a stochastic field theory for susceptible and infected hosts coupled to a diffusive pathogen field and integrate out the mediator to obtain a non local effective action with a Yukawa type interaction kernel in space and memory effects in time. This formalism yields closed analytical expressions for the basic and effective reproduction numbers, including one loop fluctuation corrections that renormalize the transmission rate and generate spatial shielding effects beyond mean field SIR models. We derive an effective screening mass and a Debye like correlation length whose critical softening signals an impending instability of the disease free state, providing a field theoretical early warning indicator for epidemic waves. Superspreading is incorporated through a heterogeneous ``epidemic charge`` distribution, which controls the divergence of the correlation length and explains the emergence of multifocal outbreak patterns even when the population average $R_0$ is subcritical. Finally, we apply the model to high resolution COVID19 incidence data across $\approx 400$ German districts, dynamically extracting the effective screening mass and the renormalized reproduction number, and show that the gauge mediated indicators anticipate major changes in epidemic regime several days before classical SIR based estimates.2026-02-14T22:31:57Z23 pages, 8 figures,PHYSA, 2026, 131940Jose de Jesus Bernal-AlvaradoDavid Delepine10.1016/j.physa.2026.131940http://arxiv.org/abs/2606.26228v2Interpreting "Interpretability" and Explaining "Explainability" in Machine Learning in Physics2026-08-16T18:24:00ZWe review the concepts of interpretability and explainability as they apply to machine learning in physics. We define interpretability as concerning the structural transparency of a model (the ability to understand or approximate its inner workings) and explainability as concerning the scientific content of a model (the ability to map it onto domain knowledge). We discuss the trade-offs each entails (interpretability vs. expressivity; explainability vs. adaptability), the contexts in which each is needed, and the intrinsic and post-hoc tools available for achieving them. Throughout, we emphasize that machine-learned models are subject to the same scientific questions as classical models, differing only in scale, and that interpretability and explainability are best understood as deliberate modeling choices rather than inherent properties. We also emphasize the importance of task specification and intervention plans as a core aspect of model design.2026-06-24T18:00:01Z31 pages, 3 figures, Part of the VERaiPHY Initiative; v2: Minor revisionsRikab GambhirLuisa Lucie-SmithJesse Thalerhttp://arxiv.org/abs/2608.15476v1A theory of spatial early warning signals for tipping points on complex networks2026-08-16T01:47:17ZSpatial early warning signals (EWSs) seek evidence of an approaching tipping point from a single snapshot of many interacting elements. Existing theory largely assumes spatial homogeneity, whereas networks introduce systematic differences among nodes that may obscure fluctuation-based warning signals. We develop a mathematical framework for spatial EWSs in stochastic dynamical systems on networks. We find that the expected spatial variance, a popular spatial EWS, decomposes exactly into a structural contribution from heterogeneity in the equilibrium state and a fluctuation contribution determined by the stationary covariance. Near a simple steady-state bifurcation, the potentially divergent covariance concentrates along the critical eigendirection: the left eigenvector determines how strongly noise excites the critical fluctuation, while the right eigenvector determines its spatial pattern. Consequently, the spatial variance has a divergent fluctuation contribution when the limiting critical eigendirection is noise-excited and spatially nonuniform after centering. In contrast, the spatial coefficient of variation generally saturates, while skewness, kurtosis, and Moran's $I$ approach network-dependent limits without a universal warning direction. We also derive results for homogeneous networks, node-wise baseline subtraction as preprocessing, and Hopf bifurcations, for which the limiting distributions are qualitatively different. These results clarify when spatial EWSs provide reliable warnings and why their performance depends on network structure, noise, and preprocessing.2026-08-16T01:47:17Z1 figureNaoki Masudahttp://arxiv.org/abs/2608.15312v1The geometry of uncertainty decomposition in profile-likelihood fits2026-08-15T16:36:37ZUncertainty decompositions in profile-likelihood fits are commonly reported through nuisance-parameter impacts, although shifting a fitted parameter and fluctuating the observation that constrains it answer different questions. Recently, Pinto et al. provided an explicit construction for uncertainty decomposition based on fluctuating the observations and argued that such a construction allows for a cleaner interpretation of systematic uncertainties in terms of physical sources. We provide a geometric description of the distinction between the two methods by using a coupled pair of fiber bundles. The geometrical approach clarifies the relationship between several methods traditionally used to estimate systematic uncertainties in high-energy physics. By describing the profiling as an information-orthogonal horizontal lift and the Schur complement as the induced metric on the parameter-of-interest manifold, physical-source uncertainties arise by mapping observation fluctuations to score covectors, raising them with the inverse total information, and pushing the resulting estimator covariance forward to the parameters of interest. This construction clarifies why nuisance-parameter impacts do not generally coincide with repeated-experiment source variances.2026-08-15T16:36:37Z15 pages, 5 figuresRafael Coelho Lopes de Sáhttp://arxiv.org/abs/2510.05143v2Functional Connectivity Networks for Transportation Delay Analysis: from Theory to Software2026-08-15T13:48:23ZWithin the endeavour of modelling and understanding the propagation of delays in transportation networks, an approach that has attracted increasing interest in the last decade is the creation of functional network representations. These graphs map elements of interest (e.g. airports or stations) as nodes, and derive pairwise propagation patterns from their dynamics through correlation and causality tests. In spite of multiple notable results, this approach still lacks a coherent framework, with decisions related to many fundamental steps being left to the judgement of the researcher. We here provide an introduction to the theory behind functional networks for transportation systems, detailing the main steps and the associated pitfalls. We further introduce a Python package, delaynet, designed to support the researcher in the reconstruction and analysis of such networks. We finally present an analysis of the propagation of delays in the Swiss train system; and discuss future research steps.2025-10-01T22:28:26Z39 pages, 21 figures, 6 tables, for documentation, see https://delaynet.readthedocs.io/Carlson Moses BüthMassimiliano Zanin10.1016/j.trip.2026.102187http://arxiv.org/abs/2608.14801v1Statistical validation of calorimeter inpainting with generative diffusion priors2026-08-14T18:14:23ZLocalized detector inefficiencies produce incomplete calorimeter data that limit the ability to perform precision measurements. We address this problem in relativistic heavy-ion collisions from a Bayesian perspective using pretrained calorimeter diffusion models as priors to reconstruct the missing signal conditioned on surrounding measurements. In this work, we conduct a systematic comparison of several diffusion-based inpainting algorithms, whose performance is evaluated using Bayesian posterior diagnostics of energy response, spatial bias, and uncertainty calibration. The reconstruction fidelity is also analyzed across collision centralities and masked region sizes. This study establishes a general validation strategy for probabilistic reconstruction of missing detector information.2026-08-14T18:14:23Z21 pages, 15 figuresHimanshu RajRoli Eshahttp://arxiv.org/abs/2512.05790v10Learnability Window in Gated Recurrent Neural Networks2026-08-14T12:43:53ZWe develop a statistical theory of temporal learnability in recurrent neural networks, quantifying the maximal temporal horizon $\mathcal{H}_N$ over which gradient-based learning can recover lag-dependent structure at finite sample size $N$. The theory is built on the effective learning rate envelope $f(\ell)$, a function that captures how gating mechanisms and adaptive optimizers jointly shape the coupling between state-space dynamics and parameter updates during Backpropagation Through Time. Under heavy-tailed ($α$-stable) fluctuations, where empirical averages concentrate at rate $N^{-1/κ_α}$ with $κ_α= α/(α-1)$, the interplay between envelope decay and statistical concentration yields explicit scaling laws for the growth of $\mathcal{H}_N$: logarithmic, polynomial, and exponential temporal learning regimes emerge according to the decay law of $f(\ell)$. These results identify envelope decay as the key determinant of temporal learnability. Slower attenuation of $f(\ell)$ enlarges $\mathcal{H}_N$, while heavy-tailed fluctuations compress it by weakening statistical concentration. Moreover, envelope geometry outweighs dataset size: slowing the envelope's decay enlarges $\mathcal{H}_N$ more than adding data, so more complex architectures that realize slower-decaying envelopes can be more data-efficient than simpler ones. Experiments across multiple gated architectures and optimizers corroborate these structural predictions.2025-12-05T15:16:59ZAccepted at Physical Review ELorenzo Livi10.1103/843n-yshjhttp://arxiv.org/abs/2601.13812v3It's Not the Tool, It's the Task: A Framework for Cognitively Activated AI Augmentation in Physics Instruction2026-08-14T08:52:21ZGenerative artificial intelligence (AI) systems can now reliably solve many standard tasks used in introductory physics courses, producing correct equations, graphs, and explanations. While this capability is often framed as an opportunity for efficiency or personalization, it also poses a subtle ethical and educational risk: students may increasingly submit correct results without engaging in the epistemic practices that define learning physics. This challenge has recently been described as the "boiling frog problem" because we may not fully recognize how rapidly AI capabilities are advancing and fail to respond with commensurate urgency. In this article, we argue that the central challenge of AI in physics education is not cheating or tool selection, but instructional design. Drawing on research on self-regulated learning, cognitive load, multiple representations, and hybrid intelligence, we propose a practical framework for cognitively activated learning activities that structures student activities before, during, and after AI use. Using an example from an introductory kinematics laboratory, we show how AI can be integrated in ways that preserve prediction, interpretation, and evaluation as core learning activities. Rather than treating AI as an answer-generating tool, the framework positions AI as an epistemic partner whose contributions are deliberately bounded and reflected upon.2026-01-20T10:10:07ZJochen KuhnStefan KüchemannDave RakestrawPatrik Vogthttp://arxiv.org/abs/2608.14045v1Bayesian inference of event-by-event collision geometry from charged-particle multiplicity in heavy-ion collisions2026-08-14T07:46:09ZWe propose the Inference-driven Participant Determination (IPD) method, a Bayesian framework for inferring event-by-event posterior distributions of the number of participants ($N_{\text{part}}$) and binary collisions ($N_{\text{coll}}$) from final-state charged-particle multiplicities in relativistic heavy-ion collisions. The joint distribution of $(N_{\text{part}}, N_{\text{coll}})$ obtained from the Monte-Carlo Glauber model is used as the prior, while negative binomial distributions calibrated to charged-particle multiplicity fluctuations define the likelihood. This approach replaces conventional hard-cut centrality classification with a probabilistic assignment based on $N_{\text{part}}$, making the multiplicity--geometry smearing explicit and reducing the impact of volume fluctuations on downstream observables. A closure test using an UrQMD-MCG hybrid model at $\sqrt{s_{NN}} = 19.6$~GeV shows that the method yields well-calibrated posterior distributions with negligible bias and improves the reconstruction of net-proton cumulants relative to conventional multiplicity-based centrality selection.2026-08-14T07:46:09ZYige HuangFu-Peng LiHanwen FengNu Xuhttp://arxiv.org/abs/2608.13633v1Unknown Unknowns: Model Misspecification in Machine Learning for Physics2026-08-13T15:08:04ZMachine learning is now a central tool for solving inverse problems in particle physics and astronomy. Models are trained on simulation and deployed on real data, raising the question not just of whether they fit, but of whether they are wrong in ways we did not anticipate: the unknown unknowns. This challenge of model misspecification is not unique to machine learning. In physics, misspecification is sometimes exactly what we want to find: new discoveries appear as failures of existing models. At other times, we want such effects absorbed into the analysis without biasing the measurement. A robust analysis is one that absorbs the misspecifications we are not interested in, while preserving sensitivity to the ones we are. Machine learning can both amplify misspecification and provide new tools to address it. We discuss the challenges of model misspecification, diagnostics for detecting it, and strategies for mitigation. No single diagnostic can confirm that a model is correctly specified: detection and mitigation are two halves of an iterative loop, in which a battery of complementary diagnostics is applied, the model is updated, and the process repeated. Robustness against unknown unknowns is ultimately less about any single technique than about a disposition: a willingness to suspect one's own model, and to design analyses that can survive being wrong in ways one did not anticipate.2026-08-13T15:08:04Z25 pages, 2 figures, part of the VERaiPHY initiativeJuan Cruz-MartinezCarolina Cuesta-LazaroAlexander HeldMichael Kaganhttp://arxiv.org/abs/2603.16946v2Automatic Termination Strategy of Inelastic Neutron-scattering Measurement Using Bayesian Optimization for Bin-width Selection2026-08-13T13:50:52ZCurrently, an excessive amount of event data is being obtained in four-dimensional inelastic neutron-scattering experiments. A method for automatic bin-width optimization of multidimensional histograms has been developed and recently validated on real inelastic neutron-scattering data. However, measuring beyond the equipment resolution leads to inefficient use of valuable beam time. To improve experimental efficiency, an automatic termination strategy is essential. We propose a Bayesian-optimization-based method to compute a stopping criterion that can support online decisions on whether to continue or terminate an experiment. In the proposed method, the bin-width optimization is performed using Bayesian optimization to efficiently compute the optimal bin widths. The experiment is terminated when the optimal bin widths become smaller than the target resolutions. In numerical experiments using real inelastic neutron-scattering data, the optimal bin widths decrease as the number of events increases. Even the optimal bin widths for data downsampled to 1/5 are comparable with the resolutions limited by the sample size, choppers, and so on. This implies excessive measurement of the inelastic neutron experiments for the moment. Moreover, we found that Bayesian optimization can reduce the search cost to approximately 10% of an exhaustive search in our numerical experiments.2026-03-16T13:29:41Z17 pages, 6 figures; revised version under review at Journal of the Physical Society of Japan (JPSJ)Kensuke MutoHirotaka SakamotoKenji NagataTaka-hisa ArimaMasato Okadahttp://arxiv.org/abs/2510.18964v2Dark Matter profiles of "in silico" galaxies: deep learning inference2026-08-13T13:32:24ZMachine learning has the potential to improve the reconstruction of the dark matter profile of galaxies with respect to traditional methods, like rotation curves. We demonstrate on the simulation suite Illustris-TNG that a steerable equivariant convolutional neural network (CNN) is able to infer the dark matter profiles within and around individual galaxies from photometric and interferometric data, improving on a standard CNN. Within the in silico environment of the simulations, our architecture is able to capture the dark matter distribution within galaxies without a parametrization of the profile. We perform an interpretability analysis to understand the internal mechanisms of the trained model and the most important data features used to estimate the dark matter profiles. The equivariant CNN recovers the dark matter profile of galaxies within the stellar mass range $[10^{10} - 10^{12} ]$ $M_{\odot}$ with excellent precision and accuracy: the mean squared error is reduced by a factor of ~ 3 from its value under the training distribution, demonstrating that the network has learnt from the data features. While this holds within the controlled 'in silico' environment of the simulation, we argue that few additional steps are needed before this method can be reliably applied to galaxies in the real field observations.2025-10-21T18:00:03ZPublished: May 27, 2026JCAP05(2026)094Martín de los RiosSerafina Di GioiaFabio IoccoRoberto Trotta10.1088/1475-7516/2026/05/094http://arxiv.org/abs/2511.23149v2Precision Measurements of Higgs Hadronic Decay Modes at the FCC-ee2026-08-13T09:20:20ZThe expected precision at the FCC-ee on the product $σ\times\mathcal{B}(H\rightarrow b\bar{b}, c\bar{c},s\bar{s},gg)$ of Higgs boson production cross sections times branching ratios of hadronic decays is presented. This study provides the first comprehensive determination of all major hadronic Higgs decay modes in a combined fit at future $e^+ e^-$ colliders, using both Higgs-strahlung ($ZH$) and Vector boson fusion ($ν\barν H$) production processes, with a full treatment of interference effects in the $ν\barν jj$ final state. It assumes four identical IDEA detectors collecting $e^+e^-$ collisions at $\sqrt{s}=240$ and $365\,$GeV. The combination of all channels across both energies, with full covariance between production and decay modes, yields a production cross-section times branching-ratio precision at the percent to per-mil level for the dominant hadronic final states ($b\bar{b}, c\bar{c},gg$). These results provide a comprehensive input to the determination of Higgs coupling projections at the FCC-ee, and they establish for the first time sensitivity to the rare decay $H\rightarrow s\bar{s}$, demonstrating that FCC-ee has the potential to provide evidence of the strange-quark Yukawa coupling.2025-11-28T12:54:47Z30 pages, 11 figuresJHEP04(2026)181Andrea Del VecchioJan EysermansLoukas GouskosGeorge IakovidisAlexis MaloizelGiovanni MarchioriMichele Selvaggi10.1007/JHEP04(2026)181http://arxiv.org/abs/2608.12988v1Machine learning correction of satellite precipitation is governed by mechanism purity, not algorithmic complexity: a proof-of-concept study in Hunan, China, with pre-registered cross-regional validation2026-08-13T09:11:10ZSatellite precipitation products such as IMERG exhibit biases that vary with terrain, season, and precipitation regime, leaving the applicability boundaries of machine learning correction unclear. This study proposes the Terrain-Moisture-Intensity (TMI) framework, centered on mechanism purity, extending the correction problem from purely algorithmic optimization to physical consistency diagnosis. A proof-of-concept study in Hunan Province employs IMERG V07, SRTM DEM, and ERA5 variables (tcwv, u10, v10). Ablation results indicate that, under the conditions of this study, terrain-moisture relationships are predominantly additive: RF-Full yields merely +0.001 R^2 gain over LR-Full, while bias rises to 1.282 mm d^-1; MAE decreases by approximately 14%, reflecting a trade-off between tail-fitting improvement and mean shift. SHAP diagnostics identify three categories of boundaries. Spatially, Central Hunan exhibits significant degradation (R^2=0.133) despite strong variable activation, consistent with mechanism fragmentation induced by mixed terrain. Temporally, u10 undergoes directional reversal between summer and spring (+0.096 to -0.156), presenting "silent failure." Extreme precipitation (>=50 mm d^-1) approximates a mechanism saturation frontier rather than isolated out-of-distribution samples, with DEM showing the largest relative amplification in SHAP disorder (+150%). The results demonstrate that machine learning correction performance is primarily constrained by mechanism purity. A pre-registered cross-regional test (Hunan, Guangxi, Guangdong) confirms this screening capability out of sample: a priori coherence proxies predict correction efficiency with a mean absolute error of 2.6 percentage points, while the transfer-versus-retraining contrast separates mechanism mismatch (coastal Guangdong) from portability (Guangxi), establishing the framework as a validated applicability screen.2026-08-13T09:11:10Z44 pages, 8 figures, 5 tables; supplementary material included (8 tables, 2 figures)Yi Xuhttp://arxiv.org/abs/2508.06456v3Comparative study of ensemble-based uncertainty quantification methods for neural network interatomic potentials2026-08-13T01:00:11ZMachine learning interatomic potentials (MLIPs) enable atomistic simulations with near first-principles accuracy at substantially reduced computational cost, making them powerful tools for large-scale materials modeling. The accuracy of MLIPs is typically validated on a held-out dataset of \emph{ab initio} energies and atomic forces. However, accuracy on these small-scale properties does not guarantee reliability for emergent, system-level behavior -- precisely the regime where atomistic simulations are most needed, but for which direct validation is often computationally prohibitive. As a practical heuristic, predictive precision -- quantified as inverse uncertainty -- is commonly used as a proxy for accuracy, but its reliability remains poorly understood, particularly for system-level predictions. In this work, we systematically assess the relationship between predictive precision and accuracy in both in-distribution (ID) and out-of-distribution (OOD) regimes, focusing on ensemble-based uncertainty quantification methods for neural network potentials, including bootstrap, dropout, random initialization, and snapshot ensembles. We use held-out cross-validation for ID assessment and calculate cold curve energies and phonon dispersion relations for OOD testing. These evaluations are performed across various carbon allotropes as representative test systems. We find that uncertainty estimates can behave counterintuitively in OOD settings, often plateauing or even decreasing as predictive errors grow. These results highlight fundamental limitations of current uncertainty quantification approaches and underscore the need for caution when using predictive precision as a stand-in for accuracy in large-scale, extrapolative applications.2025-08-08T16:58:51ZYonatan KurniawanDepartment of Physics and Astronomy, Brigham Young University, Provo, Utah, USAMingjian WenInstitute of Fundamental and Frontier Sciences, University of Electronic Science and Technology of China, Chengdu, ChinaEllad B. TadmorDepartment of Aerospace Engineering and Mechanics, University of Minnesota, Minneapolis, Minnesota, USAMark K. TranstrumCross Stream Consulting, Springville, UT, USA10.1088/2632-2153/ae9fb4