https://arxiv.org/api/4pYxpSlwIjknOK2V6wUrrLwkH7o2026-09-10T20:15:09Z121454515http://arxiv.org/abs/2608.29090v1Detecting Multiple Phase Transitions in Lattice Systems with Intrinsic Dimensions2026-08-29T06:41:03ZLattice systems with multiple nearby transitions pose two related challenges: resolving distinct transition scales and identifying the degrees of freedom primarily associated with each transition. We show that the intrinsic dimension of Monte Carlo configuration ensembles, estimated by the two-nearest-neighbors method, provides a geometric diagnostic for both problems. In the two-dimensional $q$-state clock model, the intrinsic dimension distinguishes the ordered, quasi-critical, and disordered regimes for both well-separated ($q=9$) and closely spaced ($q=5$) Berezinskii--Kosterlitz--Thouless transitions. In the $q=5$ case, the intermediate phase appears as a broad low-dimensional valley even when energy and magnetization do not separately resolve the two transitions. In the four-dimensional $U(1)$ Higgs model, we introduce channel-decomposed intrinsic dimensions based on gauge-invariant plaquette and Higgs variables. The dominant response of each channel tracks transitions associated with the corresponding degrees of freedom, while the combined channel retains features of both. We further show that intrinsic dimensions evaluated directly on gauge-variant fields are dominated by gauge-orbit directions, demonstrating the importance of removing gauge redundancy before interpreting configuration-space geometry. These results establish channel-decomposed intrinsic dimension as a geometric probe of lattice systems with multiple transitions and motivate its application to disentangling deconfinement and chiral crossover scales in full QCD.2026-08-29T06:41:03Z18 pages, 15 figuresJie MeiTetsuo HatsudaMei HuangLingxiao Wanghttp://arxiv.org/abs/2609.01643v1A Probability Model for Pentagonal Prism Dice Rolls2026-08-28T17:57:35ZIn the realm of probability theory and game design, the study of dice probabilities plays a crucial role. The featured dice are usually fair, so their probabilities can be predicted without rolling the dice. Pentagonal prismatic dice are not fair. Hence, the probabilities are more complicated. Our mathematical model is naturally derived from both the geometric properties of the dice and empirical evidence. Dice of varying sizes but with constant volume were designed. Using a 3D printer with 100 percent infill, the dice were printed and prepared. A machine that simulates dropping dice through a dice tower onto a hard surface was used to obtain about 41,000 results. From these data, a statistical model was constructed that gives probabilities of dice with varying height-to-radius ratios. The formula for the model is based on the centroid solid angle, but with two adjustment parameters due to varying energy and other requirements to transition between different resting aspects. Within the range of height-to-radius ratios tested, the data are within reasonable statistical error of the model.2026-08-28T17:57:35Z10 pages, 8 figuresPaul R. HurstJ. Naleo Hydehttp://arxiv.org/abs/2608.28375v1Localizing Global Discrepancies: Marginal Contributions and Contextual Anomaly Detection2026-08-28T14:27:33ZGlobal goodness-of-fit and discrepancy statistics can establish that a sample departs from a reference distribution without identifying which observations drive the departure. We develop a framework for this localization problem by assigning to each observation its conditional or marginal contribution across random statistical contexts. This connects resampling diagnostics and data valuation to projection theory and event-level anomaly detection. For symmetric statistics, fixed-size replacement is exactly equivalent to centered conditional localization. For U-statistics, the addition score equals the first Hoeffding/Hájek contribution; for smooth distributional functionals it is related at leading order to the influence function; and for unbiased known-background MMD it reduces exactly to the MMD witness.
This viewpoint also yields more efficient estimators. Matched-context subtraction removes fluctuations unrelated to the observation, while for pairwise MMD the event-containing terms give a simple localizer. On the LHC Olympics anomaly-detection benchmark, the pair estimator converges to the direct empirical MMD witness with the predicted 1/(Rm^2) scaling, where m is batch size and R the number of batches. At m=1000 and R=5x106 it reaches correlation 0.9993 with essentially identical AUC.
We also ask when context contains information beyond an event's own features. In a shared-latent toy model, the full single-event signal and background distributions are identical by construction, forcing isolated-event AUC=0.5. Discriminating information survives only in cross-event dependence induced by the shared latent parameter; the ensemble recovers this information, whereas an independent-latent control does not. This separates two roles of context: efficient localization of a global discrepancy and genuinely additional class information when the alternative contains shared structure.2026-08-28T14:27:33Z32 pagesTommaso dorigohttp://arxiv.org/abs/2608.28314v1Vertex reconstruction for a search for neutron-antineutron conversions with HIBEAM2026-08-28T13:25:25ZThe HIBEAM/NNBAR programme (incorporated into the FINESSE/NNBAR programme) at the European Spallation Source is proposed to search for neutrons converting to antineutrons. An important observable is the reconstructed vertex arising from charged particles produced by an antineutron annihilating on a thin target foil and which pass through a time projection chamber. This paper studies track clustering and foil-plane vertex reconstruction for this topology. Both non-machine-learning methods and graph-neural-network methods are tested and compared with each other, including deterministic clustering, trackless projection and a hybrid clustering/graph-neural-network chain with track classification and vertex refinement. We conclude that, for the geometrically simple events in the HIBEAM TPC, classical reconstruction methods perform on-par with machine-learning based methods in terms of vertex coordinate reconstruction. Custom machine-learning based methods can, however, deliver event-shape information that may be important in downstream analyses.2026-08-28T13:25:25Z18 pages, 10 figures, 7 tablesAlexander BurgmanSze Chun YiuYamna ShaikhDavid MilsteadEily MerjyLucas ÅstrandKenneth ÖsterbergFredrik OljemarkMatthias HollValentina SantoroAndré NepomucenoJoshua L. Barrowhttp://arxiv.org/abs/2608.27333v1Evaluating Quantum Kernel Methods for Track-Based Classification in High-Energy Physics2026-08-27T16:35:30ZWe present a systematic design for large-scale quantum kernel classification, demonstrated through a quantum support vector classifier (QSVC) for particle-track classification using centroid-based CLAS12 drift-chamber features. Each event is encoded into a six-qubit state via a fully entangled ZZFeatureMap, whose fidelities define a quantum kernel within a standard SVM framework. By decoupling state preparation from kernel construction and distributing evaluation across a multi-node MPI-based HPC allocation, the approach scales to 1.0x10^5 training and 4.0x10^5 test events with an exactly constructed kernel matrix, to our knowledge more than an order of magnitude larger than prior high-energy-physics quantum-kernel studies.
Benchmarked against linear, polynomial, RBF, and sigmoid SVM kernels and extremely randomized trees (ERT), the ideal QSVC achieves the highest recall (99.99%) among all models. Under a calibrated hardware noise model (FakeMumbaiV2, 500 training / 2,000 test events), AUC falls from 0.9985 to 0.9671 and peak significance improvement falls from 17.5 to ~3.5, yet recall remains at 99.51% -- indicating this signal-retention advantage is attenuated but not eliminated by circuit-level decoherence.
Geometric analysis of the quantum embedding shows near-orthogonal inter-class states with coherent intra-class neighborhoods under ideal simulation; under noise this structure compresses toward the maximally mixed state while preserving its relative ordering.
These results demonstrate a scalable, reproducible workflow for quantum kernel experimentation at HEP-relevant scale, quantifying the practical cost of realistic hardware noise on quantum-enhanced classification.2026-08-27T16:35:30ZEmmanuel BilliasNikos Chrisochoideshttp://arxiv.org/abs/2608.21549v2Probabilistic inference of surface parameters for Monin-Obukhov similarity theory2026-08-26T21:12:55ZIn simulations of atmospheric flow, the grid spacing typically exceeds the size of the roughness elements at the surface by an order of magnitude. The unresolved effects of surface morphology and roughness on the flow are represented by effective surface parameters and specified as a surface flux boundary condition, most often through a formulation based on Monin-Obukhov similarity theory (MOST). These surface parameters are known to depend on both surface and flow properties, yet they are generally estimated as deterministic quantities with no characterization of associated uncertainty. In this study, we use a Bayesian approach to infer surface parameters for MOST and quantify their uncertainties. For the aerodynamic roughness length $z_0$ inferred in isolation, a normal-normal conjugate update yields the posterior and posterior predictive distributions in closed form. We first demonstrate our method on idealized conventionally neutral boundary layers generated by large-eddy simulation, where $z_0$ is prescribed, and quantify how prior- and observation-related choices shape the inferred posterior. We then apply the method to field observations from the Atmospheric Radiation Measurement Southern Great Plains observatory, from which we infer $z_0$ distributions conditioned on month, on wind direction, and on both jointly. By leveraging a prior informed by sample statistics of all near-neutral observations in the training years, we demonstrate the advantage of the Bayesian inference method for $z_0$ relative to the state-of-practice least-squares profile-fitting method in conditions of data sparsity. Predictions for unseen observations-evaluated at the posterior mean-reduce root mean squared error and mean absolute error in data-sparse wind directions, while posterior predictive distributions consistently reduce the continuous ranked probability score by approximately $20$-$30\%$.2026-08-21T18:37:44Z19 pages, 9 figures, 1 tableEthan YoungIn ShinMichael F. Howlandhttp://arxiv.org/abs/2608.26354v1Cross-simulator transfer with foundation model summaries: Towards robust SKA-era reionization inference2026-08-26T19:37:18ZSimulation-based inference (SBI) for parameter estimation is vulnerable to model misspecification: neural summaries and density estimators trained on a specific forward model typically fail when applied to data drawn from another model, or from real observations, and no training simulator can capture the full observational pipeline of a real measurement exactly. We show that a self-supervised Vision Transformer (ViT), pretrained label-free on a fast approximate simulator, produces transferable data summaries that generalize across simulators. Without retraining, it can be reused as a frozen encoder to infer astrophysical parameters from a completely different simulator that resolves the radiative transfer explicitly, on which it has never seen either data or parameters. As a concrete use case in 21cm cosmology, SKATR, a ViT pretrained with a Joint Embedding Predictive Architecture (JEPA), serves as a foundation model for reionization inference from upcoming SKA measurements: SKATR is pretrained once on 67k low-cost, noiseless semi-numerical 21cmFAST lightcones, then frozen and applied to hydrodynamical Loreli II lightcones, where a lightweight conditional flow matching head infers five astrophysical parameters; the encoder is never shown Loreli data, its parameters, or any noise. In our comparison, SKATR yields the most precise and best-calibrated posteriors across all five parameters, matching the accuracy of the fully-supervised in-domain baseline while requiring 2.6x fewer radiative-transfer simulations. Under realistic SKA AA* noise, only SKATR remains simultaneously accurate, informative, and calibrated, outperforming even a supervised baseline retrained from scratch on noisy data. Self-supervised pretraining on computationally efficient semi-numerical simulations is therefore a viable route to calibrated, simulator- and noise-agnostic reionization inference for the SKA-era.2026-08-26T19:37:18Z12 pages, 5 figures, prepared for submission to A&AYannic PietschkeCaroline HenekaAyodele OreRomain Meriothttp://arxiv.org/abs/2605.24441v2On the Statistical Interpretation of Discoveries in LHC Data2026-08-26T19:13:11ZWe examine discovery criteria at the Large Hadron Collider (LHC) within a model-independent framework, with particular emphasis on the statistical signatures of new physics. This study is motivated by the recent shift from model-specific searches based on a small number of distributions to broad, model-agnostic strategies, which offer substantially greater sensitivity to unexpected phenomena. We revisit the well-known criterion of a local statistical significance of $5\,σ$ for the observation of new phenomena in invariant-mass distributions and discuss how this threshold should be modified to account for look-elsewhere effects arising not only from multiple bins within a given distribution, but also from the simultaneous consideration of multiple distributions. We present a simple but statistically conservative relation between local and global significances in the presence of multiple invariant-mass distributions at the LHC, which can serve as a useful first approximation for planning future measurements.2026-05-23T07:22:25Z9 pages, 2 figuresJ. Phys. G: Nucl. Part. Phys. 53 (2026), 085003S. V. ChekanovE. J. Weik10.1088/1361-6471/ae996dhttp://arxiv.org/abs/2501.01437v4On the reconstruction limits of complex networks2026-08-26T13:01:47ZNetwork reconstruction consists in retrieving the hidden interaction structure of a system from observations. Many reconstruction algorithms have been proposed, although less research has been devoted to describe their theoretical limitations. In this work, we take a first-principles approach and build on our earlier definition of reconstructability-the fraction of structural information recoverable from data. We relate this quantity to the true data-generating (TDG) process and delineate an information-theoretic reconstruction limit, i.e., the upper bound of the mutual information between the true underlying graph and any graph reconstructed from observations. These concepts lead us to a principled numerical method to assess the validity of empirically reconstructed networks, based on model selection and a quantity we introduce: the reconstruction index. This index approximates the reconstructability from data, quantifies the variability of the reconstructed network ensemble, and is shown to predict reconstruction error without requiring knowledge of the true underlying network. We characterize this method and test it on empirical time series and networks.2024-12-23T18:53:59ZCharles MurphyAriane LizotteFrançois ThibaultVincent ThibeaultPatrick DesrosiersAntoine Allardhttp://arxiv.org/abs/2608.13848v2Posterior Inference of Hamiltonian Parameters from RIXS Spectroscopy2026-08-26T03:49:43ZWe present the first application of simulation-based inference to resonant inelastic X-ray scattering spectroscopy. Using truncated marginal neural ratio estimation to efficiently restrict the prior and conditional flow matching as the joint density estimator, we infer full posteriors with a modest simulation budget for two Ni$^{2+}$ compounds---NiPS$_3$ as a representative covalent case and K$_2$NiF$_4$ as a more atomic one. We demonstrate that a vision transformer encoder whose tokenization matches the physical layout of the RIXS map yields better-covered and sharper posteriors than generic image encoders. Applying the validated method to experimental NiPS$_3$ and K$_2$NiF$_4$ data, we recover a joint posterior that reveals parameter correlations invisible to point estimators, and a posterior predictive distribution that closely matches the observed spectrum. The amortized posterior unlocks a class of analyses not previously available to the field such as nuisance-marginalized uncertainty quantification, multi-measurement posterior fusion and active experimental design.2026-08-14T00:42:07ZSamuel KleinThomas M. LinkerLouis ConreuxDaniel RatnerApurva MehtaMakoto TachibanaJiemin LiJonathan PelliciariValentina BisogniWei HeXiangpeng LuoMark P. M. DeanMarton K. LajerMichael KaganJoshua J. TurnerYongqiang ChengSean Gasiorowskihttp://arxiv.org/abs/2609.01637v1Geometry-native machine learning reconstruction of DSMC moment fields with support monitoring2026-08-26T00:31:15ZDirect simulation Monte Carlo (DSMC) resolves rarefied-gas dynamics without a constitutive closure, but finite-sample estimates of macroscopic moments converge at markedly different rates. We develop a non-intrusive, geometry-native machine learning reconstruction of the retained two-dimensional moment hierarchy: number density, two velocity components, translational temperature, three pressure-tensor components, and two heat-flux components. From three sampling blocks, the estimator corrects a structured prior learned from development data with a bounded term computed from the current observation, while preserving additive-moment consistency and the measured zero-frequency content. In cavity development tests, the final observation-conditioned estimator, whose prior is a trained MambaIR restoration network, reduces transverse-heat-flux error to 0.658 and 0.672 times that of a ten-block direct average at two rarefied conditions. For a hypersonic cylinder, a cylinder-centred estimator is fixed before evaluation on six new observation/reference pairs. It improves both global transverse heat flux and near-wall normal heat flux in every pair; the ratios of arithmetic-mean normalised root-mean-square errors (NRMSEs) are 0.846 and 0.793, and the Holm-adjusted one-sided exact probabilities are 0.03125.2026-08-26T00:31:15ZEhsan Roohihttp://arxiv.org/abs/2604.12364v2Cross-Domain Transfer with Particle Physics Foundation Models: From Jets to Neutrino Interactions2026-08-25T19:59:52ZFuture AI-based studies in particle physics will likely start from a foundation model to accelerate training and enhance sensitivity. As a step toward a general-purpose foundation model for particle physics, we investigate whether the OmniLearned and ParticleViT foundation models pretrained on diverse high-$Q^2$ simulated and real $pp$ and $ep$ collisions retain useful knowledge to a few-GeV fixed-target neutrino experiment. We process MINERvA neutrino--nucleus scattering events and evaluate pretrained models on two types of tasks: regression of available energy and binary classification of charged-current pion final states ($\mathrm{CC1π^{\pm}}$, $\mathrm{CCNπ^{\pm}}$, and $\mathrm{CC1π^{0}}$). Pretrained OmniLearned and ParticleViT models outperform similarly sized models trained from scratch at the same compute budget, with the largest gains for OmniLearned on regression and for ParticleViT on classification. When the same transformer architecture is instead initialized from unrelated text pretraining (BERT), this advantage appears only marginally for classification in terms of compute efficiency and not in any way for regression. These results suggest that particle-level foundation models acquire inductive biases that generalize across large differences in energy scale, detector technology, and underlying physics processes, pointing toward detector-agnostic inference in particle physics.2026-04-14T06:53:22Z18 pages, 13 figuresGregor KrzmancVinicius MikuniBenjamin NachmanCallum Wilkinsonhttp://arxiv.org/abs/2506.14006v2Evolutionary chemical learning in dimerization networks2026-08-25T18:37:47ZWe present a framework for chemical learning based on Competitive Dimerization Networks (CDNs) - systems in which multiple molecular species, e.g., proteins, DNA oligomers, or RNA oligomers, reversibly bind to form dimers. We show numerically that these networks can, in principle, be trained in vitro through directed evolution, enabling the implementation of complex learning tasks such as multiclass classification without digital hardware or prior knowledge of all microscopic association constants. Each molecular species functions analogously to a neuron, with binding affinities acting as tunable synaptic weights. A training protocol involving mutation, selection, and amplification of DNA-based components allows CDNs to robustly discriminate among noisy input patterns. The resulting classifiers exhibit strong output contrast and high mutual information between input and output, especially when guided by a contrast-enhancing loss function. Comparative analysis with in silico gradient descent training reveals closely correlated performance. These results establish CDNs as a promising platform for analog physical computation, bridging synthetic biology and machine learning, and advancing the development of adaptive, energy-efficient molecular computing systems.2025-06-16T21:10:12Z10 pages, 5 figures + AppendixAlexei V. TkachenkoBortolo Matteo MognettiSergei Maslovhttp://arxiv.org/abs/2503.09980v4Thermodynamic cost of inference and learning in physical neural networks2026-08-25T18:31:46ZHow much of the energy consumed by artificial neural networks is set by physics rather than by implementation? For irreversible digital hardware the reference is Landauer's principle, which charges $k_B T\ln 2$ per erased bit. We map a generic feedforward network onto a physical Hamiltonian in which each layer relation is an elastic compatibility constraint, and obtain two exact bounds. First, its equilibrium free energy is independent of the input and of every weight and bias, at all temperatures, so quasi-static inference requires no work whatsoever: no thermodynamic cost attaches to computation itself. Second, at finite speed the work exceeds the squared Wasserstein-2 distance between the initial and final thermal states, divided by the protocol duration. Relaxed to an entropic measure of distinguishability, this identifies the cost with the information separating successive inputs, about $k_B T$ per dimension of the widest layer at the fastest usable speed. Learning is fundamentally different: writing the parameters carries an irreducible cost of a few $k_B T$ each that survives the quasi-static limit. The thermodynamic price of a neural network is therefore set by its memory rather than its arithmetic. Simulations confirm both bounds: the inference work saturates the transport bound to within four percent, and accuracy collapses once the dissipated work falls below the thermal scale.2025-03-13T02:35:07Z22 pages, 3 figure, 1 table, substantial revisionAlexei V. Tkachenkohttp://arxiv.org/abs/2608.24872v1Binary Hypothesis Testing: A Robust Framework Against the Look Elsewhere Effect2026-08-25T17:52:29ZIn particle physics, discovery claims conventionally require an observed significance exceeding $5σ$. However, the interpretation of a $5σ$ result depends critically on the testing procedure, namely whether the hypothesis is tested at a single pre-specified point in parameter space or by scanning over a range of possible signal locations. This distinction gives rise to the look-elsewhere effect, a concept that is widely used but often interpreted as a simple penalty for scanning. In this work, we reinterpret the look-elsewhere effect as a correction for procedural inconsistency arising when the null distribution is generated under one procedure while the test statistic is evaluated under another. Within the framework of hypothesis testing, we examine its implications and clarify the distinct roles of the look-elsewhere effect in binary and peak-search scenarios. In particular, we show that binary test is robust against the look-elsewhere effect, whereas peak searches require an explicit correction for the search over signal locations. Using a moderate trial factor of approximately 26, calibrated from the ATLAS Higgs search, we show that a $3σ$ global significance of peak search can correspond to approximately $4σ$ significance of the binary test for the same value of the observed test statistic. This reformulation provides a clearer statistical interpretation of the look-elsewhere effect and offers a more coherent framework for understanding significance claims in particle physics.2026-08-25T17:52:29Z19 pages, 6 figuresHan ZhangXue-feng DingYu-Feng LiYi-fang WangLiang-jian WenLiang Zhan