https://arxiv.org/api/kTPna3Wm9KhmzXAqXBd025eF3Zs 2026-09-10T19:11:04Z 12145 30 15 http://arxiv.org/abs/2511.15853v2 The Ensemble Kalman Inversion Race 2026-09-01T19:42:02Z Ensemble Kalman methods were initially developed to solve nonlinear data assimilation problems in oceanography but are now popular in applications far beyond their original use cases. Of particular interest is climate model calibration. As hybrid physics and machine-learning models advance, the number of parameters and complexity of parameterizations in climate models will continue to grow. To fully realize these advances, we must move from laborious hand-tuning to calibration-driven model development in rapid iteration cycles. Thus, robust calibration of these parameters plays an increasingly important role. We focus on learning climate model parameters by minimizing the misfit between modeled and observed climate statistics in an idealized setting. Ensemble Kalman methods are a natural choice for this problem because they are derivative-free, scalable to high dimensions, and robust to noise caused by statistical observations. Given the many variants of ensemble methods proposed, an important question is: Which ensemble Kalman method should be used for climate model calibration? To answer this question, we perform systematic numerical experiments to explore the relative computational efficiencies of several ensemble Kalman methods. The numerical experiments involve statistical observations of Lorenz-type models of increasing complexity, frequently used to represent simplified atmospheric systems, with some featuring neural network parameterizations. For each test problem, several ensemble Kalman methods and a derivative-based method "race" to reach a specified accuracy, and we measure the computational cost required to achieve the desired accuracy. We investigate how prior information and the parameter or data dimensions define the computational costs of the various methods. 2025-11-19T20:18:40Z Journal of Advances in Modeling Earth Systems (2026), 18, e2025MS005629 Rebecca Gjini Matthias Morzfeld Oliver R. A. Dunbar Tapio Schneider 10.1029/2025MS005629 http://arxiv.org/abs/2604.11481v3 Emergence of Complex Web Structures 2026-09-01T14:09:40Z Complex structures often emerge from initially homogeneous or weakly correlated states. We address the apparent tension between this ordering and entropy growth through a unified framework combining semi-microscopic phase-space dynamics, transport geometry, information theory, and coarse-grained effective modeling. The key point is that entropy depends on the level of description: a coarse-grained spatial field may become more ordered as structure forms, even while the full phase-space description becomes more complex through shell crossing, multistreaming, and the activation of velocity degrees of freedom. Using a Lagrangian--Eulerian transport map, we show how density amplification is governed by the Jacobian of the deformation and how anisotropic collapse arises from the eigenvalues of a hierarchy of deformation tensors. Long-range interaction or information flow is encoded in the displacement field, so that nonlocality enters directly through transport. We connect this geometric description to a maximum-entropy Gaussian baseline and show how nonlinear transport and nonlocal coupling generate scale coupling, higher-order correlations, and non-Gaussianity. We then formulate a Landau--Ginzburg description in which the growth of seed anisotropies is interpreted as the activation of lower effective free-energy branches, providing a coarse-grained realization of self-organization. Applied to generated cosmological fields, this framework indicates that the nonlocal tidal level becomes relevant already at moderate overdensity. Although cosmological structure formation is the main realization considered here, the framework is intended more broadly as a mesoscopic language for systems in which transport, anisotropy, nonlocality, and self-organization are central. 2026-04-13T13:47:35Z 38 pages, 8 figures, 1 table, accepted Francisco-Shu Kitaura http://arxiv.org/abs/2509.21246v2 Peak separation methods for inverse photoelectron spectra: Comparing second derivative, curve fitting, and deconvolution analyses 2026-09-01T09:10:18Z Inverse photoelectron spectroscopy (IPES) is a powerful technique for probing the unoccupied electronic states of materials. It can be regarded as the inversion process of photoelectron spectroscopy (PES), which examines the occupied states. Recently developed low-energy inverse photoelectron spectroscopy (LEIPS) can significantly advance the study of unoccupied states, owing to minimal sample damage and suppressed dark counts compared to conventional IPES. However, the instrumental resolution remains at 0.2 eV, which is one order of magnitude lower than that of PES. Spectral broadening caused by the low instrumental resolution often results in overlapping peaks. Peak separation is therefore crucial in the analysis of LEIPS spectra. In this study, we compared three peak separation methods-second derivative, curve fitting, and deconvolution. These methods were applied to modeled and experimental LEIPS spectra of the lowest unoccupied molecular orbital-derived band of pentacene, which consists of two splitting peaks due to the two inequivalent molecules in the unit cell. We systematically and quantitatively evaluated the performance of each method in terms of analysis parameters and discussed its robustness to noise as well as its peak separation capabilities as a function of peak energy spacing and intensity ratio. This work offers a practical framework for peak separation in LEIPS, with extensions to PES and a wide range of spectroscopies. 2025-09-25T14:39:26Z 16 pages, 7 figures. Published in Review of Scientific Instruments Rev. Sci. Instrum. 97, 023906 (2026) Ryotaro Nakazawa Haruki Sato Hiroyuki Yoshida 10.1063/5.0303140 http://arxiv.org/abs/2609.00883v1 iPINN for Broadband CARS Phase Retrieval: A Framework for Function Approximation and Inverse Modeling Problems in Nonlinear Spectroscopy 2026-09-01T08:17:10Z Phase retrieval in broadband coherent anti-Stokes Raman spectroscopy (BCARS) is an ill-posed inverse problem. The Raman-like signal is encoded in the imaginary part of the resonant susceptibility, which mixes coherently with a non-resonant background (NRB) that varies across acquisitions. We introduce an inverse physics-informed neural network (iPINN) that predicts Lorentzian peak parameters from raw BCARS spectra and reconstructs the resonant susceptibility through a differentiable analytical forward model. A transformer encoder assigns spectral features to 24 learnable peak slots, and a multi-view consistency loss enforces invariance across NRB pattern, NRB strength, and noise. Unlike direct spectral regression approaches, the method retains accuracy under varying acquisition conditions. On a public benchmark, iPINN achieves the lowest error among the tested baselines (MAE 0.016 vs. next-best 0.046). On 28 zero-shot test spectra acquired across seven solvents and four focal positions, accuracy is depth-invariant in five of seven solvents. These results show that inverse parametric prediction with a differentiable physical decoder supports robust phase retrieval across measurement conditions. 2026-09-01T08:17:10Z Ravi Teja Vulchi Carl Messerschmidt Mohammadsadegh Vafaeinezhad Rajendhar Junjuri Tobias Meyer-Zedler Juergen Popp Thomas Bocklitz http://arxiv.org/abs/2609.01674v1 Multi-fidelity Monte Carlo estimation of floor response spectra under combined seismic and structural parameter uncertainties 2026-09-01T08:03:10Z Floor response spectra (FRS) are essential tools for the design of non-structural elements (such as equipment or components). Given the various physical phenomena influencing FRS, high-fidelity (HF) mechanical models of the primary structure may be required to estimate them. Since numerical simulations based on such models are generally computationally expensive, this paper proposes using a multi-fidelity Monte Carlo (MFMC) approach for the efficient estimation of FRS. The method relies on using observations from a fast low-fidelity (LF) model as control variables. If the absolute value of the correlation between LF and HF samples is close to 1, this approach reduces both variance and estimation error compared to a standard Monte Carlo estimate based solely on HF data samples. Through a case study involving the reactor building of the Kashiwazaki-Kariwa nuclear power plant, we demonstrate the suitability of this method for FRS estimation. It effectively reduces variance and estimation error, even when using a LF model as simple as a single-degree-of-freedom system. We also show that the method accounts for modeling uncertainties while maintaining comparable performance. Its ease of use makes it a valuable tool for practitioners. 2026-09-01T08:03:10Z Nils Baillie Baptiste Kerleguer Cyril Feau Josselin Garnier Fan Wang http://arxiv.org/abs/2601.11415v2 Auditing Frozen-Encoder Anomaly Detection Across Mechanical Systems: Representation Provenance, Calibration, and Protocol Effects 2026-09-01T06:31:06Z This version reports a reproducibility audit of the frozen-encoder experiments presented in version 1. The numerical discrimination results are reproducible from the preserved artifacts, but their original attribution to interferometric pretraining is not supported. The released checkpoint contains a nested model state that loads without missing parameters, whereas loading the outer checkpoint dictionary leaves almost the entire EfficientNet-B0 feature stack uninitialized. Preserved embeddings labelled as interferometric have norms of order $10^{-12}$, matching freshly initialized EfficientNet-B0 networks and differing by more than twelve orders of magnitude from the preserved ImageNet embeddings. A second, separately preserved near-zero embedding set produces almost the same IMS 4th-test anomaly scores ($r=0.987$) and record-level discrimination (AUC $0.9812$ versus $0.9818$). We therefore withdraw the causal claim that IMS performance demonstrates a morphological prior transferred from gravitational-wave instrumentation. We reanalyse the controlled IMS splits at matched observed false-positive rates and add multivariate classical signal baselines. The near-zero representations retain strong tail separation, particularly in the 2nd and 4th IMS runs, but this is now interpreted as an exploratory architecture-and-initialization effect coupled to Mahalanobis scoring. A separate PRONOSTIA audit shows that the original large warning times were induced by a lifetime-fraction baseline; under fixed-time evaluation, a ten-feature classical baseline outperforms the preserved encoder scores. These results illustrate how checkpoint provenance, finite-sample calibration, architecture, and target-domain baselines can create an appearance of cross-domain transfer. They also define the controls required before assigning physical meaning to frozen-representation anomaly scores. 2026-01-16T16:35:07Z 6 pages, 2 figures, 4 tables Jose Sánchez Andreu http://arxiv.org/abs/2609.00132v1 Anomaly detection for multijet scenarios 2026-08-31T18:00:01Z Signals of physics beyond the Standard Model continue to resist discovery at the LHC. Recent years have seen the proliferation of new anomaly detection techniques, promising discovery with significantly fewer model assumptions than traditional approaches. Strategies based on weak supervision have been especially successful, but so far were largely limited in scope by their reliance on a resonance manifesting a decay into a pair of jets. In this work, we demonstrate that a well-established idea from jet substructure physics - recursive soft drop - in combination with the CATHODE technique for anomaly detection can be used to simultaneously perform anomaly detection for signals with an arbitrary number of jets in the final state, greatly increasing the scope of such searches. 2026-08-31T18:00:01Z 14 pages, 7 figures, 2 tables Gregor Kasieczka Sung Hak Lim Louis Moureaux Tore von Schwartz David Shih Chitrakshee Yede http://arxiv.org/abs/2608.30838v1 Cosmic variance and ergodicity in finite systems with correlations 2026-08-31T14:08:50Z We consider the difference between ensemble and volume average in cosmology. It is known that for sufficiently weak long-range correlations the root mean square of the difference, which we call ergodicity bias, decays like $R^{-3/2}$ in the limit of large volume $R^3$. We calculate the condition this imposes on the power spectrum of a Gaussian random field. We quantify the bias for finite $R$, and show that the $R\to\infty$ limit is of little relevance for cosmological observations when the measured scales and correlations extend to the size of the observable universe. We consider curvature, density, and velocity perturbations. On large scales the bias is important in all three cases. For the density perturbations, which are observationally the most relevant, the relative bias first exceeds 100% at the separation $r=177$ Mpc, and is larger than 100% for all $r>560$ Mpc. It should be taken into account when comparing ensemble and volume averages for large-scale structure. The bias is also large for the cosmic microwave background temperature perturbations on large angular scales, but this is not relevant for observations, as their analysis does not involve volume averaging. 2026-08-31T14:08:50Z 18+3 pages, 16 figures Dipayan Mukherjee Syksy Rasanen http://arxiv.org/abs/2609.00094v1 Modernization and Statistical Validation of a Multilayer TRIM.SP Code for Low-Energy Muon and Ion Implantation 2026-08-31T14:01:24Z We describe the staged modernization of a local multilayer TRIM.SP implementation, denoted TRIM.SP-NL. The work included behavior-preserving refactoring, correction of localized bookkeeping and initialization defects, support for more than five elements per layer, reproducible build modes, and a runtime-selectable random-number-generator (RNG) interface. The RANLUX RNG remains the default reference backend, while PCG32 and xoshiro256** provide faster alternatives without changing the input-deck format. The code validation combined exact regression testing with statistical comparisons over 13 representative target multilayer configurations, comprising 143 energy-configuration points per backend and 429 simulations in total. Across the test suite, PCG32 and xoshiro256** reduced runtime by factors of 1.98 and 2.01, respectively. The implanted, backscattered, and transmitted fractions, as well as the mean implantation depths, range straggling, and representative depth profiles showed no systematic RNG dependence. Together with a Node.js-based graphical interface for setup, scans, and result collection, the Fortran engine forms the open-source TRIM.SP-NL Workbench. 2026-08-31T14:01:24Z 9 pages, 5 figures Zaher Salman Ryan M. L. McFadden Thomas Prokscha http://arxiv.org/abs/2609.01651v1 Parameter inference from a non-stationary unknown process using statistical feature-based slow feature analysis 2026-08-31T05:37:07Z Non-stationary phenomena are ubiquitous, with examples to be found in climatological measurements, brain activity, and the behavior of financial markets. Starting with a time series from a non-stationary process, a key challenge is to infer the time-varying parameters that underlie the non-stationarity in these systems, without requiring a generative model of the dynamics to be learned. This problem is referred to as Parameter Inference from a Non-stationary Unknown Process (PINUP). Here we introduce a PINUP method called feature-based Slow Feature Analysis (f -SFA) comprising the computation of time-series features across sliding windows, followed by dimension reduction using slow feature analysis (SFA). This allows us to detect slow variation in a potentially wide range of statistical properties of the measured dynamics on a timescale determined by the window length. Crucially, using a comprehensive time-series feature set avoids the subjectivity of feature selection, while the SFA slowness constraint overcomes the bias towards irrelevant correlated features seen with variance-based dimension reduction. The performance of f -SFA surpasses that of four benchmark PINUP methods across a diverse range of non-stationary chaotic processes, and we explore the impact of various parameters on performance, including observation noise, parameter timescales, parameter amplitudes, and unseen parameter values. Further, applying f -SFA to sleep polysomnography data, we show that it is able to infer a time-varying parameter underlying the non-stationary sleep recordings that closely tracks depth of sleep. To our knowledge, this work presents the first comparative study of PINUP methods, and we demonstrate that f -SFA is a simple, effective, and noise-robust approach for quantifying non-stationarity from time series, that can be applied in a range of fields. 2026-08-31T05:37:07Z Kieran S. Owens Masako Tamaki Ben D. Fulcher http://arxiv.org/abs/2608.30100v1 Unimodality and Radial Monotonicity of the Magnetic Resonance Fingerprinting T1/T2 Matching Objective 2026-08-31T00:17:45Z Magnetic resonance fingerprinting (MRF) estimates tissue parameters by matching an acquired MR signal time course to entries in a Bloch- or EPG-simulated dictionary. However, no study has yet proven the uniqueness of the matching results. In two earlier studies by exhaustive objective mapping I showed that for two widely used MRF sequences the normalized-correlation objective exhibits a single dominant peak at the true T1/T2 values and decreases smoothly away from that peak. These empirical properties motivated the fast MRF-ZOOM search algorithm even without using a pre-generated signal dictionary, but their theoretical basis has remained incomplete. The purpose of this work is to develop a mathematical framework for the MRF matching objective under normalized-correlation matching. 2026-08-31T00:17:45Z Ze Wang http://arxiv.org/abs/2608.15364v2 Determining low-$\ell$ p-mode frequency shifts in Sun-like stars: Enhancing the cross-correlation technique with filters 2026-08-30T17:35:06Z Acoustic mode frequencies in the Sun and Sun-like stars change due to magnetic activity, on time-scales much larger than the star's rotation and much smaller than its evolution. Given the poor S/N of the observed stellar p-modes, it is challenging to measure the changes of individual mode frequencies. Typically, power spectra of different time series segments are cross-correlated to estimate a mean p-mode frequency change, which ends up averaging over the individual mode contributions. We seek to enhance the cross-correlation method, by introducing a novel and computationally cheap method, thus enabling us to disentangle p-mode frequency changes for different spherical harmonic degree $\ell$. Assuming that the inclination angle and rotation rate are already measured, filters are designed, which enable the isolation of $δω_\ell$, frequency changes of modes with a given $\ell$, while preventing bias creeping in from neighbouring modes. Monte-Carlo simulations are performed to quantify uncertainty in the estimation of $δω_\ell$. We validate our method against well-studied solar data (SOHO/VIRGO and BiSON) and demonstrate its applicability to the solar-like Kepler star KIC 8006161. 2026-08-15T18:37:42Z 7 pages, 9 figures, accepted A&A Letters Samarth G. Kashyap Laurent Gizon Jesper Schou Rachel Howe http://arxiv.org/abs/2608.29760v1 A generalized likelihood model for segmented muon counters 2026-08-30T12:45:03Z Measurements of the muonic component of extensive air showers constrain cosmic-ray mass composition and hadronic interactions at energies beyond those accessible at accelerators. Arrays of segmented detectors with binary readout are widely used for this purpose: they sample the muon density at different distances from the shower core to reconstruct the muon lateral distribution function (LDF). Each detector response is summarized by the number of activated segments, $k$, whose probability distribution provides the likelihood relating the observation to the expected muon content. Signal pile-up, detector inefficiency, corner-clipping muons, and background signals shape this distribution, and neglecting them can bias the reconstruction. Existing analytical models include pile-up but otherwise assume an ideal detector response. In this work, we develop a unified statistical framework that incorporates inefficiency, corner clipping, and background through a small set of physically interpretable parameters. We derive exact expressions for the detector response and the likelihood required for muon-LDF reconstruction, together with a simple binomial approximation that preserves the main statistical properties of the exact distribution. Dedicated Monte Carlo simulations are used to assess the impact of the assumptions underlying the analytical treatment and show that it is negligible over the parameter range considered. They also show that the exact and approximate likelihoods yield similar performance in terms of estimator bias and confidence-interval coverage. Although motivated by the Underground Muon Detector of the Pierre Auger Observatory, the framework applies more broadly to segmented particle detectors with binary readout in which particle content is inferred from the number of activated segments. 2026-08-30T12:45:03Z 16 pages, 10 figures Joaquín de Jesús Juan Manuel Figueira Federico Sanchez Darko Veberic http://arxiv.org/abs/2608.29658v1 ButterMamba: Butterworth-Enhanced Spatial-Temporal Mamba for Efficient Traffic Flow Prediction 2026-08-30T08:46:11Z Accurate traffic flow prediction is fundamental to intelligent transportation systems, playing a pivotal role in urban mobility optimization and smart city development. While Graph Neural Networks (GNNs) integrated with time series forecasting have emerged as promising solutions, two critical limitations persist: (1) the quadratic complexity of attention-based architectures hinders real-time deployment in large-scale networks, and (2) high-frequency noise in sensor data significantly degrades prediction reliability. These challenges are particularly acute in metropolitan scenarios where both computational efficiency and noise robustness are paramount. To address these limitations, we introduce \textbf{ButterMamba}, a novel and efficient framework based on State Space Models (SSMs). ButterMamba consists of two key components: (1) a Butterworth Spectral Filtering module that preprocesses the data by removing high-frequency noise, allowing the model to focus on significant underlying trends, and (2) a Spatial-Temporal State Mixer that uses a parallel Mamba architecture to efficiently capture both long-range temporal dependencies and complex spatial correlations across the road network. By decoupling noise filtering from spatial-temporal modeling, ButterMamba achieves superior predictive accuracy with linear computational complexity. Extensive experiments on three public datasets demonstrate that ButterMamba not only outperforms existing state-of-the-art models in terms of prediction accuracy but also considerably reduces training time and memory usage. 2026-08-30T08:46:11Z 9 pages, 6 figures Limiao Zhang Yuhui Lu Jie Gao Hao Jiang Haiping Ma Xingyi Zhang http://arxiv.org/abs/2603.21247v2 Accelerate Vector Diffusion Maps by Landmarks 2026-08-30T06:56:26Z We propose a landmark-constrained algorithm, LA-VDM (Landmark Accelerated Vector Diffusion Maps), to accelerate the Vector Diffusion Maps (VDM) framework built upon the Graph Connection Laplacian (GCL), which captures pairwise connection relationships within complex datasets. LA-VDM introduces a novel two-stage normalization that effectively address nonuniform sampling densities in both the data and the landmark sets. Under a manifold model with the frame bundle structure, we show that we can accurately recover the parallel transport with landmark-constrained diffusion from a point cloud, and hence asymptotically LA-VDM converges to the connection Laplacian. The performance and accuracy of LA-VDM are demonstrated through experiments on simulated datasets and an application to nonlocal image denoising. 2026-03-22T14:09:55Z Sing-Yuan Yeh Yi-An Wu Hau-Tieng Wu Mao-Pei Tsui