https://arxiv.org/api/kTPna3Wm9KhmzXAqXBd025eF3Zs2026-09-10T19:11:04Z121453015http://arxiv.org/abs/2511.15853v2The Ensemble Kalman Inversion Race2026-09-01T19:42:02ZEnsemble Kalman methods were initially developed to solve nonlinear data assimilation problems in oceanography but are now popular in applications far beyond their original use cases. Of particular interest is climate model calibration. As hybrid physics and machine-learning models advance, the number of parameters and complexity of parameterizations in climate models will continue to grow. To fully realize these advances, we must move from laborious hand-tuning to calibration-driven model development in rapid iteration cycles. Thus, robust calibration of these parameters plays an increasingly important role. We focus on learning climate model parameters by minimizing the misfit between modeled and observed climate statistics in an idealized setting. Ensemble Kalman methods are a natural choice for this problem because they are derivative-free, scalable to high dimensions, and robust to noise caused by statistical observations. Given the many variants of ensemble methods proposed, an important question is: Which ensemble Kalman method should be used for climate model calibration? To answer this question, we perform systematic numerical experiments to explore the relative computational efficiencies of several ensemble Kalman methods. The numerical experiments involve statistical observations of Lorenz-type models of increasing complexity, frequently used to represent simplified atmospheric systems, with some featuring neural network parameterizations. For each test problem, several ensemble Kalman methods and a derivative-based method "race" to reach a specified accuracy, and we measure the computational cost required to achieve the desired accuracy. We investigate how prior information and the parameter or data dimensions define the computational costs of the various methods.2025-11-19T20:18:40ZJournal of Advances in Modeling Earth Systems (2026), 18, e2025MS005629Rebecca GjiniMatthias MorzfeldOliver R. A. DunbarTapio Schneider10.1029/2025MS005629http://arxiv.org/abs/2604.11481v3Emergence of Complex Web Structures2026-09-01T14:09:40ZComplex structures often emerge from initially homogeneous or weakly correlated states. We address the apparent tension between this ordering and entropy growth through a unified framework combining semi-microscopic phase-space dynamics, transport geometry, information theory, and coarse-grained effective modeling. The key point is that entropy depends on the level of description: a coarse-grained spatial field may become more ordered as structure forms, even while the full phase-space description becomes more complex through shell crossing, multistreaming, and the activation of velocity degrees of freedom. Using a Lagrangian--Eulerian transport map, we show how density amplification is governed by the Jacobian of the deformation and how anisotropic collapse arises from the eigenvalues of a hierarchy of deformation tensors. Long-range interaction or information flow is encoded in the displacement field, so that nonlocality enters directly through transport. We connect this geometric description to a maximum-entropy Gaussian baseline and show how nonlinear transport and nonlocal coupling generate scale coupling, higher-order correlations, and non-Gaussianity. We then formulate a Landau--Ginzburg description in which the growth of seed anisotropies is interpreted as the activation of lower effective free-energy branches, providing a coarse-grained realization of self-organization. Applied to generated cosmological fields, this framework indicates that the nonlocal tidal level becomes relevant already at moderate overdensity. Although cosmological structure formation is the main realization considered here, the framework is intended more broadly as a mesoscopic language for systems in which transport, anisotropy, nonlocality, and self-organization are central.2026-04-13T13:47:35Z38 pages, 8 figures, 1 table, acceptedFrancisco-Shu Kitaurahttp://arxiv.org/abs/2509.21246v2Peak separation methods for inverse photoelectron spectra: Comparing second derivative, curve fitting, and deconvolution analyses2026-09-01T09:10:18ZInverse photoelectron spectroscopy (IPES) is a powerful technique for probing the unoccupied electronic states of materials. It can be regarded as the inversion process of photoelectron spectroscopy (PES), which examines the occupied states. Recently developed low-energy inverse photoelectron spectroscopy (LEIPS) can significantly advance the study of unoccupied states, owing to minimal sample damage and suppressed dark counts compared to conventional IPES. However, the instrumental resolution remains at 0.2 eV, which is one order of magnitude lower than that of PES. Spectral broadening caused by the low instrumental resolution often results in overlapping peaks. Peak separation is therefore crucial in the analysis of LEIPS spectra. In this study, we compared three peak separation methods-second derivative, curve fitting, and deconvolution. These methods were applied to modeled and experimental LEIPS spectra of the lowest unoccupied molecular orbital-derived band of pentacene, which consists of two splitting peaks due to the two inequivalent molecules in the unit cell. We systematically and quantitatively evaluated the performance of each method in terms of analysis parameters and discussed its robustness to noise as well as its peak separation capabilities as a function of peak energy spacing and intensity ratio. This work offers a practical framework for peak separation in LEIPS, with extensions to PES and a wide range of spectroscopies.2025-09-25T14:39:26Z16 pages, 7 figures. Published in Review of Scientific InstrumentsRev. Sci. Instrum. 97, 023906 (2026)Ryotaro NakazawaHaruki SatoHiroyuki Yoshida10.1063/5.0303140http://arxiv.org/abs/2609.00883v1iPINN for Broadband CARS Phase Retrieval: A Framework for Function Approximation and Inverse Modeling Problems in Nonlinear Spectroscopy2026-09-01T08:17:10ZPhase retrieval in broadband coherent anti-Stokes Raman spectroscopy (BCARS) is an ill-posed inverse problem. The Raman-like signal is encoded in the imaginary part of the resonant susceptibility, which mixes coherently with a non-resonant background (NRB) that varies across acquisitions. We introduce an inverse physics-informed neural network (iPINN) that predicts Lorentzian peak parameters from raw BCARS spectra and reconstructs the resonant susceptibility through a differentiable analytical forward model. A transformer encoder assigns spectral features to 24 learnable peak slots, and a multi-view consistency loss enforces invariance across NRB pattern, NRB strength, and noise. Unlike direct spectral regression approaches, the method retains accuracy under varying acquisition conditions. On a public benchmark, iPINN achieves the lowest error among the tested baselines (MAE 0.016 vs. next-best 0.046). On 28 zero-shot test spectra acquired across seven solvents and four focal positions, accuracy is depth-invariant in five of seven solvents. These results show that inverse parametric prediction with a differentiable physical decoder supports robust phase retrieval across measurement conditions.2026-09-01T08:17:10ZRavi Teja VulchiCarl MesserschmidtMohammadsadegh VafaeinezhadRajendhar JunjuriTobias Meyer-ZedlerJuergen PoppThomas Bocklitzhttp://arxiv.org/abs/2609.01674v1Multi-fidelity Monte Carlo estimation of floor response spectra under combined seismic and structural parameter uncertainties2026-09-01T08:03:10ZFloor response spectra (FRS) are essential tools for the design of non-structural elements (such as equipment or components). Given the various physical phenomena influencing FRS, high-fidelity (HF) mechanical models of the primary structure may be required to estimate them. Since numerical simulations based on such models are generally computationally expensive, this paper proposes using a multi-fidelity Monte Carlo (MFMC) approach for the efficient estimation of FRS. The method relies on using observations from a fast low-fidelity (LF) model as control variables. If the absolute value of the correlation between LF and HF samples is close to 1, this approach reduces both variance and estimation error compared to a standard Monte Carlo estimate based solely on HF data samples. Through a case study involving the reactor building of the Kashiwazaki-Kariwa nuclear power plant, we demonstrate the suitability of this method for FRS estimation. It effectively reduces variance and estimation error, even when using a LF model as simple as a single-degree-of-freedom system. We also show that the method accounts for modeling uncertainties while maintaining comparable performance. Its ease of use makes it a valuable tool for practitioners.2026-09-01T08:03:10ZNils BaillieBaptiste KerleguerCyril FeauJosselin GarnierFan Wanghttp://arxiv.org/abs/2601.11415v2Auditing Frozen-Encoder Anomaly Detection Across Mechanical Systems: Representation Provenance, Calibration, and Protocol Effects2026-09-01T06:31:06ZThis version reports a reproducibility audit of the frozen-encoder experiments presented in version 1. The numerical discrimination results are reproducible from the preserved artifacts, but their original attribution to interferometric pretraining is not supported. The released checkpoint contains a nested model state that loads without missing parameters, whereas loading the outer checkpoint dictionary leaves almost the entire EfficientNet-B0 feature stack uninitialized. Preserved embeddings labelled as interferometric have norms of order $10^{-12}$, matching freshly initialized EfficientNet-B0 networks and differing by more than twelve orders of magnitude from the preserved ImageNet embeddings. A second, separately preserved near-zero embedding set produces almost the same IMS 4th-test anomaly scores ($r=0.987$) and record-level discrimination (AUC $0.9812$ versus $0.9818$).
We therefore withdraw the causal claim that IMS performance demonstrates a morphological prior transferred from gravitational-wave instrumentation. We reanalyse the controlled IMS splits at matched observed false-positive rates and add multivariate classical signal baselines. The near-zero representations retain strong tail separation, particularly in the 2nd and 4th IMS runs, but this is now interpreted as an exploratory architecture-and-initialization effect coupled to Mahalanobis scoring. A separate PRONOSTIA audit shows that the original large warning times were induced by a lifetime-fraction baseline; under fixed-time evaluation, a ten-feature classical baseline outperforms the preserved encoder scores. These results illustrate how checkpoint provenance, finite-sample calibration, architecture, and target-domain baselines can create an appearance of cross-domain transfer. They also define the controls required before assigning physical meaning to frozen-representation anomaly scores.2026-01-16T16:35:07Z6 pages, 2 figures, 4 tablesJose Sánchez Andreuhttp://arxiv.org/abs/2609.00132v1Anomaly detection for multijet scenarios2026-08-31T18:00:01ZSignals of physics beyond the Standard Model continue to resist discovery at the LHC. Recent years have seen the proliferation of new anomaly detection techniques, promising discovery with significantly fewer model assumptions than traditional approaches. Strategies based on weak supervision have been especially successful, but so far were largely limited in scope by their reliance on a resonance manifesting a decay into a pair of jets. In this work, we demonstrate that a well-established idea from jet substructure physics - recursive soft drop - in combination with the CATHODE technique for anomaly detection can be used to simultaneously perform anomaly detection for signals with an arbitrary number of jets in the final state, greatly increasing the scope of such searches.2026-08-31T18:00:01Z14 pages, 7 figures, 2 tablesGregor KasieczkaSung Hak LimLouis MoureauxTore von SchwartzDavid ShihChitrakshee Yedehttp://arxiv.org/abs/2608.30838v1Cosmic variance and ergodicity in finite systems with correlations2026-08-31T14:08:50ZWe consider the difference between ensemble and volume average in cosmology. It is known that for sufficiently weak long-range correlations the root mean square of the difference, which we call ergodicity bias, decays like $R^{-3/2}$ in the limit of large volume $R^3$. We calculate the condition this imposes on the power spectrum of a Gaussian random field. We quantify the bias for finite $R$, and show that the $R\to\infty$ limit is of little relevance for cosmological observations when the measured scales and correlations extend to the size of the observable universe. We consider curvature, density, and velocity perturbations. On large scales the bias is important in all three cases. For the density perturbations, which are observationally the most relevant, the relative bias first exceeds 100% at the separation $r=177$ Mpc, and is larger than 100% for all $r>560$ Mpc. It should be taken into account when comparing ensemble and volume averages for large-scale structure. The bias is also large for the cosmic microwave background temperature perturbations on large angular scales, but this is not relevant for observations, as their analysis does not involve volume averaging.2026-08-31T14:08:50Z18+3 pages, 16 figuresDipayan MukherjeeSyksy Rasanenhttp://arxiv.org/abs/2609.00094v1Modernization and Statistical Validation of a Multilayer TRIM.SP Code for Low-Energy Muon and Ion Implantation2026-08-31T14:01:24ZWe describe the staged modernization of a local multilayer TRIM.SP implementation, denoted TRIM.SP-NL. The work included behavior-preserving refactoring, correction of localized bookkeeping and initialization defects, support for more than five elements per layer, reproducible build modes, and a runtime-selectable random-number-generator (RNG) interface. The RANLUX RNG remains the default reference backend, while PCG32 and xoshiro256** provide faster alternatives without changing the input-deck format. The code validation combined exact regression testing with statistical comparisons over 13 representative target multilayer configurations, comprising 143 energy-configuration points per backend and 429 simulations in total. Across the test suite, PCG32 and xoshiro256** reduced runtime by factors of 1.98 and 2.01, respectively. The implanted, backscattered, and transmitted fractions, as well as the mean implantation depths, range straggling, and representative depth profiles showed no systematic RNG dependence. Together with a Node.js-based graphical interface for setup, scans, and result collection, the Fortran engine forms the open-source TRIM.SP-NL Workbench.2026-08-31T14:01:24Z9 pages, 5 figuresZaher SalmanRyan M. L. McFaddenThomas Prokschahttp://arxiv.org/abs/2609.01651v1Parameter inference from a non-stationary unknown process using statistical feature-based slow feature analysis2026-08-31T05:37:07ZNon-stationary phenomena are ubiquitous, with examples to be found in climatological measurements, brain activity, and the behavior of financial markets. Starting with a time series from a non-stationary process, a key challenge is to infer the time-varying parameters that underlie the non-stationarity in these systems, without requiring a generative model of the dynamics to be learned. This problem is referred to as Parameter Inference from a Non-stationary Unknown Process (PINUP). Here we introduce a PINUP method called feature-based Slow Feature Analysis (f -SFA) comprising the computation of time-series features across sliding windows, followed by dimension reduction using slow feature analysis (SFA). This allows us to detect slow variation in a potentially wide range of statistical properties of the measured dynamics on a timescale determined by the window length. Crucially, using a comprehensive time-series feature set avoids the subjectivity of feature selection, while the SFA slowness constraint overcomes the bias towards irrelevant correlated features seen with variance-based dimension reduction. The performance of f -SFA surpasses that of four benchmark PINUP methods across a diverse range of non-stationary chaotic processes, and we explore the impact of various parameters on performance, including observation noise, parameter timescales, parameter amplitudes, and unseen parameter values. Further, applying f -SFA to sleep polysomnography data, we show that it is able to infer a time-varying parameter underlying the non-stationary sleep recordings that closely tracks depth of sleep. To our knowledge, this work presents the first comparative study of PINUP methods, and we demonstrate that f -SFA is a simple, effective, and noise-robust approach for quantifying non-stationarity from time series, that can be applied in a range of fields.2026-08-31T05:37:07ZKieran S. OwensMasako TamakiBen D. Fulcherhttp://arxiv.org/abs/2608.30100v1Unimodality and Radial Monotonicity of the Magnetic Resonance Fingerprinting T1/T2 Matching Objective2026-08-31T00:17:45ZMagnetic resonance fingerprinting (MRF) estimates tissue parameters by matching an acquired MR signal time course to entries in a Bloch- or EPG-simulated dictionary. However, no study has yet proven the uniqueness of the matching results. In two earlier studies by exhaustive objective mapping I showed that for two widely used MRF sequences the normalized-correlation objective exhibits a single dominant peak at the true T1/T2 values and decreases smoothly away from that peak. These empirical properties motivated the fast MRF-ZOOM search algorithm even without using a pre-generated signal dictionary, but their theoretical basis has remained incomplete. The purpose of this work is to develop a mathematical framework for the MRF matching objective under normalized-correlation matching.2026-08-31T00:17:45ZZe Wanghttp://arxiv.org/abs/2608.15364v2Determining low-$\ell$ p-mode frequency shifts in Sun-like stars: Enhancing the cross-correlation technique with filters2026-08-30T17:35:06ZAcoustic mode frequencies in the Sun and Sun-like stars change due to magnetic activity, on time-scales much larger than the star's rotation and much smaller than its evolution. Given the poor S/N of the observed stellar p-modes, it is challenging to measure the changes of individual mode frequencies. Typically, power spectra of different time series segments are cross-correlated to estimate a mean p-mode frequency change, which ends up averaging over the individual mode contributions. We seek to enhance the cross-correlation method, by introducing a novel and computationally cheap method, thus enabling us to disentangle p-mode frequency changes for different spherical harmonic degree $\ell$. Assuming that the inclination angle and rotation rate are already measured, filters are designed, which enable the isolation of $δω_\ell$, frequency changes of modes with a given $\ell$, while preventing bias creeping in from neighbouring modes. Monte-Carlo simulations are performed to quantify uncertainty in the estimation of $δω_\ell$. We validate our method against well-studied solar data (SOHO/VIRGO and BiSON) and demonstrate its applicability to the solar-like Kepler star KIC 8006161.2026-08-15T18:37:42Z7 pages, 9 figures, accepted A&A LettersSamarth G. KashyapLaurent GizonJesper SchouRachel Howehttp://arxiv.org/abs/2608.29760v1A generalized likelihood model for segmented muon counters2026-08-30T12:45:03ZMeasurements of the muonic component of extensive air showers constrain cosmic-ray mass composition and hadronic interactions at energies beyond those accessible at accelerators. Arrays of segmented detectors with binary readout are widely used for this purpose: they sample the muon density at different distances from the shower core to reconstruct the muon lateral distribution function (LDF). Each detector response is summarized by the number of activated segments, $k$, whose probability distribution provides the likelihood relating the observation to the expected muon content. Signal pile-up, detector inefficiency, corner-clipping muons, and background signals shape this distribution, and neglecting them can bias the reconstruction. Existing analytical models include pile-up but otherwise assume an ideal detector response. In this work, we develop a unified statistical framework that incorporates inefficiency, corner clipping, and background through a small set of physically interpretable parameters. We derive exact expressions for the detector response and the likelihood required for muon-LDF reconstruction, together with a simple binomial approximation that preserves the main statistical properties of the exact distribution. Dedicated Monte Carlo simulations are used to assess the impact of the assumptions underlying the analytical treatment and show that it is negligible over the parameter range considered. They also show that the exact and approximate likelihoods yield similar performance in terms of estimator bias and confidence-interval coverage. Although motivated by the Underground Muon Detector of the Pierre Auger Observatory, the framework applies more broadly to segmented particle detectors with binary readout in which particle content is inferred from the number of activated segments.2026-08-30T12:45:03Z16 pages, 10 figuresJoaquín de JesúsJuan Manuel FigueiraFederico SanchezDarko Veberichttp://arxiv.org/abs/2608.29658v1ButterMamba: Butterworth-Enhanced Spatial-Temporal Mamba for Efficient Traffic Flow Prediction2026-08-30T08:46:11ZAccurate traffic flow prediction is fundamental to intelligent transportation systems, playing a pivotal role in urban mobility optimization and smart city development. While Graph Neural Networks (GNNs) integrated with time series forecasting have emerged as promising solutions, two critical limitations persist: (1) the quadratic complexity of attention-based architectures hinders real-time deployment in large-scale networks, and (2) high-frequency noise in sensor data significantly degrades prediction reliability. These challenges are particularly acute in metropolitan scenarios where both computational efficiency and noise robustness are paramount. To address these limitations, we introduce \textbf{ButterMamba}, a novel and efficient framework based on State Space Models (SSMs). ButterMamba consists of two key components: (1) a Butterworth Spectral Filtering module that preprocesses the data by removing high-frequency noise, allowing the model to focus on significant underlying trends, and (2) a Spatial-Temporal State Mixer that uses a parallel Mamba architecture to efficiently capture both long-range temporal dependencies and complex spatial correlations across the road network. By decoupling noise filtering from spatial-temporal modeling, ButterMamba achieves superior predictive accuracy with linear computational complexity. Extensive experiments on three public datasets demonstrate that ButterMamba not only outperforms existing state-of-the-art models in terms of prediction accuracy but also considerably reduces training time and memory usage.2026-08-30T08:46:11Z9 pages, 6 figuresLimiao ZhangYuhui LuJie GaoHao JiangHaiping MaXingyi Zhanghttp://arxiv.org/abs/2603.21247v2Accelerate Vector Diffusion Maps by Landmarks2026-08-30T06:56:26ZWe propose a landmark-constrained algorithm, LA-VDM (Landmark Accelerated Vector Diffusion Maps), to accelerate the Vector Diffusion Maps (VDM) framework built upon the Graph Connection Laplacian (GCL), which captures pairwise connection relationships within complex datasets. LA-VDM introduces a novel two-stage normalization that effectively address nonuniform sampling densities in both the data and the landmark sets. Under a manifold model with the frame bundle structure, we show that we can accurately recover the parallel transport with landmark-constrained diffusion from a point cloud, and hence asymptotically LA-VDM converges to the connection Laplacian. The performance and accuracy of LA-VDM are demonstrated through experiments on simulated datasets and an application to nonlocal image denoising.2026-03-22T14:09:55ZSing-Yuan YehYi-An WuHau-Tieng WuMao-Pei Tsui