https://arxiv.org/api/U9Sl/Tl6qJFzMGe+dvQFFG0fv6Y2026-09-10T21:18:14Z121456015http://arxiv.org/abs/2608.21491v2NestyNet. IV. Laws Chosen by Nothing in Advance2026-08-25T08:46:07ZDifferential-equation (DE) discovery tends to break down precisely where much of physics begins. Fields are coupled, governing laws are nonlinear in the state, amplitudes, coordinates, or operators of interest, yet derivatives must remain consistent across fields, channels, and differentiation orders. NestyNet-DE addresses this via two complementary DE search strategies, both employing analytic derivatives from segmented neural surrogates: a sparse-library route linear in the outer coefficients, and a new operator-factorized route that searches directly over equation structure rather than a fixed library, recovering compositional laws the sparse route misses. The recovered laws can be strongly nonlinear in states, fields, coordinates, and their couplings. The framework also handles multi-dataset shared-support discovery, complex and vector laws, and Hamiltonian discovery from phase-space trajectories. Beyond a law's form, the same data yield its geometry, the Lie point symmetries of the recovered equation. We apply the pipeline to 30 years of daily ephemerides of $308$ main-belt asteroids and recover the reduced-Kepler hierarchy (areal law, inverse-square force, and reduced Hamiltonian), where a discovered rotational symmetry fixes the centrifugal coefficient rather than fitting it. We also present a benchmark of 57 real-valued ODEs and 26 complex-valued systems, including the Schrödinger and Dirac equations, with Maxwell's equations as a coupled vector-PDE case study. Here, discovery is not shackled to a fixed library, nor does it end at the equation. It composes laws that no dictionary anticipated, and closes the loop from data, to a law chosen by nothing in advance, to the geometry that explains and integrates it.2026-08-21T12:57:24Z20 pages, 2 figuresRodrigo IbataWassim TenachiFoivos DiakogiannisNeil IbataAnirudh Shankarhttp://arxiv.org/abs/2608.21051v2NestyNet. III. Symbolic Regression from Analytic Neural Surrogates2026-08-25T08:42:50ZMany physical laws are simple only after the right representation, decomposition or internal coordinate has been found, but discovering that structure from data is combinatorially hard. This task is symbolic regression (SR), the search for closed-form expressions that fit data without assuming a fixed model class. Here we present NestyNet-SR. A neural surrogate with analytic derivatives is used to detect separability, recursively reducing multivariate problems to simpler neural atoms. These atoms are distilled into closed form by a tiered symbolic-search stack, whose final tier is a novel factorized symbolic search that separates structure from calibration. Composing candidate internal coordinates freely, it scores each coordinate by how well calibrated functions of it (e.g., polynomials, power laws, sinusoids) fit the data, so the constants of those calibrated maps, however deeply nested in the final expression, are fitted rather than searched. The method supports multi-dataset regression, automated feature discovery, and dimensional-analysis pruning. On the SRBench AI~Feynman benchmark, NestyNet-SR achieves exact symbolic recovery of all 120 noiseless equations, the first such result, and under noise a statistical audit certifies which structures survive. As a real-data vignette, given only the separate mass-model components of SPARC-survey galaxies, the algorithm discovers the baryonic acceleration coordinate, reproduces the established mass-to-light and acceleration scales and the non-unique form of the radial acceleration relation, and adds held-out-galaxy generalization, a calibrated symmetry abstention, and a posterior for the local slope of the law. Analytic derivatives thus provide a practical route from neural surrogates to interpretable closed-form empirical laws.2026-08-21T12:48:08Z24 pages, 6 figuresRodrigo IbataWassim TenachiFoivos DiakogiannisNeil IbataAnirudh Shankarhttp://arxiv.org/abs/2608.05862v2NestyNet. I. Physics Functions Are Hard to Fit with Neural Networks: A Framework for Accurate Surrogates and Analytic Derivatives2026-08-25T08:30:19ZMany of the smooth functions that matter most in physics are precisely the ones that standard neural network methods struggle to fit accurately. Here we present NestyNet, a coupled model-and-optimizer framework capable of fitting such targets to high accuracy while also delivering their gradients, Hessians, Laplacians, and antiderivatives analytically and at low cost. This makes it a natural substrate for scientific machine learning tasks. The model is a deterministic segmented analytic surrogate, and its optimizer is a second-order Levenberg--Marquardt scheme whose damping and linear solves are tailored to the stiff, strongly correlated parameter geometries induced by multiscale and sharply structured targets typical in physics and other scientific applications. On the AI Feynman benchmark of 120 physics equations, NestyNet achieves median improvement factors of $2\,100\times$ for function values, $1\,400\times$ for first derivatives, and $780\times$ for second derivatives relative to standard neural networks trained with first-order optimization (Adam). Even after refining those fits with quasi-Newton (L-BFGS) optimization, the corresponding improvements are $540\times$, $450\times$, and $250\times$. Owing to the analytic design it is up to $\approx 44\times$ faster than vectorized automatic-differentiation (autograd) baselines, with a margin growing with model size. The same analytic-derivative framework also supports vector- and complex-valued targets, measurement uncertainties in both inputs and outputs, and constraints, and its modules can be composed flexibly to build scientifically useful model architectures, all without reverting to autograd. Together, these components provide a practical modular framework for fitting difficult scientific surrogates while delivering accurate differential operators for subsequent analysis.2026-08-06T10:41:20Z25 pages, 6 figuresRodrigo IbataWassim TenachiFoivos DiakogiannisNeil IbataAnirudh Shankarhttp://arxiv.org/abs/2608.20956v2Identifying the structure of dynamical transitions in logistic map2026-08-25T08:18:43ZNonlinear dynamical systems manifest rich variety of dynamical states and transitions driven by fluctuations. To understand the pattern of fluctuations during dynamical transitions, we investigate the structural features of chaos to order transition in logistic map. We determine fluctuations as amplitude jumps and encode them onto a complex network where nodes represent amplitude levels and links represent transitions between distinct amplitude bins. We discover that global network measures identify points of period doubling, regimes of periodicity and chaos, including interior crises events. Using local network measures, we also unravel novel peculiar parabolic-shaped patterns in the orbit diagram that we show are reminiscent of the distribution of stable and unstable periodic points in the bifurcation diagram.2026-08-21T10:23:12ZAswin BalajiShruti TandonShwetha VisweshNorbert MarwanJuergen KurthsR. I. Sujithhttp://arxiv.org/abs/2608.24164v1Decoupling candidate dual AGN from chance superpositions in the GOTHIC survey via a deep-learning framework2026-08-25T07:30:33ZDual active galactic nuclei (DAGN) mark a critical phase in the evolution of merging galaxies and the pairing of supermassive black holes, yet they remain difficult to identify in large imaging surveys because of projection effects and limited spatial resolution. Compact foreground stars and unresolved substructure can mimic dual nuclei through chance superposition, complicating automated detection. We revisit the 46,061 galaxies flagged but rejected as DAGN candidates by the GOTHIC pipeline, primarily because the two nuclei fell within the SDSS fibre aperture or exceeded its separation threshold. We train a supervised deep-learning framework based on the YOLOv11 oriented-bounding-box architecture on annotated SDSS imaging to separate genuine dual nuclei from foreground stellar contaminants and other spurious alignments. The final model attains a validation precision of 0.919, recall of 0.905, and $F_1$ of 0.912 for the dual-nuclei class, and yields 29,605 dual-nucleus candidates after removing star-dominated and blended detections. Structured visual inspection indicates that $54.5$--$62\%$ are consistent with genuine dual nuclei, implying $\sim(1.4$--$1.8)\times10^{4}$ plausible systems. Cross-calibrating the YOLO separation against the deterministic GOTHIC centroid measurement and restricting to the compact regime ($d \le 6.87''$) gives a conservative subset of $\sim 13{,}672$ candidates, reaching calibrated separations of $\sim 0.56''$. Spectroscopy of the most compact ($\le 1$~kpc) systems shows they are dominated by passive, absorption-line galaxies with no resolved double-peaked emission, so confirmation requires higher-resolution follow-up. The catalogue is a statistically refined list of candidates, not confirmed DAGN. Nonetheless, deep-learning detection substantially reduces contamination and expands the plausible DAGN census.2026-08-25T07:30:33ZSubmitted to MNRAS. Supplementary Material merged in the main textBhavesh MukhejaSnehanshu SahaAnwesh BhattacharyaMousumi DasFrançoise CombesSudhanshu Barwayhttp://arxiv.org/abs/2512.14729v2Topological Dirac cluster synchronization on directed hypergraphs2026-08-24T21:40:48ZTopological higher-order synchronization reveals collective phenomena demonstrating how topology shapes dynamics, yet providing a general theoretical framework for designing and controlling dynamical states on higher-order networks remains a challenge of the field. Here we propose a dynamical systems framework for topological cluster synchronization on directed hypergraphs that can be used to design synchronization patterns on nodes and hyperedges of directed hypergraphs. By encoding oriented hyperedges into the hypergraph topological Dirac Hamiltonian, we obtain a spectrum whose isolated eigenstates correspond to distinct topological synchronization cluster states defined jointly on nodes and hyperedges. By selecting any isolated eigenstate, the system can be driven toward the associated dynamical state reflecting a specific partition of the hypergraph without modifying the underlying hypergraph structure. We numerically demonstrate the ability to design different topological cluster synchronization states on directed-hypergraph block models and empirical systems-including higher-order contact networks and the ABIDE functional brain network. Our results establish a general and interpretable route for controlling collective dynamics in directed higher-order systems.2025-12-10T04:32:40Z(22 pages, 9 figures)Yupeng GuoAhmed A. A. ZaidXueming LiuGinestra Bianconihttp://arxiv.org/abs/2606.15360v4Matched generating elements in maximum entropy density reconstruction2026-08-24T15:15:49ZMoment-constrained maximum entropy (MaxEnt) reconstructs a density as an exponential family whose sufficient statistics are chosen constraint functions. We study this choice separately from the numerical coordinates used to solve the dual problem. An odd-only family on a symmetric support forces f(x)f(-x) to be constant, so parity matching is necessary. A logarithmic-rational constraint produces the Student/Cauchy family and remains population-defined when required integer moments do not exist; a parity-matched one-parameter path provides a reproducible fractional-exponent alternative containing the monomials. The numerical study uses a common damped-Newton solver and audits the six-distribution benchmark of Rajan et al. Stable parity-block divided differences, residual replay in the original constraint coordinates, and independent adaptive integration identify a raw-coordinate false convergence on D5. Under the corrected computation, the selected PATP arm is within 1.22 times the oracle on D5 but retains selected-to-oracle error ratios of 30.7, 8.88, and 3.33 on D2, D4, and D6 at n=12. These checks preserve the negative selector conclusion without treating a numerical artifact as statistical evidence. The Pearson row is reproduced, whereas the published GOPoly rows are not under the printed protocol. Matched LogRat results show family capacity; neither tested selector supports automatic choice among constraint families.2026-06-13T15:45:18ZNumerical correction to the PATP arm: stable parity-block divided-difference coordinates with original-constraint replay replace a raw-power quadrature false convergence. D5 is resolved, while selected-to-oracle gaps remain on D2, D4, and D6, so automatic selection remains unresolved. 26 pages, three figures, eight tables. Code: https://github.com/SZabolotnii/Ku-MaxEnt-code-supplementSerhii Zabolotniihttp://arxiv.org/abs/2608.22873v1Differentiable Parametric Simulation and Reconstruction Models in Parnassus2026-08-24T06:59:34ZParnassus is a framework for fast detector simulation and reconstruction, directly mapping truth-level particles onto reconstructed objects. Such models can be built from deep generative networks trained on paired samples, which are fit automatically to a target detector, or from parametric prescriptions of the kind used by Delphes, which are constructed by hand. We remove this asymmetry by making the parametric models fully differentiable so that their parameters can be fit to a target sample by gradient descent. We demonstrate closure by fitting a parametric model to samples from a known configuration of itself, recovering the generating parameters and characterizing the degeneracies among them, and we present a first fit to CMS full simulation. The resulting models are interpretable, inexpensive, and run in the standard Parnassus pipeline which is fully Python based and GPU enabled.2026-08-24T06:59:34Z14 pages, 4 figuresAbdelrahman ElabdEilam GrossDmitrii KobylianskiiRunze LiBenjamin Nachmanhttp://arxiv.org/abs/2304.06522v3Dynamics-Based Intrinsic Signal Model for High-Dimensional, Small-Sample Data2026-08-24T02:10:01ZSignal extraction is difficult when the number of variables $N$ is much larger than the number of observations $M$. We address this problem under the working hypothesis that an empirical dataset consists of states sampled from underlying multivariate dynamics. Instead of treating the $M$ observations as points in an $N$-dimensional variable space, we treat the $N$ variables as points in an $M$-dimensional sample-coordinate space, interpreted as an effective time-delay coordinate space of the latent dynamics. This representation allows a large $N$ to provide many points for estimating the variable distribution even when $M$ is small. Transposed singular value decomposition (SVD) and the unsupervised feature-selection method of Taguchi are used to extract variable-side deviations from an estimated Gaussian background as signal candidates. As $M$ is reduced, the effective separation between sampled states increases; contributions from finite-correlation components are expected to decay, whereas sufficiently long-correlation components can persist toward the small-sample limit. We define these persistent components as intrinsic signals and estimate them by extrapolation toward $M=0$. We first tested the method on high-dimensional, small-sample data explicitly generated by a randomised coupling strength globally coupled map (RCS-GCM), for which the long- and short-correlation components were known. The extracted intrinsic signals corresponded to the known long-correlation variables. We then applied the method to The Cancer Genome Atlas (TCGA) pan-kidney gene-expression data, which are not ordinarily treated as data explicitly generated by a dynamical system. Using an SVD component associated with the pathologic-M category, variable-side signal components were extracted from the 20,531-dimensional data even under small-sample conditions.2023-04-13T13:18:36Z19 pages, 15 figuresYoh-ichi MototakeY-h. Taguchihttp://arxiv.org/abs/2608.22515v1Interpretable statistical feature engineering for early disruption prediction in the short pulse ADITYA tokamak2026-08-23T17:29:46ZReliable early disruption prediction is critical for the safe operation and real-time control of tokamaks. However, machine learning based prediction frameworks have predominantly targeted medium and long pulse devices, with comparatively limited attention given to short pulse tokamaks where available warning time is inherently constrained. In this work, an interpretable machine learning framework is developed for feature engineering and early prediction of disruptions in the ADITYA using the initial plasma evolution information, prior to the activation of the negative converter of the ohmic transformer power supply. Statistical descriptors comprising the mean, variance, skewness, kurtosis and wavelet energy entropy are extracted from routinely available plasma diagnostics over different operation time windows. Decision tree based feature selection is employed to identify physically meaningful disruption precursors and to reduce feature dimensionality. These selected features are used to train a random forest classifier. The proposed framework achieves stable predictive performance across different analysis windows, with a maximum ROC-AUC of 0.87 for 0-35 ms and 0-40 ms windows. Comparable and in some cases improved, performance is obtained using the reduced feature set, demonstrating that the selected statistical descriptors retain the essential information required for disruption prediction. The proposed methodology provides an interpretable and computationally efficient framework for real time disruption prediction in short pulse tokamaks and establishes that carefully engineered statistical descriptors can effectively replace raw time series inputs for early disruption prediction, thereby offering a practical pathway toward real time plasma control in short pulse tokamaks similar to ADITYA and ADITYA-U.2026-08-23T17:29:46Z27 pagesJyoti AgarwalKavit PatelBhaskar ChaudhuryAbhishek SharmaShrichand JakharManika Sharmahttp://arxiv.org/abs/2608.22463v1Noise Effects on Ordinal Pattern Statistics via Majorization2026-08-23T15:42:21ZThe Bandt-Pompe permutation entropy framework, alongside the complexity-entropy causality plane, has become a standard tool for characterizing the dynamical properties of time series. However, observational noise distorts ordinal pattern probability distributions in ways that can systematically misplace time series within the causality plane, compromising dynamical classification. This effect is particularly relevant for geophysical signals, which are typically poorly and irregularly sampled, and have a low signal-to-noise level. In this work, we characterize the distortions on ordinal pattern statistics using the formalism of majorization. We provide theoretical results and propose corrective strategies that restore discriminability under realistic measurement conditions. To achieve this, we introduce methodology that allows the characterization of noisy dynamical series and further allows the quantification of observational noise without the need of a fitting procedure. Finally, we illustrate our methodology by analyzing paleomagnetic records to determine if the geological evolution of the Earth dipole is better described by a stochastic or chaotic system.2026-08-23T15:42:21ZFacundo Sapienzahttp://arxiv.org/abs/2602.24282v2Unfolding without Iterations, Adversaries, or Surrogates2026-08-22T14:39:40ZCorrecting measurements for detector effects and constructing appropriate public data representations is a pressing problem in LHC physics. Current methods solve this inverse problem by relying on iterations, minimax optimization, or a surrogate forward mapping. We introduce Adversary-free Unfolding SanS Iteration or Emulation (AUSSIE), which dispenses with these mechanisms while remaining asymptotically correct. AUSSIE replaces the second OmniFold step with a new loss function that directly yields solutions with minimal dependence on the reference simulation. We showcase AUSSIE on various unfolding tasks, including full-phase-space jet substructure.2026-02-27T18:56:55Z24 pages, 14 figures, 4 tables. v2: Accepted in SciPost PhysicsSciPost Phys. 21, 055 (2026)Ayodele OreTilman Plehn10.21468/SciPostPhys.21.3.055http://arxiv.org/abs/2405.07359v2Forecasting with an N-dimensional Langevin Equation and a Neural-Ordinary Differential Equation2026-08-21T17:22:47ZAccurate prediction of electricity day-ahead prices is essential in competitive electricity markets. Although stationary electricity-price forecasting techniques have received considerable attention, research on non-stationary methods is comparatively scarce, despite the common prevalence of non-stationary features in electricity markets. Specifically, existing non-stationary techniques will often aim to address individual non-stationary features in isolation, leaving aside the exploration of concurrent multiple non-stationary effects. Our overarching objective here is the formulation of a framework to systematically model and forecast non-stationary electricity-price time series, encompassing the broader scope of non-stationary behavior. For this purpose we develop a data-driven model that combines an N-dimensional Langevin equation (LE) with a neural-ordinary differential equation (NODE). The LE captures fine-grained details of the electricity-price behavior in stationary regimes but is inadequate for non-stationary conditions. To overcome this inherent limitation, we adopt a NODE approach to learn, and at the same time predict, the difference between the actual electricity-price time series and the simulated price trajectories generated by the LE. By learning this difference, the NODE reconstructs the non-stationary components of the time series that the LE is not able to capture. We exemplify the effectiveness of our framework using the Spanish electricity day-ahead market as a prototypical case study. Our findings reveal that the NODE nicely complements the LE, providing a comprehensive strategy to tackle both stationary and non-stationary electricity-price behavior. The framework's dependability and robustness is demonstrated through different non-stationary scenarios by comparing it against a range of basic naive methods.2024-05-12T18:45:30Z26 pages, 7 figuresChaos, 34, 043105 (2024)Antonio Malpica-MoralesMiguel A. Durán-OlivenciaSerafim Kalliadasis10.1063/5.0189402http://arxiv.org/abs/2609.00020v1Chern--Simons Fluctuations and Information Geometry in Discrete Electromagnetism2026-08-21T13:35:26ZWe construct helicity-conditioned statistical states for a Whitney-discretized electromagnetic field on a closed oriented three-manifold. The simplicial de Rham complex provides exact discrete gauge symmetry, while the Whitney inner product separates exact, harmonic, and coexact sectors. The spatial Abelian Chern--Simons functional is gauge invariant and depends only on the coexact potential. After fixing harmonic modes, we introduce a helicity-biased Gaussian ensemble on the reduced electromagnetic phase space and derive explicit formulas for its admissible parameters, partition function, mean helicity, relative entropy, and Fisher information. The distribution uniquely minimizes relative entropy under a prescribed mean-helicity constraint. Its helicity susceptibility equals the variance of the discrete Chern--Simons functional and controls the local distinguishability of neighboring statistical states. Because magnetic helicity is generally not conserved under unconstrained Maxwell dynamics, these states represent conditioned inference rather than dynamical equilibrium.2026-08-21T13:35:26ZJean-Pierre Magnothttp://arxiv.org/abs/2512.12142v2MeltwaterBench: Deep learning for spatiotemporal downscaling of surface meltwater2026-08-20T20:23:27ZThe Greenland ice sheet is melting at an accelerated rate due to processes that are not fully understood and hard to measure. The distribution of surface meltwater can help understand these processes and is observable through remote sensing, but current maps of meltwater face a trade-off: They are either high-resolution in time or space, but not both. We develop a deep learning model that creates gridded surface meltwater maps at daily 100m resolution by fusing data streams from remote sensing observations and physics-based models. In particular, we spatiotemporally downscale regional climate model (RCM) outputs using synthetic aperture radar (SAR), passive microwave (PMW), and a digital elevation model (DEM) over the Helheim Glacier in Eastern Greenland from 2017-2023. Using SAR-derived meltwater as "ground truth", we show that a deep learning-based method that fuses all data streams is over 10 percentage points more accurate over our study area than existing non deep learning-based approaches that only rely on a regional climate model (83% vs. 95% Acc.) or passive microwave observations (72% vs. 95% Acc.). Alternatively, creating a gridded product through a running window calculation with SAR data underestimates extreme melt events, but also achieves notable accuracy (90%) and does not rely on deep learning. We evaluate standard deep learning methods (UNet and DeepLabv3+), and publish our spatiotemporally aligned dataset as a benchmark, MeltwaterBench, for intercomparisons with more complex data-driven downscaling methods. The code and data are available at github.com/blutjens/hrmelt.2025-12-13T02:43:05Zin review at Journal of Advances in Modeling Earth SystemsBjörn LütjensPatrick AlexanderRaf AntwerpenTil WidmannGuido CervoneMarco Tedesco