https://arxiv.org/api/RCqO3PcCugRi+tAguq490shG4Ps2026-09-12T21:51:13Z112164515http://arxiv.org/abs/2608.23977v2IC-ThermBench: An Open, Progressive Benchmark for Generalizable 2.5D/3D-IC Thermal Learning2026-09-07T02:38:42ZStandardized benchmarks are fundamental to reliable progress in AI for EDA, including learning-based thermal modeling. However, existing thermal prediction studies often rely on different datasets, simulators, data splits, preprocessing pipelines, and metrics, while most datasets and implementations remain unavailable, making fair and reproducible comparison difficult. We introduce IC-ThermBench, an open and progressive benchmark that combines established 3D-IC steady-state, transient, and industrial package tasks with a new 50,000-sample 2.5D chiplet extension designed to evaluate progressively broader represented physical variation and cross-package OOD transfer. Five Generalization Scopes cover 3D-IC fixed-design prediction, Within-Family Generalization under layout, material, and boundary-condition variation, and Cross-Package OOD transfer to unseen package systems. We evaluate eight representative baselines under common data, splits, labels, and metrics. Performance degrades gradually from S2 to S4 as represented physical support broadens, but Cross-Package OOD produces a much sharper degradation: the best RMSE and MAE increase from 1.216 and 0.938~K at S4 to 15.99 and 15.00~K at S5, respectively. With only 10 labeled samples per OOD case, target-domain adaptation reduces the best MAE to 2.60~K. IC-ThermBench further provides a unified generation, training, inference, and evaluation pipeline, enabling reproducible and fair comparison of existing and new thermal predictor.2026-08-25T02:10:20ZProject site: https://github.com/Day333/ThermalBenchDavid HangWenkai YangKuiye DingHaiyang XinJacky Weihttp://arxiv.org/abs/2609.06874v1MedGSSR: Generalizable Medical Image Super-Resolution 3D Reconstruction via Hierarchical Feed-forward Gaussian Splatting2026-09-06T23:33:02ZHigh-resolution volumetric medical imaging is critical for clinical diagnosis, yet acquisition is often limited by scanner hardware, scan time, and for CT, radiation dose. Medical 3D Super-Resolution (Med3DSR) offers a computational alternative, but existing methods commonly rely on per-subject optimization, pretrained priors, or coordinate-based implicit representations, which compromise anatomical fidelity and limit efficiency. To address these limitations, we present MedGSSR, a fully end-to-end feed-forward framework that represents volumes as an explicit 3D Gaussian field for Med3DSR. Unlike coordinate-based implicit functions, our explicit 3D Gaussian representation naturally enhances signal continuity and local high-frequency fidelity. Specifically, MedGSSR explicitly decouples the reconstruction process into coarse-grained structural preservation and fine-grained textural refinement through the proposed Pyramid Anatomical Encoder and a Hierarchical Gaussian Projector. To support arbitrary-scale super-resolution, we introduce sub-voxel Gaussian decomposition and a Differentiable Gaussian Voxelizer that directly queries the continuous 3D intensity field, reducing discretization artifacts. Extensive experiments on MRI and CT benchmarks demonstrate that MedGSSR significantly outperforms state-of-the-art methods. Notably, our framework exhibits robust generalizability across unseen datasets without requiring per-subject optimization, enabling fast inference and high-fidelity volumetric super-resolution in practical clinical settings. Our project webpage, including code, is at https://william2ai.github.io/medgssr2026-09-06T23:33:02ZECCV 2026Chengkai WangLuoyu HongYiting ZhaoJiamin WangXiang FengFeiwei QinZhenzhong KuangXuefei YinAli BashashatiYanming Zhuhttp://arxiv.org/abs/2603.26517v2Hyperelastic constitutive model discovery with differentiable finite elements and structure-preserving neural networks2026-09-06T09:47:36ZThe discovery of constitutive laws from experimentally accessible measurements is a central problem in nonlinear computational mechanics. Many data-driven constitutive identification approaches rely either on paired strain-stress data or on full-field displacement measurements, both of which are difficult to obtain in realistic three-dimensional settings. We present a differentiable finite element framework for the discovery of hyperelastic material laws from partial observations, including boundary-only displacement measurements and global reaction forces. The method embeds the nonlinear finite element equilibrium problem directly into the learning loop, so that candidate strain-energy densities are assessed through the deformation fields and reactions they induce. This formulation enforces mechanical equilibrium as a constraint and allows the loss function to be evaluated only at observed locations. To ensure physical admissibility and promote numerical solvability throughout training, the constitutive response is represented by Hyperelastic Neural Networks, a structure-preserving neural class that enforces residual energy and stress-free conditions, frame indifference, isotropic material symmetry, polyconvexity, coercivity, and controlled volumetric growth by construction. The resulting PDE-constrained learning problem is solved using a quasi-Newton strategy combined with continuation and solver-aware backtracking. Numerical experiments in two- and three-dimensional finite elasticity demonstrate accurate recovery of hyperelastic isotropic responses from boundary-only data, robustness to measurement noise, and generalization across geometries, loading conditions, and boundary conditions.2026-03-27T15:27:04ZComputational Mechanics, 2026Francesco Regazzoni10.1007/s00466-026-02850-2http://arxiv.org/abs/2503.21450v6CMADiff: Cross-Modal Aligned Diffusion for Controllable Protein Generation2026-09-06T03:16:36ZAI-assisted protein design has emerged as a critical tool for advancing biotechnology, as deep generative models have demonstrated their reliability in this domain. However, most existing models primarily utilize protein sequence or structural data for training, neglecting the physicochemical properties of proteins.Moreover, they are deficient to control the generation of proteins in intuitive conditions. To address these limitations,we propose CMADiff here, a novel framework that enables controllable protein generation by aligning the physicochemical properties of protein sequences with text-based descriptions through a latent diffusion process. Specifically, CMADiff employs a Conditional Variational Autoencoder (CVAE) to integrate physicochemical features as conditional input, forming a robust latent space that captures biological traits. In this latent space, we apply a conditional diffusion process, which is guided by BioAligner, a contrastive learning-based module that aligns text descriptions with protein features, enabling text-driven control over protein sequence generation. Validated by a series of evaluations including AlphaFold3, the experimental results indicate that CMADiff outperforms protein sequence generation benchmarks and holds strong potential for future applications. The implementation and code are available at https://github.com/HPC-NEAU/PhysChemDiff.2025-03-27T12:41:48ZFurther improvement is needed in the contentChangjian ZhouYuexi QiuXinyue ZhangJiaqing ZhangHeng-Da ChengJiafeng LiJia SongWensheng Xianghttp://arxiv.org/abs/2609.06314v1A review of weakly enforced Dirichlet boundary conditions in computational flow analysis2026-09-06T00:26:36ZStrongly enforced Dirichlet boundary conditions require highly refined near-wall meshes to resolve steep velocity and thermal gradients. This introduces high computational costs, especially for practical flow simulations. Weakly enforced boundary conditions alleviate this burden by acting as a variationally consistent near-wall model. By allowing a controlled slip at the solid wall, weak enforcement recovers accurate flow quantities on coarse boundary-layer meshes across both incompressible and compressible regimes. Furthermore, weak boundary conditions serve as the fundamental enabling technology for immersogeometric analysis. Because the weak operator evaluates boundary integrals independently of the background mesh, high-fidelity flow analysis can be performed directly on complex geometries without fitting a mesh to the surface. This capability has facilitated direct geometry-to-analysis workflows for boundary-representation CAD models, raw point clouds, photogrammetric reconstructions, and segmented medical images. This review examines the unified mathematical development of the weak boundary condition framework from scalar advection-diffusion equations to the full Navier-Stokes equations, illustrating its versatility and robustness across incompressible and compressible flows, whether in traditional boundary-fitted, sliding-interface, or advanced immersogeometric applications.2026-09-06T00:26:36ZMing-Chen HsuMonu JaiswalYuri Bazilevshttp://arxiv.org/abs/2606.07463v2Amortized Neural Optimization for Pre-Layout Signal Integrity Design Space Exploration using Differentiable Surrogates2026-09-05T21:31:38ZPre-layout design space exploration (DSE) for high-speed signal integrity (SI) analysis is often limited by the computational cost of simulations and iterative optimization algorithms within modern electronic design automation (EDA) workflows. While machine learning surrogate models accelerate the simulation step, optimizing designs still requires utilizing iterative black-box search methods. This iterative nature scales poorly and makes multi-corner sweeps computationally expensive. As a solution, this paper proposes amortized neural optimization (ANO) for pre-layout SI design. ANO avoids iterative black-box inference and utilizes a fully differentiable neural network surrogate model from which it extracts gradient information to train a global optimization policy. This does not solve the optimization problem repeatedly at inference, but learns the process offline, which amortizes the computational cost. Once the ANO policy is trained, it maps different channel contexts to near-optimal design parameters in a single deterministic forward pass. The efficiency and accuracy of the ANO framework are demonstrated based on three SI design scenarios, including DDR5 decision feedback equalization (DFE), 9-dimensional SerDes Tx/Rx co-equalization, and DDR3 DQS differential pair eye diagram optimization under intra-pair skew constraints. Trading roughly 10% in optimality compared to instance-specific black-box algorithms results in speedups of three to four orders of magnitude. For a large-scale 320,000-instance multi-corner SerDes sweep optimization, ANO reduces what would have taken days of computation using iterative search algorithms to a single batched forward pass that completes in milliseconds. This transforms computationally expensive SI optimization into real-time and interactive pre-layout DSE.2026-06-05T17:13:44Z16 pages, 20 figures, 8 tables. This work has been submitted to the IEEE for possible publication. Copyright may be transferred without notice, after which this version may no longer be accessibleJulian WithöftWerner JohnEmre EcikRalf BrüningJürgen Götzehttp://arxiv.org/abs/2608.25369v2Forecasting Global Volatility with Predictive Spillover Networks: A Neuro-Econometric Spatio-Temporal Transformer for Asynchronous Financial Markets2026-09-05T15:28:02ZForecasting global realized volatility requires a model that can learn from interconnected markets without treating zero-coded exchange closures as observed zero volatility. We develop PGA-Trans-HAR, a neuro-econometric architecture that combines a rolling ridge-VAR/GFEVD predictive-connectedness network, masked spatio-temporal attention, and a frozen HAR anchor. Missing observations used to estimate the rolling econometric prior are completed only within the trailing information set available at the forecast origin. An asymmetric source mask prevents closed markets from transmitting zero-coded closure signals. A learned, market-specific gate allocates weight between the econometric prior and data-driven spatial attention. A bounded inverse-softplus correction then refines the HAR forecast while preserving positivity. We evaluate eight international equity indices from 2006 to 2022 at 1-, 5-, and 22-union-calendar-day forecast leads, where the target is the one-day realized volatility observed at the corresponding future date. The design uses five-seed ensembles, select-and-refit estimation, structural ablations, HAC-adjusted Diebold--Mariano tests, and block-bootstrap Model Confidence Sets. PGA-Trans-HAR records the lowest cross-market average MAE at the 1-day forecast lead and the lowest average MSE and MAE at the 5- and 22-union-calendar-day forecast leads. Relative to HAR, both losses decline for all eight markets at the 1- and 5-day forecast leads and for seven markets at the 22-day forecast lead. The evidence shows that combining an origin-aligned econometric network with masked attention can improve multi-market volatility forecasts in asynchronous financial environments.2026-08-26T04:41:50ZXinlin ZhaoHaotian Qiaohttp://arxiv.org/abs/2603.28325v4Building evidence-based knowledge bases from full-text literature for disease-specific biomedical reasoning2026-09-05T04:28:38ZBiomedical knowledge resources often either preserve evidence as unstructured text or compress it into flat triples that omit study design, provenance, and quantitative support. Here we present EvidenceNet, a disease-specific dataset of record-level evidence collections and corresponding graph representations derived from full-text biomedical literature. EvidenceNet uses a large language model (LLM)-assisted pipeline to extract experimentally grounded findings as structured evidence records, normalize biomedical entities, score evidence quality, and connect related records through typed semantic relations. We release EvidenceNet-HCC with 7,872 evidence records and a corresponding graph with 10,328 nodes and 49,756 edges, and EvidenceNet-CRC with 6,622 records and a corresponding graph with 8,795 nodes and 39,361 edges. Technical validation shows high component fidelity, including 98.3% field-level extraction accuracy, 100.0% high-confidence entity-link accuracy, 87.5% fusion integrity, and 90.0% semantic relation-type accuracy. Downstream analyses show that the data support retrieval-augmented question answering and graph-based tasks such as future link prediction and target prioritization. These results establish EvidenceNet as a disease-specific biomedical knowledge base dataset for evidence-aware analysis and reuse.2026-03-30T11:53:45Z30 pages, 5 figures, 12 tablesChang ZongJinyu ChenSicheng LvSi-tu XueHuilin ZhengJian WanLei Zhanghttp://arxiv.org/abs/2609.05663v1What LLM Trading Agents Actually Do in Production: A Six-Month, Population-Scale Record from Two Fleets2026-09-04T18:56:31ZWe present a continuous, population-scale measurement record of autonomous language-model trading agents operating in production across two systems with one design lineage: DX Terminal Pro (3,505 user-funded vaults trading real ETH in Base memecoin markets for 21 days, February to March 2026) and the DXAP live alpha fleet (500 to 599 user-created agents all-history, 91 to 117 concurrently active, trading Hyperliquid perpetuals, June to August 2026). The record spans roughly six months, 7.5M single-model invocations with about 300K onchain actions, and a further 231,638 multi-tool turns producing 14,596 fills. Four findings carry the paper. First, the operating layer determines behavior more than anything written in strategy text: a risk slider explains leverage (+0.425 per level), agent fixed effects absorb 60% of variance, and a leaderboard render boundary causally routes selection (regression discontinuity 1.75x at the top-3 cut). Second, sizing is volatility-blind: median leverage is 5.0x in every volatility sextile, and one posture-slider cell (11% of the book) holds 62% of liquidations. Third, agents capture almost none of the upside they reach: 43.2% of positions saw at least +300 bps of favorable excursion within 24h, yet 49.3% of those closed with a negative trade return; a mechanical bracket recovers +39.0 bps per position. Fourth, neither fleet shows a directional edge. The DXAP fleet is not profitable and trails a matched Hyperliquid retail benchmark (41% vs. 50% roundtrip win rate). A paired-replay league of frontier models on 416 captured production scenarios finds decision quality statistically indistinguishable at this horizon, while choice stability differs sharply across model families. Every headline survives day-clustered inference, permutation nulls, and a common-fee restatement; the paper closes with a 17-rule methodology canon bought with our own retractions.2026-09-04T18:56:31Z17 pages, 8 figures, 3 tables. Artifacts, figures, and aggregate data: https://github.com/ProjectDXAI/continuous-record-llm-trading-agentsT. J. BartonChris ConstantakisPatti HausemanAnnie MousAlaska HoffmanBrian BergeronHunter Goodreauhttp://arxiv.org/abs/2606.24556v2Development of a Programming Based Kinetic Model for Two Stage Composting of Solid Waste2026-09-04T18:11:13ZAs the world is moving toward sustainable development, there is an important need to adopt sustainable waste management solutions of biodegradable solid waste, such as composting, which offers significant advantages over traditional methods like landfilling and incineration by reducing greenhouse gas emissions, enriching soil fertility, and minimizing landfill waste. However, optimizing the composting process governed by factors like aeration, moisture, and carbon-to-nitrogen ratio often relies on complex mathematical models that are difficult to interpret and apply. To overcome this challenge, a user-friendly programming-based two-stage kinetic model has been developed which simplifies composting efficiency analysis, with the first stage covering the initial 28 days of decomposition and the second stage evaluating further degradation, making the process more accessible and actionable for sustainable waste management.2026-06-23T13:23:51ZZarif Tanzim AzizMd. Mahfuzur RahmanQuazi Hamidul Barihttp://arxiv.org/abs/2509.12596v2A Computational Pipeline for Patient-Specific Modeling of Thoracic Aortic Aneurysm: From Medical Image to Finite Element Analysis2026-09-04T17:02:52ZThe aorta is the body's largest arterial vessel, serving as the primary pathway for oxygenated blood within the systemic circulation. Aortic aneurysms consistently rank among the top twenty causes of mortality in the United States. Thoracic aortic aneurysm (TAA) arises from abnormal dilation of the thoracic aorta and remains a clinically significant disease, ranking as one of the leading causes of death in adults. A thoracic aortic aneurysm ruptures when the integrity of all aortic wall layers is compromised due to elevated blood pressure. Currently, three-dimensional computed tomography (3D CT) is considered the gold standard for diagnosing TAA. The geometric characteristics of the aorta, which can be quantified from medical imaging, and stresses on the aortic wall, which can be obtained by finite element analysis (FEA), are critical in evaluating the risk of rupture and dissection. Deep learning based image segmentation has emerged as a reliable method for extracting anatomical regions of interest from medical images. Voxel based segmentation masks of anatomical structures are typically converted into structured mesh representation to enable accurate simulation. Hexahedral meshes are commonly used in finite element simulations of the aorta due to their computational efficiency and superior simulation accuracy. Due to anatomical variability, patient specific modeling enables detailed assessment of individual anatomical and biomechanics behaviors, supporting precise simulations, accurate diagnoses, and personalized treatment strategies. Finite element (FE) simulations provide valuable insights into the biomechanical behaviors of tissues and organs in clinical studies. Developing accurate FE models represents a crucial initial step in establishing a patient-specific, biomechanically based framework for predicting the risk of TAA.2025-09-16T02:50:06ZJiasong ChenLinchen QianRuonan GongChristina SunTongran QinThuy PhamCaitlin MartinMohammad ZafarJohn ElefteriadesWei SunLiang Lianghttp://arxiv.org/abs/2503.08904v3Towards Efficient Parametric State Estimation in Circulating Fuel Reactors with Shallow Recurrent Decoder Networks2026-09-04T10:48:47ZThe recent developments in data-driven methods have paved the way to new methodologies to provide accurate state reconstruction of engineering systems; nuclear reactors represent particularly challenging applications for this task due to the complexity of the strongly coupled physics involved and the extremely harsh and hostile environments, especially for new technologies such as Generation-IV reactors. Data-driven techniques can combine different sources of information, including computational proxy models and local noisy measurements on the system, to robustly estimate the state. This work leverages the novel Shallow Recurrent Decoder architecture to infer the entire state vector (including neutron fluxes, precursors concentrations, temperature, pressure and velocity) of a reactor from three out-of-core time-series neutron flux measurements alone. In particular, this work extends the standard architecture to treat parametric time-series data, ensuring the possibility of investigating different accidental scenarios and showing the capabilities of this approach to provide an accurate state estimation in various operating conditions. This paper considers as a test case the Molten Salt Fast Reactor (MSFR), a Generation-IV reactor concept, characterised by strong coupling between the neutronics and the thermal hydraulics due to the liquid nature of the fuel. The promising results of this work are further strengthened by the possibility of quantifying the uncertainty associated with the state estimation, due to the considerably low training cost. The accurate reconstruction of every characteristic field in real-time makes this approach suitable for monitoring and control purposes in the framework of a reactor digital twin.2025-03-11T21:32:28ZStefano RivaCarolina IntroiniJ. Nathan KutzAntonio Cammihttp://arxiv.org/abs/2510.12368v2Constrained Sensing and Reliable State Estimation with Shallow Recurrent Decoders on a TRIGA Mark II Reactor2026-09-04T10:47:41ZShallow Recurrent Decoder networks are a novel data-driven methodology able to provide accurate state estimation in engineering systems, such as nuclear reactors. This deep learning architecture is a robust technique designed to map the temporal trajectories of a few sparse measures to the full state space, including unobservable fields, which is agnostic to sensor positions and able to handle noisy data through an ensemble strategy, leveraging the short training times and without the need for hyperparameter tuning. The architecture was successfully applied to the Molten Salt Fast Reactor concept; now, this work considers the performance of Shallow Recurrent Decoders on a deployed reactor concept. The underlying model is represented by a fluid dynamics model of the TRIGA Mark II research reactor; the architecture will use both synthetic temperature data coming from the numerical model and leveraging experimental temperature data recorded during a previous campaign. The objectives of this work are therefore: 1) presenting the first application of SHRED to a deployed nuclear reactor (TRIGA Mark II); 2) the integration of hybrid synthetic/experimental data within the SHRED framework; 3) a systematic assessment of SHRED robustness under physically constrained, low-dynamics sensor locations; and 4) a quantitative evaluation of SHRED self-correction capability when model-to-data discrepancies are present. Indeed, this approach is capable of accurately reconstruct every field of interest in real-time, using both synthetic (with an average relative error lower than 4\% in euclidean norm) and experimental data (with a RMSE of 1.52 K on the temperature, compared of 1.85 K of the CFD), making it suitable for interpretable monitoring and control purposes in the context of building digital twins for nuclear reactors.2025-10-14T10:31:34ZStefano RivaCarolina IntroiniJosè Nathan KutzAntonio Cammihttp://arxiv.org/abs/2608.04464v2Trie-Constrained Token Prediction with Hierarchy-Aware Semantic Alignment for HS Code Prediction2026-09-04T04:12:42ZHarmonized System (HS) code prediction (HSP) from commodity text is essential to international trade, and its importance continues to grow in port logistics. For the purposes of such prediction, recently, large language models (LLMs) have been actively investigated, owing especially to their strong language-understanding capabilities. However, their high computational cost limits deployment in constrained environments such as container terminals. Small language models (SLMs) offer a practical alternative, but their smaller scale makes them prone to generating invalid HS codes and to overlooking the hierarchical semantics between commodity text and HS codes. To address these limitations, this study proposes TRIE-HSA, which combines trie-constrained token prediction with hierarchy-aware semantic alignment (HSA). This method constrains the SLM to predict only valid digits under the HS taxonomy and aligns commodity text representations with the hierarchical structure of HS codes. In extensive experiments on data collected from an operational container terminal, TRIE-HSA improved average HS6 accuracy by 49.96 %p over zero-shot inference and exceeded the strongest task-specific benchmark by 11.94 %p. These results demonstrate that accurate and structurally valid HSP is achievable with fewer than 10 billion parameters. Therefore, TRIE-HSA offers a practical basis for deployment of HSP in port logistics operations that cannot support largescale LLMs.2026-08-05T05:38:50ZMinseop KimTaekhyun ParkKikun ParkHyerim Baehttp://arxiv.org/abs/2602.13769v4OR-Agent: Bridging Evolutionary Search and Structured Research for Automated Heuristic Design2026-09-04T02:55:55ZAutomating heuristic design in complex, experiment-driven domains requires more than iterative mutation of solution algorithms. Current LLM-based evolutionary methods often rely on stochastic mutation loops that lack long-term strategic planning and a formal mechanism to learn from historical failures, leading to inefficient exploration and redundant trials. To address this, we present OR-Agent, a multi-agent research framework designed for automated heuristic design in optimization problems with rich experimental environments. OR-Agent organizes heuristic search as tree-based workflow that explicitly models branching hypothesis generation and systematic backtracking. Furthermore, to address the lack of adaptive learning in current agents, we introduce a hierarchical, optimization-inspired reflection system in which short-term reflections act as verbal gradients, long-term reflections as verbal momentum, and memory compression as semantic weight decay - collectively forming a principled mechanism for governing research dynamics. Extensive experiments on classical combinatorial optimization problems (e.g., TSP, CVRP, bin packing) and simulation-based cooperative driving scenarios demonstrate that OR-Agent outperforms strong evolutionary search baselines. All code and experimental data are publicly available at https://github.com/qiliuchn/OR-Agent.2026-02-14T13:32:03ZQi LiuRuochen HaoCan LiWanjing Ma