https://arxiv.org/api/XTwVUTGNPQSIl9mk+kiucYi8r0M2026-07-20T22:51:40Z628253015http://arxiv.org/abs/2510.11539v5Simultaneous Calibration of Noise Covariance and Kinematics for State Estimation of Legged Robots via Bi-level Optimization2026-07-17T03:49:57ZAccurate state estimation is critical for legged and aerial robots operating in dynamic, uncertain environments. A key challenge lies in specifying process and measurement noise covariances, which are typically unknown or manually tuned. In this work, we introduce a bi-level optimization framework that jointly calibrates covariance matrices and kinematic parameters in an estimator-in-the-loop manner. The upper level treats noise covariances and model parameters as optimization variables, while the lower level executes a full-information estimator. Differentiating through the estimator allows direct optimization of trajectory-level objectives, resulting in accurate and consistent state estimates. We validate our approach on quadrupedal and humanoid robots, demonstrating significantly improved estimation accuracy and uncertainty calibration compared to hand-tuned baselines. Our method unifies state estimation, sensor, and kinematics calibration into a principled, data-driven framework applicable across diverse robotic platforms.2025-10-13T15:39:21ZDenglin ChengJiarong KangXiaobin Xionghttp://arxiv.org/abs/2404.13301v3Sequential subspace methods on Stiefel manifold optimization2026-07-17T02:59:43ZWe investigate the minimization of a quadratic function over Stiefel manifolds (the set of all orthogonal $r$- frames in $\mathbf{R}^n$), which has applications in high-dimensional semi-supervised classification tasks. To reduce the computational complexity, we employ sequential subspace methods(SSM) to transform the high-dimensional problem to a series of low-dimensional ones. In this paper, our goal is to achieve an optimal solution of high quality, referred to as a ''qualified critical point". Qualified critical points are defined as those where the associated multiplier matrix meets specific upper-bound conditions. These points exhibit near-global optimality in quadratic optimization problems.
In the context of a general quadratic, SSM generates a sequence of qualified critical points through low-dimensional surrogate regularized models. The convergence to a qualified critical point is guaranteed, when each SSM subspace is constructed from the following vectors: (i) a set of orthogonal unit vectors associated with the current iterate, (ii) a set of vectors representing the gradient of the objective, and (iii) a set of eigenvectors links to the smallest $r$ eigenvalues of the system matrix. Furthermore, incorporating Newton direction vectors into the subspaces can significantly accelerate the convergence of SSM.2024-04-20T07:14:17Z29 pagesPengwen ChenChung-Kuan ChengChester Holtzhttp://arxiv.org/abs/2512.02769v2Reinforcement learning for irreversible reinsurance problems: the randomized singular control approach2026-07-17T01:56:39ZThis paper studies the continuous-time reinforcement learning for stochastic singular control with the application to an infinite-horizon irreversible reinsurance problem. The singular control is equivalently characterized as a pair of regions of time and the augmented states, called the singular control law. To encourage the exploration in the learning procedure, we propose a randomization method by considering an auxiliary singular control and entropy regularization. The exploratory singular control problem is formulated as a two-stage optimal control problem, in which the time-inconsistency issue arises in the outer problem. Existence of equilibrium singular control law for the time-inconsistent outer problem is rigorously established. Taking advantage of the solution structure, we utilize a proper parameterization and neural networks to devise the actor-critic reinforcement learning algorithm. In the numerical experiment, we show the superior convergence of parameter iterations based on the randomized equilibrium policy and illustrate how the exploration may advance the learning performance.2025-12-02T13:49:13ZZongxia LiangXiaodong LuoXiang Yuhttp://arxiv.org/abs/2507.15234v2An Optimization-Based Framework for Solving Forward-Backward Stochastic Differential Equations: Convergence Analysis and Error Bounds2026-07-17T01:37:41ZForward-backward stochastic differential equations have recently become a key focus in the computational field, and their role in continuous-time stochastic optimal control and reinforcement learning has grown increasingly prominent. In this paper, we develop an optimization-based framework for solving coupled forward-backward stochastic differential equations, which naturally arise in stochastic optimal control through the stochastic maximum principle and related Hamiltonian systems. We introduce an integral-form objective function and prove its equivalence to the error between consecutive Picard iterates. Our convergence analysis establishes that minimizing this objective generates sequences that converge to the true solution. We provide explicit upper and lower bounds that relate the objective value to the error between trial and exact solutions. We validate the proposed objective and its theoretical interpretation using two analytical test cases, and further illustrate its numerical applicability on a nonlinear stochastic optimal control problem with up to 1000 dimensions.2025-07-21T04:30:17ZYutian WangYuan-Hua NiXun Lihttp://arxiv.org/abs/2607.15532v1Logic, Optimization, and Artificial Intelligence2026-07-17T00:46:08ZLogic and optimization can, in combination, make valuable contributions to rule-based AI. Logic is the obvious medium for encoding a rule base and drawing inferences from it, while optimization provides a powerful technology for computing inferences. Their combination has taken on new relevance amid a growing concern for transparency in AI. which is important for reproducibility, explainability, trustworthiness, and fairness. Rule-based AI provides a natural solution to transparency that is becoming increasingly practical due to today's highly advanced optimization methods. This article surveys several areas of logic-optimization partnership, including probabilistic logic, Bayesian logic, belief logics and Dempster-Shafer theory, nonmonotonic (default) logic, many-valued logics, and inference of logical formulas from noisy data based on Boolean regression. It shows how to compute projections, the fundamental problem of both logic and optimization, using decision diagrams and logic-based Benders decomposition. It describes the use of postoptimality analysis to explain how conclusions are reached, further enhancing transparency, as well as the role of optimization in answer set programming modulo theories. The paper concludes by suggesting possible future research directions.2026-07-17T00:46:08ZJ. N. Hookerhttp://arxiv.org/abs/2603.17401v2Dynamical Properties of Safety Filters for Linear Systems and Affine Control Barrier Functions2026-07-17T00:16:13ZThis letter studies the dynamical properties of safety filters designed based on Control Barrier Functions (CBF). This mechanism, which is popular in safety-critical applications, takes a nominal controller and minimally modifies it to render it safe. Although CBF-based safety filters make the closed-loop system safe, characterizing their additional dynamical properties, such as stability, boundedness, or existence of spurious equilibria, remains a challenging problem. Here, we address this problem for the case of linear systems and an affine CBF constraint. We provide conditions under which the closed-loop system presents undesired equilibria, unbounded trajectories, or the origin is globally exponentially stable.2026-03-18T06:23:00ZPol MestresShima Sadat MousaviAaron D. Ameshttp://arxiv.org/abs/2607.15505v1Fast and Scalable Caputo Fractional Gradient Descent via Perturbation-Preserving Memory Compression2026-07-16T23:27:12ZFractional gradient descent (FGD) incorporates long-range memory through Caputo-type operators and has been shown to improve stability in ill-conditioned and nonconvex optimization problems. Despite these advantages, its practical use remains limited, mainly due to the high computational cost of evaluating history-dependent convolutions, which scales quadratically with the number of iterations. In this paper, we focus on making Caputo-based optimization computationally viable without sacrificing its intrinsic memory structure. We begin by expressing the fractional descent direction as a discrete convolution over past gradients, which provides a unified view of the method. Based on this formulation, we introduce two complementary mechanisms to reduce the cost of the memory term. The first uses a sum-of-exponentials (SOE) approximation of the power-law kernel, leading to efficient recursive updates. The second approach, newly proposed in this paper as dyadic hierarchical discrete convolution (DHDC), compresses the gradient history through a multiscale aggregation strategy. Rather than treating these approximations as purely numerical accelerations, we interpret them as perturbations of the ideal Caputo operator. This viewpoint allows us to analyze how the compressed memory affects the optimization dynamics. Under standard $μ$-strong convexity and $L$-smoothness assumptions, we show that the resulting method still exhibits monotone descent and linear convergence, provided that the approximation error remains controlled.2026-07-16T23:27:12ZHwanseo LeeJunseo LeeHyunju Kimhttp://arxiv.org/abs/2511.05227v3Semiconvexity of (weak) Kantorovich potentials in the Lorentzian optimal transport problem2026-07-16T22:06:36ZWe study semiconvexity properties of (weak) Kantorovich potentials for the Lorentzian optimal transport problem with the standard cost function $c$. We show that, in general, this regularity - known in the Riemannian context - does not extend to the Lorentzian setting. Nevertheless, we provide a general regularity result for $c$-convex functions and show that, under suitable general assumptions on the measures, this yields semiconvexity of the (weak) potentials at least on an open set of full measure. This, in turn, allows us to conclude the existence and uniqueness of an optimal transport map.2025-11-07T13:23:40ZAlec Metschhttp://arxiv.org/abs/2607.15481v1Tropical Bi-Objective Pseudolinear Optimization as Parametric Mean-Payoff Games2026-07-16T21:56:48ZWe extend the parametric mean-payoff game framework of Parsons et al. to bi-objective tropical pseudolinear optimization with general two-sided constraints. The problem is to simultaneously minimize two tropical pseudolinear objectives over the feasible set of a general two-sided system U otimes x oplus b is less than or equal to V otimes x oplus d, we characterize the Pareto front via a parametric mean-payoff game in two parameters ( lambda 1, lambda 2). The feasibility region R is convex and the Pareto front P is a convex piecewise-linear curve with finitely many breakpoints, these properties are natural extensions of the single-parameter case to two parameters. In addition, we give as a new result, the joint denominator bound: the cycle coefficients satisfy k 1( gamma ) + k 2( gamma ) is less than or equal to 2 for any elementary cycle gamma, yielding | Delta | is less than or equal to 2 for every 2 times 2 Newton system, except in the fully decoupled case, and implying that all breakpoints have half-integer coordinates for integer data. Optimality and infeasibility certificates are given in terms of the cycle structure of the parametric game. Two algorithms are developed, a directional bisection algorithm (O(n squared (n+m) log M) per direction) and a Newton scheme tracing the complete Pareto front via 2 times 2 linear solves in at most | S | steps, independent of M. The directional bisection algorithm is pseudo-polynomial in n, m and M. The Newton scheme is independent of M but requires up to |S| steps, where |S| is exponential in n. Lastly, we give numerical experiments on random instances to confirm the directional bisection complexity bound exactly; the Newton scheme's worst-case bound is not attained by random instances but is shown to be tight via explicit adversarial constructions.2026-07-16T21:56:48Z38 pages, 8 figuresIbrahim HassanAbdulhadi AminuLawal Muhammadhttp://arxiv.org/abs/2607.15471v1A Unified Variational Framework for Optimal Transport with Lagrangian Costs2026-07-16T21:41:12ZWe investigate optimal transport distances induced by general Lagrangian action functionals. Extending the classical Monge - Kantorovich and Benamou - Brenier theories, we derive a unified variational framework that connects several equivalent formulations of the induced transport distance, including Lagrangian, Eulerian, convex optimization, Hamilton - Jacobi dual, and Hamiltonian flow formulations. Under standard convexity assumptions on the Lagrangian, we establish the equivalence of these formulations through variational arguments and convex duality. The resulting optimality system reveals a natural Hamiltonian structure on the Wasserstein space, providing a direct link between optimal transport, Hamiltonian dynamics, and optimal control. The proposed framework extends the classical quadratic-cost theory to general Lagrangian costs and offers a unified perspective for the analysis of transport metrics generated by action functionals.2026-07-16T21:41:12Z19 pagesHailiang Liuhttp://arxiv.org/abs/2607.15470v1pyoptexplain: A Python Library for Post-Optimality Analysis and Explanation of Optimization Models2026-07-16T21:41:09ZOptimization models are built in a variety of modeling languages and solved by a variety of solvers, but once a solution exists, the information needed to understand it is fragmented: each solver exposes a partial, differently named set of native diagnostics, and the modeling language has already canonicalized the formulation the user wrote. We present pyoptexplain, a practitioner-first Python library for post-optimality analysis of optimization models that sits above this layer. It adapts a model authored in any of five modeling front ends, namely cvxpy, Pyomo, gurobipy, docplex, and OR-Tools, into a normalized internal representation, solves it through a choice of backends, and answers the why and what-if questions of an optimization decision through one uniform interface. The design rests on two observations. First, a post-optimality quantity requested from different backends for the same problem can come back as an exception, a structurally meaningless zero or a basis-dependent value that disagrees across solvers, so reporting whatever one solver returns is unreliable. pyoptexplain reports a quantity only when both the representation and the chosen backend can justify it, and does not approximate unavailable information. Second, repeated scenario analysis can amortize its cost by extracting the model once and reusing a warm solver session across a batch of scenarios. pyoptexplain builds a single scalable what-if interface, uniform across its modeling languages and backends and returning a certified report for every scenario, at a cost within a small constant factor of the bare solver. A reproducible computational study substantiates both claims. Source code is available at https://github.com/h-fellahi/pyoptexplain and installation can be done through the Python Package Index https://pypi.org/project/pyoptexplain/.2026-07-16T21:41:09ZHussein Fellahihttp://arxiv.org/abs/2512.24069v2Time-varying Mixing Matrix Design for Energy-efficient Decentralized Federated Learning2026-07-16T20:58:35ZWe consider the design of mixing matrices to minimize the operation cost for decentralized federated learning (DFL) in wireless networks, with focus on minimizing the maximum per-node energy consumption. As a critical hyperparameter for DFL, the mixing matrix controls both the convergence rate and the needs of agent-to-agent communications, and has thus been studied extensively. However, existing designs mostly focused on minimizing the communication time, leaving open the minimization of per-node energy consumption that is critical for energy-constrained devices. This work addresses this gap through a theoretically-justified solution for mixing matrix design that aims at minimizing the maximum per-node energy consumption until convergence, while taking into account the broadcast nature of wireless communications. Based on a novel convergence theorem that allows arbitrarily time-varying mixing matrices, we propose a multi-phase design framework that activates time-varying communication topologies under optimized budgets to trade off the per-iteration energy consumption and the convergence rate while balancing the energy consumption across nodes. Our evaluations based on real data have validated the efficacy of the proposed solution in combining the low energy consumption of sparse mixing matrices and the fast convergence of dense mixing matrices.2025-12-30T08:24:28ZXusheng ZhangTuan NguyenTing Hehttp://arxiv.org/abs/2607.15452v1All Games Have Equilibria2026-07-16T20:47:39ZResearch on Nash equilibrium existence for infinite games has grown into a patchwork of technical preconditions and counterexamples. This paper presents a unified program in equilibrium theory by revising the predominant model of mixed strategies based on countable additivity. A game is specified by a nonempty set of players and, for each player, a nonempty action set and a bounded von Neumann-Morgenstern utility function. Every such game is shown to admit a Nash equilibrium in finitely additive mixed strategies. In addition, the equilibrium correspondence for any such game is shown to be nonempty, compact-valued, and upper hemicontinuous, and the same is true for equilibria obtained as limits of finite approximations. Techniques developed in this paper show that infinite games long treated as intractable become amenable to direct equilibrium analysis.2026-07-16T20:47:39ZM. Ali KhanArthur Paul PedersenMaxwell B. Stinchcombehttp://arxiv.org/abs/2607.15432v1Proactive Inpatient Bed Requests for Emergency Department Admissions2026-07-16T20:06:05ZEmergency department (ED) boarding occurs when admitted patients remain in the ED while awaiting inpatient beds. Boarding is a major driver of ED crowding and has been associated with poor patient outcomes. We propose a framework to help EDs reduce boarding time and length of stay by using information about current patients and bed availability to proactively request inpatient beds before admission decisions are finalized.
We formulate the problem as a Markov decision process in which predictions of each patient's admission probability and time to disposition are aggregated to guide early inpatient bed requests. This formulation leads to three data-driven policies based on approximate dynamic programming, reinforcement learning, and a newsvendor-type approach. Using a simulation model based on data from a large ED, we evaluate these policies across a wide range of settings. The simulation study shows that proactive aggregate bed requests can reduce average boarding times for admitted patients by 30-70\% and average length of stay for all ED patients by 6-15\%, while creating only modest idle time for prepared inpatient beds. The newsvendor heuristic provides the most attractive tradeoff between ED performance and inpatient bed idle time, whereas the reinforcement learning heuristic produces smoother bed-request patterns when stability in downstream hospital processes is especially important.
Our work shows how EDs can use prediction tools to make proactive bed-request decisions that improve ED operations while helping managers balance reductions in ED delays against inpatient bed idle time. Our findings also illustrate the value of evaluating both simple myopic heuristics and more sophisticated reinforcement learning-based approaches, since each can offer distinct advantages depending on the performance measures and implementation constraints most important to managers.2026-07-16T20:06:05Z51 pages, 8 figuresQIan ChengNilay Tanik ArgonAniruddhan GanesaramanSerhan Ziyahttp://arxiv.org/abs/2607.15420v1Harnessing GPU Acceleration in Large-Scale Process Optimization2026-07-16T19:51:21ZThis paper presents a proof-of-concept workflow for equation-oriented process optimization that runs entirely on a GPU. Process optimization models often incorporate complex interconnected unit operations, dynamics, and uncertainties, resulting in large nonlinear programs that can be computationally demanding for conventional CPU-based solvers. Although emerging GPU-based solvers offer substantial computational benefits, their application to process optimization has been limited by the lack of GPU-compatible process modeling tools. We address this gap by prototyping the GPU-compatible process optimization models using an existing GPU-capable optimization software stack, including ExaModels (algebraic modeling system), MadNLP (optimization solver), and cuDSS (linear solver). ExaModels formulates the process optimization problem in a GPU-compatible way by exposing its repeated algebraic structure, while MadNLP and cuDSS solve the resulting nonlinear program on the GPU. This workflow is demonstrated on a CO2 absorber design problem under feed uncertainty, in which a shared column diameter is minimized subject to equilibrium and hydraulic constraints in all scenarios. For the largest case with 5,000 scenarios and 1.5 million variables, the GPU workflow achieves a speedup of approximately 21\times over a single-threaded CPU baseline using JuMP, Ipopt, and MA57.2026-07-16T19:51:21ZSubmitted to FOCAPO-CPC 2027Boxun HuangDavid Y. ShuMichel SchanenMihai AnitescuRahul GandhiSungho Shin