https://arxiv.org/api/qzah/ghydlzC7GlBnTTl86rtpX0 2026-07-23T00:18:59Z 2414 180 15 http://arxiv.org/abs/2603.09219v1 AlgoXpert Alpha Research Framework. A Rigorous IS WFA OOS Protocol for Mitigating Overfitting in Quantitative Strategies 2026-03-10T05:40:23Z Transitioning a strategy from backtest to live trading is a common failure point for quantitative systems due to parameter overfitting, selection bias, and sensitivity to regime changes. This paper presents the AlgoXpert Alpha Research Framework, a standardized protocol that evaluates strategies across three stages: In Sample (IS), which focuses on stable parameter regions instead of single optima; Walk Forward Analysis (WFA) using rolling windows and purge gaps to reduce information leakage, supported by majority pass and catastrophic veto rules; and Out of Sample (OOS) testing under strict parameter lock with no further tuning. The framework applies a defense in depth structure that includes structural safeguards such as cliff veto, execution controls such as spread and leverage guards, and equity protection mechanisms such as circuit breakers and a kill switch. A case study on USDJPY M5 intraday data demonstrates how to detect overfitting through performance decay and drawdown behavior across chronological stages. A post validation comparison of four alpha variants (v1 to v4) shows rank reversal when the objective changes from maximizing Sharpe to minimizing maximum drawdown, highlighting the trade off between risk adjusted performance and tail risk control. 2026-03-10T05:40:23Z Alpha Research Framework; Walk-Forward Analysis; Purged Validation; Pa rameter Stability; Backtest Overfitting; Selection Bias; Execution-Aware Backtesting; Stress Testing; Kill Switch; Out-of-Sample Verification. 19 Pages, 2 figures The Anh Pham Bao Chan Nguyen Nguyet Nguyen Thi http://arxiv.org/abs/2603.08553v1 Generative Adversarial Regression (GAR): Learning Conditional Risk Scenarios 2026-03-09T16:16:59Z We propose Generative Adversarial Regression (GAR), a framework for learning conditional risk scenarios through generators aligned with downstream risk objectives. GAR builds on a regression characterization of conditional risk for elicitable functionals, including quantiles, expectiles, and jointly elicitable pairs. We extend this principle from point prediction to generative modeling by training generators whose policy-induced risk matches that of real data under the same context. To ensure robustness across all policies, GAR adopts a minimax formulation in which an adversarial policy identifies worst-case discrepancies in risk evaluation while the generator adapts to eliminate them. This structure preserves alignment with the risk functional across a broad class of policies rather than a fixed, pre-specified set. We illustrate GAR through a tail-risk instantiation based on jointly elicitable $(\mathrm{VaR}, \mathrm{ES})$ objectives. Experiments on S\&P 500 data show that GAR produces scenarios that better preserve downstream risk than unconditional, econometric, and direct predictive baselines while remaining stable under adversarially selected policies. 2026-03-09T16:16:59Z Saeed Asadi Jonathan Yu-Meng Li http://arxiv.org/abs/2603.08552v1 Nonconcave Portfolio Choice under Smooth Ambiguity 2026-03-09T16:16:25Z We study continuous-time portfolio choice with nonlinear payoffs under smooth ambiguity and Bayesian learning. We develop a general framework for dynamic, non-concave asset allocation that accommodates nonlinear payoffs, broad utility classes, and flexible ambiguity attitudes. Dynamic consistency is obtained by a robust representation that recasts the ambiguity-averse problem as ambiguity-neutral with distorted priors. This structure delivers explicit trading rules by combining nonlinear filtering with the martingale approach and nests standard concave and linear-payoff benchmarks. As a leading application, delegated management with convex incentives illustrates that ambiguity aversion shifts beliefs toward adverse states, limits the range of states that would otherwise trigger more aggressive risk taking, and reduces volatility through lower risky exposure. 2026-03-09T16:16:25Z 36 pages, 8 figures Emanuele Borgonovo An Chen Massimo Marinacci Shihao Zhu http://arxiv.org/abs/2603.19288v1 Joint Return and Risk Modeling with Deep Neural Networks for Portfolio Construction 2026-03-09T01:49:51Z Portfolio construction traditionally relies on separately estimating expected returns and covariance matrices using historical statistics, often leading to suboptimal allocation under time-varying market conditions. This paper proposes a joint return and risk modeling framework based on deep neural networks that enables end-to-end learning of dynamic expected returns and risk structures from sequential financial data. Using daily data from ten large-cap US equities spanning 2010 to 2024, the proposed model is evaluated across return prediction, risk estimation, and portfolio-level performance. Out-of-sample results during 2020 to 2024 show that the deep forecasting model achieves competitive predictive accuracy (RMSE = 0.0264) with economically meaningful directional accuracy (51.9%). More importantly, the learned representation effectively captures volatility clustering and regime shifts. When integrated into portfolio optimization, the proposed Neural Portfolio strategy achieves an annual return of 36.4% and a Sharpe ratio of 0.91, outperforming equal weight and historical mean-variance benchmarks in terms of risk-adjusted performance. These findings demonstrate that jointly modeling return and covariance dynamics can provide consistent improvements over traditional allocation approaches. The framework offers a scalable and practical alternative for data-driven portfolio construction under nonstationary market conditions. 2026-03-09T01:49:51Z Keonvin Park http://arxiv.org/abs/2603.07881v1 A Distributed Method for Cooperative Transaction Cost Mitigation 2026-03-09T01:35:17Z Funds at large portfolio management firms may consist of many portfolio managers (PMs), each managing a portion of the fund and optimizing a distinct objective. Although the PMs determine their trades independently, the trade lists may be netted and executed by the firm. These net trades may be sufficiently large to impact the market prices, so the PMs may realize prices on their trades that are different from the observed midpoint price of the assets before execution. These transaction costs generally reduce the returns of a portfolio over time. We propose a simple protocol, based on methods from distributed convex optimization, by which a firm can communicate estimated transaction costs to its PMs, and the PMs can potentially revise their trades to realize reduced transaction costs. This protocol does not require the PMs to disclose their method of determining trades to the firm or to each other, nor does it require the PMs to communicate their trade lists with each other. As the number of adjustment rounds grows, the trades converge to the ones that are optimal for the firm. As a practical matter we observe that even just a few rounds of adjustment lead to substantial savings for the firm and the PMs. 2026-03-09T01:35:17Z 28 pages, 4 figures Nikhil Devanathan Logan Bell Dylan Rueter Stephen Boyd http://arxiv.org/abs/2603.07692v1 Understanding the Long-Only Minimum Variance Portfolio 2026-03-08T15:47:09Z For a covariance matrix coming from a factor model of returns, we investigate the relationship between the long-only global minimum variance portfolio and the asset exposures to the factors. In the case of a 1-factor model, we provide a rigorous and explicit description of the long-only solution in terms of the parameters of the covariance matrix. For $q>1$ factors, we provide a description of the long-only portfolio in geometric terms. The results are illustrated with empirical daily returns of US stocks. 2026-03-08T15:47:09Z 25 pages, 6 figures Nick L. Gunther Alec N. Kercheval Ololade Sowunmi http://arxiv.org/abs/2603.02455v2 The Gibbs Posterior and Parametric Portfolio Choice 2026-03-06T22:36:35Z Parametric portfolio policies may experience estimation risk. I develop a generalized Bayesian framework that updates priors, delivering a posterior distribution over characteristic tilts and out-of-sample returns that is the unique belief-updating rule consistent with the investor's utility function, requiring no model for the return generating process. The Gibbs posterior is the closest distribution to the prior in Kullback-Leibler divergence subject to utility maximization. The posterior's scaling parameter $λ$ controls the weight placed on data relative to the prior. I develop a KNEEDLE algorithm to select optimal $λ^*$ in-sample by trading off posterior precision against numerical fragility, eliminating the need for out-of-sample validation. I apply this to U.S. equities (1955-2024), and confirm characteristic-based gains concentrate pre-2000. I find that $λ^*$ varies meaningfully with risk aversion and depends on higher-order moments. 2026-03-02T22:54:55Z Christopher G. Lamoureux http://arxiv.org/abs/2603.15652v1 P vs NP Problem in Portfolio Optimization: Integrating the Markowitz-CAPM Framework with Cardinality Constraints and Black-Scholes Derivative Pricing 2026-03-06T09:53:33Z This paper makes the Millennium Prize problem P vs NP operational in quantitative finance by studying cardinality-constrained portfolio selection. Starting from the convex Markowitz mean-variance program with CAPM-based expected returns (Rf plus beta times ERP), we impose a hard sparsity rule that limits the portfolio to K assets out of approximately 94 industry portfolios (Damodaran). The constraint couples discrete subset selection with continuous weight optimization, yielding a mixed-integer quadratic program and an NP-hard search space that grows combinatorially with n and K. We therefore evaluate scalable approximation schemes (greedy screening, Monte Carlo sampling, and genetic algorithms) under a replication-oriented protocol with random-seed control, distributional performance summaries (median and quantiles), runtime profiling, and convergence diagnostics. Dependence structure is documented via correlation and covariance diagnostics and positive-semidefinite checks to link algorithm behavior to the geometry implied by the risk matrix. To support the title's derivatives component, we add a European call option priced by the Black-Scholes model and map it into CAPM-consistent moments using delta-based linearization, validated with a bump test and moneyness/maturity sensitivity. Results highlight how the cardinality constraint reshapes the attainable efficient frontier, why stability and computational-cost trade-offs matter more than single-best runs, and how common-factor dependence can limit diversification in K-sparse solutions. The study provides a reproducible template for NP-hard portfolio optimization with transparent inputs and extensible derivative overlays. 2026-03-06T09:53:33Z Working paper (preprint). Uses ~94 Damodaran industry portfolios to study cardinality-constrained Markowitz-CAPM portfolio optimization (MIQP/NP-hard) with Monte Carlo and genetic algorithm approximations. Includes correlation/covariance diagnostics, efficient frontier and Sharpe summaries, runtime/seed reproducibility, and a Black-Scholes option overlay with a delta bump-test check Davit Gondauri http://arxiv.org/abs/2503.08272v3 Dynamically optimal portfolios for monotone mean--variance preferences 2026-03-06T08:56:01Z Monotone mean-variance (MMV) utility is the minimal modification of the classical Markowitz utility that respects rational ordering of investment opportunities. This paper provides, for the first time, a complete characterization of optimal dynamic portfolio choice for the MMV utility in asset price models with independent returns. The task is performed under minimal assumptions, weaker than the existence of an equivalent martingale measure and with no restrictions on the moments of asset returns. We interpret the maximal MMV utility in terms of the monotone Sharpe ratio (MSR) and show that the global squared MSR arises as the nominal yield from continuously compounding at the rate equal to the maximal local squared MSR. The paper gives simple necessary and sufficient conditions for mean-variance (MV) efficient portfolios to be MMV efficient. Several illustrative examples contrasting the MV and MMV criteria are provided. 2025-03-11T10:40:48Z 39 pages, 1 figure Mathematics of Operations Research, 2026 Aleš Černý Johannes Ruf Martin Schweizer 10.1287/moor.2025.1136 http://arxiv.org/abs/2312.13057v3 Cross-Currency Heath-Jarrow-Morton Framework in the Multiple-Curve Setting 2026-03-04T22:15:17Z We provide a general HJM framework for forward contracts written on abstract market indices with arbitrary fixing and payment adjustments, and featuring collateralization in any currency denominations. In view of this, we first provide a thorough study of cross-currency markets in the presence of collateral and incompleteness. Then we give a general treatment of collateral dislocations by describing the instantaneous cross-currency basis spreads by means of HJM models, for which we derive appropriate drift conditions. The framework obtained allows us to simultaneously cover forward-looking risky IBOR rates, such as EURIBOR, and backward-looking rates based on overnight rates, such as SOFR. Due to the discrepancies in market conventions of different currency areas created by the benchmark transition, this is pivotal for describing portfolios of interest-rate products that are denominated in multiple currencies. As an example of contract simultaneously depending on all the risk factors that we describe within our framework, we treat cross-currency swaps using our proposed abstract indices. 2023-12-20T14:31:49Z 54 pages, 3 figures Alessandro Gnoatto Silvia Lavagnini http://arxiv.org/abs/2603.16904v1 Quantum-Assisted Optimal Rebalancing with Uncorrelated Asset Selection for Algorithmic Trading Walk-Forward QUBO Scheduling via QAOA 2026-03-04T15:15:55Z We present a hybrid classical-quantum framework for portfolio construction and rebalancing. Asset selection is performed using Ledoit-Wolf shrinkage covariance estimation combined with hierarchical correlation clustering to extract n = 10 decorrelated stocks from the S&P 500 universe without survivorship bias. Portfolio weights are optimised via an entropy-regularised Genetic Algorithm (GA) accelerated on GPU, alongside closed-form minimum-variance and equal-weight benchmarks. Our primary contribution is the formulation of the portfolio rebalancing schedule as a Quadratic Unconstrained Binary Optimisation (QUBO) problem. The resulting combinatorial optimisation task is solved using the Quantum Approximate Optimisation Algorithm (QAOA) within a walk-forward framework designed to eliminate lookahead bias. This approach recasts dynamic rebalancing as a structured binary scheduling problem amenable to variational quantum methods. Backtests on S&P 500 data (training: 2010-2024; out-of-sample test: 2025, n = 249 trading days) show that the GA + QAOA strategy attains a Sharpe ratio of 0.588 and total return of 10.1%, modestly outperforming the strongest classical baseline (GA with 10-day periodic rebalancing, Sharpe 0.575) while executing 8 rebalances versus 24, corresponding to a 44.5% reduction in transaction costs. Multi-restart QAOA (4096 measurement shots per run) exhibits concentrated probability mass on high-quality schedules, indicating stable convergence of the variational procedure. These findings suggest that hybrid classical-quantum architectures can reduce turnover in portfolio rebalancing while preserving competitive risk-adjusted performance, providing a structured testbed for near-term quantum optimisation in financial applications. 2026-03-04T15:15:55Z Abraham Itzhak Weinberg http://arxiv.org/abs/2509.01393v2 Adaptive Alpha Weighting with PPO: Enhancing Prompt-Based LLM-Generated Alphas in Quant Trading 2026-03-04T02:58:52Z This paper introduces a reinforcement learning framework that employs Proximal Policy Optimization (PPO) to dynamically optimize the weights of multiple large language model (LLM)-generated formulaic alphas for stock trading strategies. Formulaic alphas are mathematically defined trading signals derived from price, volume, sentiment, and other data. Although recent studies have shown that LLMs can generate diverse and effective alphas, a critical challenge lies in how to adaptively integrate them under varying market conditions. To address this gap, we leverage a DeepSeek model to generate fifty alphas for ten stocks, and then use PPO to adjust their weights in real time. Experimental results indicate that the PPO-optimized strategy does not consistently deliver the highest cumulative returns across all stocks, but it achieves comparatively higher Sharpe ratios and smaller maximum drawdowns in most cases. When compared with baseline strategies, including equal-weighted, buy-and-hold, random entry/exit, and momentum approaches, PPO demonstrates more stable risk-adjusted performance. The findings highlight the importance of reinforcement learning in the allocation of alpha weights and show the potential of combining LLM-generated signals with adaptive optimization for robust financial forecasting and trading. 2025-09-01T11:39:54Z This paper has been accepted by International Journal of Data Science and Analytics Qizhao Chen Hiroaki Kawashima http://arxiv.org/abs/2603.03213v1 Dynamic Tracking Error and the Total Portfolio Approach 2026-03-03T18:06:56Z The Total Portfolio Approach and Strategic Asset Allocation are widely viewed as competing frameworks for institutional portfolio management. We argue they differ in a single governance parameter: the tracking error constraint. Using U.S. equity and bond data from 2000 to 2026, with portfolio simulations spanning 2004 to 2026, we show that Sharpe ratios are statistically indistinguishable across the full constraint spectrum while the volatility of realized tracking error varies approximately 12-fold. The cost of constraints spikes during crises, when forward returns are richest and governance pressure to de-risk is strongest. Dynamic tracking error subsumes both approaches and provides boards with a more productive framework for investment governance. 2026-03-03T18:06:56Z 56 pages, 7 exhibits Ashwin Alankar Allan Maymin Philip Maymin Myron Scholes Sujiang Zhang http://arxiv.org/abs/2603.00738v1 Exploratory Randomization for Discrete-Time Risk-Sensitive Benchmarked Investment Management with Reinforcement Learning 2026-02-28T17:05:39Z This paper bridges reinforcement learning (RL) and risk-sensitive stochastic control by introducing a tractable exploration mechanism for policy search in risk-sensitive portfolio management, with known and unknown model parameters, that yields an endogenous relative-entropy regularization. We construct a discrete-time risk-sensitive benchmarked investment model. This model combines a factor-based asset universe with periodic portfolio rebalancing. Exploration is incorporated through user-specified Gaussian perturbations to baseline (exploitative) controls. The risk-sensitive stochastic control problem is solved analytically using the Free Energy-Entropy Duality. The Duality recasts the control problem as a linear-quadratic-Gaussian game and introduces a natural penalty for exploration. This approach yields simple sufficiency conditions for optimality. It also induces intuitive bounds on exploration based on risk sensitivity, asset covariance, and rebalancing frequency. Additionally, the optimal investment strategy can be interpreted through the lens of fractional Kelly strategies. By connecting risk-sensitive control theory and RL, this work provides a principled parametric family for policy-gradient implementations, guiding the design of RL methods. 2026-02-28T17:05:39Z 36 pages Sebastien Lleo Wolfgang Runggaldier http://arxiv.org/abs/2508.09429v2 Optimal Control of Reserve Asset Portfolios for Stablecoins 2026-02-28T05:14:38Z Stablecoins promise par convertibility, yet issuers must balance immediate liquidity against yield on reserves to keep the peg credible. We study this treasury problem as a continuous-time control task with two instruments: reallocating reserves between cash and short-duration government bills, and setting a spread fee for either minting or burning the coin. Mint and redemption flows follow mutually exciting processes that reproduce clustered order flow. Peg deviations arise when immediate cash coverage is insufficient relative to outstanding supply, and the market price relaxes toward this liquidity-coverage fair value. We develop a stochastic model predictive control framework that incorporates moment closure for event intensities. Using Pontryagin's Maximum Principle, we show that the optimal reallocation control exhibits a soft-thresholding structure: no rebalancing occurs when the shadow-cost differential lies within a deadzone set by transaction costs, and reallocation scales linearly beyond that threshold up to a capacity-imposed saturation limit. Introducing settlement windows leads to a sampled-data implementation with a simple threshold (soft-thresholding) structure for rebalancing. We also establish a monotone stress-response property: as expected outflows intensify or windows lengthen, the optimal policy shifts predictably toward cash. In simulations covering various stress test scenarios, the controller preserves most bill carry in calm markets, builds cash quickly when stress emerges, and avoids unnecessary rotations under transitory signals. The proposed policy is implementation-ready and aligns naturally with operational cut-offs. Our results translate empirical flow risk into auditable treasury rules that improve peg quality without sacrificing avoidable carry. 2025-08-13T02:07:35Z Alexander Hammerl