https://arxiv.org/api/xcytEDFAJHwv8Y6ReWsrlcgC9Vo2026-07-22T19:06:09Z24147515http://arxiv.org/abs/2212.07944v4Variable Clustering via Distributionally Robust Nodewise Regression2026-06-01T03:22:24ZWe study a multi-factor block model for variable clustering and connect it to regularized subspace clustering through a distributionally robust version of nodewise regression. To solve the latter problem, we derive a convex relaxation, provide a data-driven approach for selecting the size of the robust region, and develop an ADMM algorithm for efficient implementation. We validate our method in extensive numerical studies and demonstrate its superior performance.2022-12-15T16:23:25ZICML 2026Kaizheng WangXiao XuXun Yu Zhouhttp://arxiv.org/abs/2606.13697v1On Reference-Regulated Multiperiod Mean-Variance Portfolio Optimization in High Dimensions2026-05-31T14:05:52ZThe multiperiod mean-variance (MV) portfolio optimization serves as a vital expansion of Markowitz's static MV portfolio selection framework. Just like its static counterpart, the multiperiod MV portfolio remains susceptible to estimation errors. We propose a reference-regulated multiperiod mean-variance (RRMV) framework that penalizes deviations from a reference policy. Therefore, this new optimization successfully combines the advantages of dynamic strategies and reference portfolios. A key contribution of this paper is the characterization of the out-of-sample Sharpe ratio under high-dimensional asymptotics with estimation errors in both the mean vector and the covariance matrix. We show how the reference penalty and the investment horizon jointly affect the optimized portfolio performance, and how regularization operates differently from the single-period portfolio optimization. Extensive simulation and real data studies demonstrate that the proposed framework improves the stability and out-of-sample Sharpe ratios of multiperiod policies significantly.2026-05-31T14:05:52ZYutao DengJianjun GaoWeichen Wanghttp://arxiv.org/abs/2605.17628v2A Penalty-Free Pipeline for Direct Quantum-Annealer Portfolio Optimization2026-05-31T08:51:45ZCardinality-constrained portfolio selection is routinely cast as a quadratic unconstrained binary optimization (QUBO) and submitted to a quantum processing unit (QPU) for direct annealing. We show that this standard penalty encoding is the binding constraint for direct-QPU execution on current D-Wave Pegasus and Zephyr hardware. Expanding the exact cardinality penalty contributes a dense rank-one term that makes the logical interaction graph complete regardless of the covariance, producing chain-break fractions from 83% at small universes up to 92% at the full forty-nine-industry Fama--French universe, and zero feasible raw samples at every tested scale. Topology-aware sparsification reduces chain breaks to near zero, but any sparsifier that removes off-diagonal entries also dilutes the cardinality constraint; an ablation reveals that this sparsify-and-project pipeline is dominated by the classical projector, not the QPU. We propose removing the penalty entirely: sample an objective-only QUBO built from expected returns and the risk-scaled covariance on hardware, and enforce cardinality classically through a deterministic feasibility projector. Across 4,468 saved embedding records on live Pegasus and Zephyr hardware, spanning equities up to forty-nine assets and football-betting instances up to forty-eight, this penalty-free pipeline reduces mean chain-break fractions from 71%--92% down to at most 0.04%, and post-processed regret is at most 0.03% relative to greedy classical references at every tested scale. We do not claim quantum advantage; the penalty encoding, not the sparse hardware topology, is the limiting factor for direct-QPU portfolio optimization at currently accessible scales.2026-05-17T19:50:04ZLuis Lozanohttp://arxiv.org/abs/2307.03391v4On Unified Adaptive Black-Litterman Mean-Variance Portfolio Management2026-05-30T16:42:05ZThis paper proposes a unified adaptive portfolio-management framework that combines factor-based view generation, Black-Litterman (BL) posterior estimation, EWMA covariance estimation, and mean-variance optimization. The key mechanism is a dynamic sliding window that adjusts the estimation horizon according to realized portfolio volatility, thereby updating factor estimates, BL posterior expected returns, and portfolio weights over time. In a ten-year empirical study of the top 100 market-capitalization constituents of the S&P 500 with turnover transaction costs, the proposed method outperforms dynamic mean-variance optimization without BL views and provides stronger downside risk control, while its relative performance remains benchmark-dependent.2023-07-07T05:30:05Z6 pages, 1 figure, 2 tablesChi-Lin LiChung-Han Hsiehhttp://arxiv.org/abs/2605.20636v2Continuous Timing Signals for Growth-Defensive Style Allocation: Factor Attribution, Risk Matching, and Out-of-Sample Evidence2026-05-29T03:38:46ZThis paper studies conditional allocation between a growth/technology ETF basket, denoted by $G$, and a defensive income/value-oriented ETF basket, denoted by $D$. The objective is not to discover a new standalone alpha factor, but to examine whether known style exposures can be dynamically allocated using macro-market timing signals. Fama-French five-factor plus momentum attribution shows that the relative portfolio $G-D$ is a recognizable style portfolio: its market beta is 0.273, its HML beta is -0.552, its momentum beta is 0.117, and its annualized alpha is 1.95\% with a Newey-West t-statistic of only 0.81. The empirical object is therefore interpreted as a growth-versus-defensive style allocation problem rather than a new return anomaly.
The allocation framework replaces discrete regime labels and if-then trading rules with a continuous smooth score. The score combines rate relief, SPY drawdown depth, high-VIX stress relief, and a growth-crowding penalty. Interaction terms are smoothed with softplus functions, the total score is mapped to G/D weights through a hyperbolic tangent function, and realized weights are smoothed with EWMA. In the main aligned comparison window from June 28, 2017 to May 15, 2026, with 10bp transaction costs, the selected smooth-score policy uses a 50\% maximum active tilt and obtains a 19.24\% CAGR, a Sharpe ratio of 1.01, and a maximum drawdown of -31.63\%. It improves over 50/50 G/D, matched TNX-only, matched core-only, SPY, and volatility-matched 100\% G benchmarks. It does not, however, exceed 100\% G or the best high-G static portfolios in raw CAGR. Walk-forward and post-2022 validations provide additional evidence of drawdown reduction and risk-adjusted allocation value. Overall, the evidence supports continuous, interpretable style timing, while also showing that high static growth exposure remains a strong benchmark.2026-05-20T02:45:07Z19 pages, 10 figures, 18 tablesZheli Xionghttp://arxiv.org/abs/2606.00143v1Regime-Adaptive Continual Learning for Portfolio Management2026-05-29T02:24:45ZFinancial markets are inherently non-stationary, exhibiting frequent regime shifts and structural changes that render traditional Portfolio Management (PM) approaches ineffective. Existing remedies, such as rolling-window retraining and naive online fine-tuning, are hindered by high computational costs and insufficient knowledge utilization, respectively, resulting in low returns and limited adaptability. Continual learning (CL) offers a promising paradigm by enabling trading agents to accumulate and transfer knowledge across sequential tasks. In this paper, we propose \textbf{Re}gime-aware \textbf{C}ontinual \textbf{A}daptive \textbf{P}ortfolio management (\textbf{ReCAP}), a novel framework that integrates CL into PM to address the challenges of dynamic financial environments. ReCAP employs an adaptive regime detection module to segment historical market data into variable-length regimes, enabling regime-specific learning of policy vectors and the construction of a policy library. During continual trading, a regime-gate module adaptively combines policy vectors from the library based on the current market state, facilitating rapid adaptation to newly detected regimes. Only the regime-gate and the current regime's policy vector are continually updated to preserve useful knowledge effectively. Extensive experiments on five real-world datasets demonstrate that ReCAP consistently outperforms popular baselines, achieving superior returns in long-term investment horizons and rapid adaptation to regime shifts.2026-05-29T02:24:45ZAccepted by KDD 2026Chaofan PanLingfei RenLinbo XiongYonghao LiWei WeiXin Yanghttp://arxiv.org/abs/2605.30464v1Distributional Portfolio Optimization (DPO): A Unified Framework for Distributions over Weights, Returns, and Parameters2026-05-28T18:38:56ZClassical portfolio optimization treats expected returns, covariances, and allocations as deterministic. Modern practice replaces at least one by a distribution: a posterior over parameters, a law of future returns, a stochastic allocation policy, or a distributional-robustness set. We call distributional portfolio optimization (DPO) the unified framework in which weights, returns, and parameters are all modeled as probability measures, organized around the joint coupling Gamma_theta(dw,dr) and its marginal triple (W,R,P). The contribution is synthetic and structural: we organize Bayesian, robust, chance-constrained, stochastic-allocation, and distributional reinforcement-learning portfolio methods through this coupling and prove boundary results connecting them, including a portfolio specialization of Wasserstein-CVaR duality, a static no-randomization theorem, a Bayesian credible-radius calibration of Wasserstein DRO, a Gaussian-isotropic second-order conservatism bound, a conditional two-sided rate W_1 = Theta(n^{-(1+alpha)/2}) governed by the local boundary Holder exponent alpha in [0,1], and a risk-shifted distributional Bellman contraction. A controlled experiment shows that across factor models at K in {10,25,50}, the credible-radius rule lands within 3-7 bp of the oracle out-of-sample tail risk and beats a 24-month validation-tuned radius while spending no validation data. On a K=25 DJIA backtest, equal-weight, no-view Black-Litterman, and Ledoit-Wolf shrinkage attain higher Sharpe than every distributional method; the operational claim is therefore confined to calibration-without-validation and turnover, not raw-return dominance.2026-05-28T18:38:56ZMiquel Noguer i Alonsohttp://arxiv.org/abs/2605.29413v1From Classical Optimization to Bayesian Integration: A Comprehensive Analysis of Systematic Portfolio Management2026-05-28T06:02:21ZThis paper compares a series of contemporary portfolio construction approaches by employing ten U.S. stocks (TSLA, WMT, BAC, GS, LLY, MRK, GOOG, META, AAPL and XOM) in a time frame from September 2023 to December 2025. The paper explores both basic mean-variance optimization, constrained optimization, Fama French five factor regression modeling, Monte Carlo simulation, and the Black-Litterman model to determine how constraints to a solution, risk factors to a strategy, simulated approximations, and specific market views may all impact the outcome of portfolio allocation, performance and stability. Overall, the results show that standard optimization may result in highly concentrated portfolios, while constrained optimization leads to changes in portfolio allocations by altering the efficient frontier, five factor regression models suggest that a basic investment style of defensive large value and profitability exposure, Monte Carlo approximation is a viable technique to arrive at mean-variance optimal portfolios provided the simulations are high enough especially under a box constraint, the Black Litterman portfolio approach produces more economically intuitive allocations and greater stability compared to standard mean-variance optimization as the approach balances equilibrium returns with investor views.2026-05-28T06:02:21ZAjay Kumar VermaShravya Barkamhttp://arxiv.org/abs/2603.09301v2Constructing a Portfolio Optimization Benchmark Framework for Evaluating Large Language Models2026-05-27T11:20:48ZThis study introduces a benchmark framework for evaluating the financial decision-making capabilities of large language models (LLMs) through portfolio optimization problems with mathematically explicit solutions. Unlike existing financial benchmarks that emphasize language-processing tasks, the proposed framework directly tests optimization-based reasoning in investment contexts. A large set of multiple-choice questions is generated by varying objectives, candidate assets, and investment constraints, with each problem designed to include a unique correct solution and systematically constructed alternatives. Experimental results comparing GPT-4, Gemini 1.5 Pro, and Llama 3.1-70B reveal distinct performance patterns: GPT achieves the highest accuracy in risk-based objectives and remains stable under constraints, Gemini performs well in return-based tasks but struggles under other conditions, and Llama records the lowest overall performance. These findings highlight both the potential and current limitations of LLMs in applying quantitative reasoning to finance, while providing a scalable foundation for developing LLM-based services in portfolio management.2026-03-10T07:35:31ZPoster presented at the AI for Finance Symposium '25, The 6th ACM International Conference on AI in Finance (ICAIF '25)Hanyong ChoJang Ho Kimhttp://arxiv.org/abs/2603.09303v2Investor risk profiles of large language models2026-05-27T11:19:42ZThis paper investigates how large language models (LLMs) form and express investor risk profiles, a critical component of retail investment advising. We examine three LLMs (GPT, Gemini, and Llama) and assess their responses to a standardized risk questionnaire under varying prompts. In particular, we establish each model's default investment profile by analyzing repeated responses per model. We observe that LLMs are generally longterm investors but exhibit different tendencies in risk tolerance: Gemini has a moderate risk level with highly consistent responses, Llama skews more conservative, and GPT appears moderately aggressive with the greatest variation in answers. Moreover, we find that assigning specific personas such as age, wealth, and investment experience leads each LLM to adjust its risk profile, although the extent of these adjustments differs across the models.2026-03-10T07:38:26ZPoster presented at the AI for Finance Symposium '25, The 6th ACM International Conference on AI in Finance (ICAIF '25)Hanyong ChoGeumil BaeJang Ho Kimhttp://arxiv.org/abs/2605.27977v1Deep Learning Forecasting of the U.S. Aggregate Bond Index2026-05-27T05:15:28ZThis study looks at the statistical properties and predictability using deep learning methods of the U.S. aggregate bond index in daily observations spanning 2018 to February 2026. We first establish that index levels are extremely persistent and consistent with unitroot behavior (Dickey and Fuller), while log returns are covariance-stationary with weak linear dependence and pronounced volatility clustering characteristic of ARCH-type processes (Engle; Bollerslev). Motivated by the trade-off between stationarity and information retention, we construct a "stationary but maximally persistent" representation via fractional differencing (Granger and Joyeux; Hosking) following the procedure of López de Prado, and evaluate shorthorizon forecast using two neural paradigms: (i) Multilayer Perceptrons (MLPs) trained on lagged vectors with joint lag-length and hyperparameter tuning (Hornik et al.; Rumelhart et al.); and (ii) Convolutional Neural Networks (CNNs) trained on Gramian Angular Field (GAF) image encodings (Wang and Oates). Empirically, MLPs match the strong naive persistence benchmark on levels, collapse toward near-zero forecasts on returns, and achieve the strongest incremental performance on the fractionally differenced series, where moderate dependence remains but unit-root drift is attenuated. In contrast, CNN-GAF models deliver consistently negative out-of-sample R 2 across all three representations. Overall, the results imply that, for short-horizon forecasting of broad bond indices, the primary determinant of predictive performance is the transformation of the series-its degree of stationarity and memory-rather than architectural complexity. Lag-based models remain competitive under persistence, while GAFbased CNNs are better suited to pattern-based tasks than to persistence-dominated next-step prediction.2026-05-27T05:15:28ZAjay Kumar VermaJul Jon Ramirez GeneralYvan Landry Ndzonde Fonkouhttp://arxiv.org/abs/2605.27945v1Stochastic Volatility, Jumps, and Rates: A Unified Framework for Option Pricing and Term-Structure Simulation2026-05-27T04:36:23ZThis study develops an integrated stochastic modeling framework for pricing short and medium-maturity equity options and assessing interest-rate risk using the Heston (1993), Bates (1996), and CIR (1985) models. We calibrate the Heston model using both the Lewis (2001) Fourier inversion and the Carr-Madan (1999) FFT approach, finding near-identical parameter sets, which is consistent with the calibration stability reported in recent studies such as Agazzotti et al. (2025). Extending the model to Bates shows that jump intensities converge to values effectively equal to zero for 60-day maturities, echoing empirical findings that jumps contribute marginally to short-term smile fitting. We further compare our calibration approach with the joint volatility-surface and variance-term-structure framework proposed by Yoo (2025), confirming that standard Heston/Bates calibration remains robust for the maturities considered. Finally, we calibrate the CIR short-rate model to the Euribor term structure, generating positive and economically consistent forward-rate scenarios in line with recent stochastic-rate option-pricing research by Jeon and Kim (2025). Overall, our results show that continuous stochastic volatility dominates near-term pricing dynamics, while stochastic interest rates materially influence valuations beyond one year.2026-05-27T04:36:23ZNunik Srikandi PutriAjay Kumar VermaNeo Paul Lesupihttp://arxiv.org/abs/2605.27848v1Regime-Based Portfolio Allocation Using Hidden Markov Models and Reinforcement Learning2026-05-27T02:04:31ZThis study develops a regime-aware portfolio allocation framework that integrates Markov switching models with Reinforcement Learning (RL) to dynamically allocate across equities (SPY), long-term Treasuries (TLT), and gold (GLD). Using daily ETF data from 2004-2025, we first characterize market behavior through a discrete Markov chain and then estimate a three-state Gaussian Hidden Markov Model (HMM) selected by the Bayesian Information Criterion (BIC). The estimated regimes-low-volatility, transitional, and high-volatility-exhibit strong persistence and state-dependent return dynamics consistent with recent findings on nonlinear market states (Ardia et al., 2024; Gupta & Pierdzioch, 2023). State-conditional analysis shows that SPY dominates in stable regimes, while TLT and GLD provide protection during stressed periods, motivating regime-conditioned allocation rules.
We evaluate rule-based rotation and RL-driven strategies using a 30% out-of-sample test window with a one-day execution lag to avoid look-ahead bias. Both HMM-based allocations outperform a passive SPY benchmark, while the RL policy achieves the highest risk-adjusted performance, delivering the strongest Sharpe ratio and materially lower drawdowns, yet remains fully interpretable through discrete regime-dependent actions. Sensitivity analysis confirms the robustness of the three-state specification relative to two-state alternatives. Overall, the results demonstrate that RL can systematically enhance HMM-based regime detection, providing a transparent, adaptive, and empirically grounded framework for tactical asset allocation. The combined HMM-RL system provides a transparent, rules-based approach to tactical allocation that improves risk-adjusted performance relative to standard benchmark strategies.2026-05-27T02:04:31ZAjay Kumar VermaNunik Srikandi PutriNeo Paul Lesupihttp://arxiv.org/abs/2509.05676v3Carbon-Sensitive Fund Construction and Hedging for Green Unit-Linked Life Insurance2026-05-26T13:48:01ZWe study the problem of hedging unit linked life insurance policies whose benefits depend on an investment fund that incorporates environmental criteria in its selection process. Offering these products poses two key challenges: constructing a green investment fund and developing a hedging strategy for policies written on that fund. We address these two problems separately. First, we design a portfolio selection rule driven by firms' carbon intensity that endogenously selects assets and avoids ad hoc pre-screens based on ESG scores. The effectiveness of our new portfolio selection method is tested using real market data. Second, we consider an insurance company issuing unit linked policies written on this fund. Such contracts are exposed to market, carbon, and mortality risk, which the insurance company seeks to hedge. Due to market incompleteness, we address the hedging problem via a quadratic approach aimed at minimizing the variance of the hedging costs. Finally, we also make a numerical analysis to assess the performance of the hedging strategy. For our simulation study, we use an efficient weak second-order scheme that allows for variance reduction.2025-09-06T10:55:09Z40 pagesKatia ColaneriAlessandra CretarolaEdoardo LombardoDaniele Mancinellihttp://arxiv.org/abs/2605.26740v1A Unified Theory of Ownership Concentration, Overlap, and Dependence2026-05-26T09:13:19ZOwnership concentration is not a scalar. For a normalized investor-stock matrix $A$, it has three irreducible layers: concentration across investors, concentration across stocks, and dependence in the joint assignment of investors to stocks. This paper develops a unified quadratic framework for those layers and shows that the same residual operator that measures static overlap also governs linearized market transmission. Raw micro concentration $M(A) = \sum_{i,j} A_{ij}^2$ admits exact row and column decompositions, support bounds, and fixed-marginal extremal characterizations on the transportation polytope. Benchmark-adjusted dependence $\mathcal{X}(A) = \sum_{i,j} (A_{ij} - p_i s_j)^2 / (p_i s_j)$ admits two exact decompositions: it is a size-weighted average of investor-level deviations from the market portfolio and, symmetrically, of stock-level deviations from the investor base. The paper also proves a multiscale aggregation law: under any partition of investors, total dependence splits exactly into between-group dependence and within-group heterogeneity. Spectrally, $\mathcal{X}(A)$ equals the sum of squared nontrivial singular values of the whitened matrix $D_p^{-1/2} A D_s^{-1/2}$. The residual operator $L$ then yields two dynamic consequences: idiosyncratic fire-sale vulnerability is bounded by the dominant overlap mode $ρ(A)$, while aggregate benchmark-relative alpha variance has worst-case capacity $ρ(A)^2$ and isotropic average-case capacity $\mathcal{X}(A)$. The fixed-marginal geometry also motivates a feasible-range sparsity score that benchmarks observed micro concentration against the sharp minimum and maximum implied by the marginals. The resulting framework separates scale concentration, feasible sparsity, overlap, and linear transmission in a way that is mathematically transparent and empirically usable for work on crowding, fragility, and systemic risk.2026-05-26T09:13:19ZMiquel Noguer i AlonsoIro Tasitsiomi