https://arxiv.org/api/FAfIA2KxYsQIc1cw4Er++SYlvuk 2026-07-21T14:01:47Z 2412 60 15 http://arxiv.org/abs/2606.11318v1 Mean-Variance Optimization in Ambiguous Financial Markets with Learning 2026-06-09T18:01:53Z We consider a continuous time investment problem in a multi-asset Black-Scholes market with the following features: The assets' drifts are not known and constitute a source of model ambiguity. However, there is a prior distribution (knowledge) on the possible drifts. Our investor is ambiguity averse and wants to maximize a mean-variance criterion for the terminal wealth where ambiguity aversion is incorporated in a smooth way. We consider here the criterion introduced in Maccheroni et al. 2013 where the variance is decomposed and each part is weighted differently to account for different levels of market risk and model ambiguity aversion. We use a novel approach to find the optimal dynamic investment strategy within the class of all adapted strategies which allow for learning. We also present a number of numerical results which help to understand how the model parameters affect the optimal investment strategy. In general it turns out that ambiguity averse investors invest less in the risky assets. 2026-06-09T18:01:53Z Nicole Bäuerle Anne MacKay http://arxiv.org/abs/2606.09420v1 Benchmarking Deep Time Series Models for Equity Portfolios 2026-06-08T12:36:12Z Benchmarking forecasting architectures for daily equity portfolios is not just a prediction exercise. It also asks which model remains usable after preferences, costs, and portfolio constraints are imposed. We build a CRSP daily-stock benchmark for 15 deep and statistical time-series architectures over 2018--2024. The protocol combines common-window decile portfolios, stochastic multi-criteria acceptability analysis, a deployment-adjusted acceptability index, and a constrained quadratic portfolio layer with capacity, beta, industry, risk, leverage, and turnover controls. The index starts from the SMAA rank-acceptability distribution and downweights models whose criteria-level wins produce high portfolio regret; its Gibbs form is characterized as an entropic update from the SMAA prior. Empirically, no architecture dominates the raw benchmark: TransEnc-8 has the largest rank-1 acceptability, 0.352, and no model exceeds about 0.36. Rankings vary with preferences, market state, feature universe, and transaction costs. In the promoted five-model constrained-portfolio comparison, TransEnc-8 is selected throughout, while return-oriented raw rankings can favor TS-RIDGE. Broad-universe decile signals can survive costs, but the baseline constrained-QP net Sharpe at 20 bps is negative for every promoted model. The benchmark supports model selection and diagnosis rather than a standalone trading-strategy claim. 2026-06-08T12:36:12Z 51 pages, 28 figures, 43 tables; includes appendices Aoxin Zhang Yuhan Cheng Kwanting Leung http://arxiv.org/abs/2606.09104v1 Addressing Market Regime Changes and Heavy-Tailed Returns in Portfolio Optimization via Bayesian VAR and Elliptical Black-Litterman 2026-06-08T06:58:11Z Deep reinforcement learning (DRL) frameworks for portfolio optimization have shown promise for their ability to learn allocation rules dynamically from market data. However, these models fail to account for fat-tailed returns, which characterize actual market behavior with more frequent extreme events. Furthermore, historical data is treated homogeneously, without accounting for temporal importance, leading models to fail during regime changes. We propose a new BAVAR-BLED algorithm that combines methods derived from Bayesian-Averaging Vector Autoregressive (BAVAR) and the Black-Litterman model using Elliptical Distributions (BLED) within a TD3 architecture. BAVAR captures a set of vector autoregressive representations that consider multi-scale temporal features, enabling adaptive allocation decisions based on regime-aware estimates of return expectations and dispersion matrices. These estimates serve as prior inputs to BLED, a model that uses Student's t-distributions, allowing for more realistic fat tail return estimates. The BAVAR-BLED algorithm uses transformer networks for view construction and CNNs for risk-aversion estimates, which modify dynamic allocation decisions based on market conditions. An evaluation of 29 Dow Jones Industrial Average constituents over a decade-long market period shows that BAVAR-BLED significantly outperforms state-of-the-art methods, achieving Sharpe and Sortino ratios of 1.72 and 2.70, respectively, and total returns of 57.26%. 2026-06-08T06:58:11Z 9 pages, 3 figures, 4 tables. Extends our prior work [Mikriukov et al., ICIC 2025] on Black-Litterman under Elliptical Distributions (BLED). Manuscript under review Daniil Mikriukov University of Liverpool Xi'an Jiaotong-Liverpool University Ruoyu Sun Xi'an Jiaotong-Liverpool University Angelos Stefanidis Xi'an Jiaotong-Liverpool University Jionglong Su Xi'an Jiaotong-Liverpool University Zhengyong Jiang Xi'an Jiaotong-Liverpool University http://arxiv.org/abs/2606.08791v1 Evaluating AI Investment Strategies 2026-06-07T19:16:50Z We study the problem of auditing a black-box algorithmic decision-maker from observable inputs and outputs alone. Our main result is an exact decomposition: under precisely characterized conditions, the cumulative \emph{regret} of a dynamic policy equals the sum of per-period covariances between the cost vector and the policy's decision. This extends the single-period identity of Aldridge~(2026) to the full multi-period setting of stochastic dynamic programming. We prove the identity holds exactly under i.i.d. costs and mean-unbiased Markov policies, derive closed-form bias corrections for non-stationary and time-varying cases, and establish the discounted-horizon analog. A Bellman recursion for the covariance regret functional connects the result to standard reinforcement learning algorithms; for rolling-window policies, the estimation-error bias is $O(d/w)$. The decomposition has direct implications for algorithmic auditing in strategic environments: in platform mechanism design, it provides a welfare-based audit metric without access to the agent's private type; in repeated games, covariance reduction is a sufficient condition for policy improvement; in procurement and ad auctions, the bias correction quantifies welfare loss from strategic misreporting. The associated trajectory estimator is consistent, asymptotically normal with HAC variance, and computable in $O(T \cdot nd)$ time. This makes the proposed approach a tractable, model-free audit tool for platform mechanisms, algorithmic portfolio strategies, and any sequential decision system subject to external performance review. 2026-06-07T19:16:50Z 33 pages Irene Aldridge http://arxiv.org/abs/2606.08569v1 Stock Investment: The p-index Approach 2026-06-07T10:51:47Z This paper has used European put option to construct the p-index risk measure to evaluate the performance of different investment strategies in China's SSE 50 index and the US SP500 index during 2018-2023. The p-index measures the insurance fee for each insured dollar to guarantee that the asset achieves at least a delta rate of return on a specified future date. It is found that with the fair price strategy, one-week and one-month holding periods can earn more, and among seven economic sectors, materials sector stocks generated highest annualized rates of return: 11.04% (one-week period), 11.93% (two-week period) and 10.18% (one-month period). With momentum and contrarian strategies of one-week holding period, the p-ratio-efficient-contrarian strategy produced the highest annualized rate of return (9.97%), followed by the p-index-inefficient-momentum strategy (9.01%) and the p-index-efficient-contrarian strategy (6.48%), the MCIRS method employing the p-index consistently delivered higher returns than its beta-based approach, and efficient (outperforming) stocks failed to sustain their momentum while inefficient (underperforming) stocks exhibited no mean reversion. It is also found that the p-index-efficient-contrarian strategy outperformed in low-sentiment (low-volume) regimes, while the p-index-inefficient-momentum strategy outperformed during high-sentiment (high-volume) periods. For the five hundred stocks of the US S&P 500 index during 2018-2023, it is found that efficient stocks sustained their momentum while inefficient stocks exhibited mean reversion. The p-index-efficient-momentum strategy produced the highest annualized rate of return (3.69%), followed by the p-ratio-inefficient-contrarian strategy (3.67%) and the beta-efficient-momentum strategy (3.48%). 2026-06-07T10:51:47Z arXiv admin note: text overlap with arXiv:2510.11074 Xinzhao Xie Bopei Nie Kuo-Ping Chang http://arxiv.org/abs/2606.08283v1 Macro Economists in the Machine: A Multi-Agent LLM Framework for Commodity-Related ETF Portfolio Construction 2026-06-06T18:07:28Z We test whether large language models (LLMs) add value in commodity portfolio construction when the information set and implementation rules are held fixed across strategies. A Hawkish Agent (inflation-tightening prior), a Dovish Agent (growth-easing prior), a Debate Agent, and a deterministic z-score Rule Agent each receive identical FRED macro z-scores and route their tilt signals through the same portfolio engine. Across 124 weekly rebalancing dates spanning the 2023 U.S. rate peak and the 2024-2025 soft landing, all three LLM strategies outperform the Rule Agent in Sharpe terms; the Hawkish and Debate Agents record the largest gains (ΔSharpe = +0.044 and +0.040, both p < 0.10 under a block bootstrap) and preserve a net-of-cost advantage over the passive inverse-volatility benchmark at one-way trading costs up to 30 basis points, while the Rule Agent's thin margin over passive disappears at approximately 5 basis points.The Debate Agent does not outperform the best single agent (ΔSharpe = -0.004, p = 0.769); its contribution is bias correction -- averaging out the Dovish Agent's miscalibrated prior -- rather than deliberation-generated return. The performance advantage is concentrated in the soft-landing sub-period, the evaluation window spans a single rate cycle, and the reported $p$-values are unadjusted for multiple comparisons. Within these limits, the results suggest that an LLM acting as a constrained macro-interpretation function can add modest but economically meaningful value over a transparent rule layer, though the margin is small and its persistence beyond this sample is unknown. 2026-06-06T18:07:28Z 45 pages, 4 figures Yiqing Wang Dehao Dai Ding Ma Kerui Geng http://arxiv.org/abs/2606.07727v1 Benchmarking Quantum Algorithmic Resilience for CVaR Portfolio Optimization: The Expressibility-Coherence Trade-off 2026-06-05T17:07:59Z Quantum combinatorial optimization offers theoretical advantages for complex financial modeling, but physical implementation on Noisy Intermediate Scale Quantum (NISQ) devices is severely constrained by hardware topology. This study presents a hardware benchmarking analysis between a Hardware Efficient Variational Quantum Neural Network (HE-VQNN) and the Warm Start Quantum Approximate Optimization Algorithm (WS-QAOA) for a hybrid Mean Variance and Conditional Value at Risk (CVaR) portfolio objective. By implementing a novel classical quantum hybrid proxy matrix to bypass the CVaR auxiliary qubit bottleneck, we map up to 16 assets from the NIFTY 50 index onto an IBM heavy hex processor. We systematically quantify algorithmic resilience to the "SWAP tax" incurred during routing. Empirical results reveal a critical operational trade-off: WS-QAOA provides exact theoretical mapping but suffers catastrophic hardware decoherence due to exponential nonlocal gate overhead. Conversely, HE-VQNN preserves hardware coherence but lacks the mathematical expressibility to capture dense tail risk asset correlations. This study exposes the limitations of dense financial optimization on current architectures forces an nonviable choice between algorithmic inexpressibility and hardware decoherence. This is indicative of a deeper limitation as to what can and cannot be done with NISQ computers lacking in all-to-all connectivity. 2026-06-05T17:07:59Z 10 pages, 11 figures. Master's thesis research conducted at the School of Quantum Technology, Defence Institute of Advanced Technology (DIAT), Pune Prashik N. Somkuwar K. Srinivasan G. Raghavan http://arxiv.org/abs/2606.07450v1 Information Networks of Stock Prices 2026-06-05T16:54:11Z The collective movement of stock prices harbors complex interdependencies that are conventionally simplified only through a linear lens. This paper explores computed structural network representations in the Indonesian capital market by testing the limits of Pearson correlation and Mutual Information (MI) in unveiling the spectral dynamics of the market. Across 2,328 rolling observation windows from 2015 to 2025, we examine 24 methodological configurations that combine three dependency estimators (Pearson, MI adaptive binning, and MI-kNN), two graph filtering schemes (Minimum Spanning Tree/MST and Planar Maximally Filtered Graph/PMFG), and four community decoders. The empirical results unveil a fundamental reality: topological richness does not always resonate with sectoral classification precision. The Pearson, MST, and Infomap configuration is shown to remain the most robust foundation for recovering conventional sectoral taxonomy. Nevertheless, when deeper observation demands the exposition of local structures and the weave of heterogeneous communities, the architectural relaxation through PMFG demonstrates its superiority. In the realm of residual information detection, MI adaptive binning appears far more proportional than kNN; histogram-based regularization successfully tames empirical noise without sweeping away traces of non-linear dependency. Ultimately, the synergy of MI and PMFG is not positioned to dethrone the dominance of linear correlation, but rather to provide an essential analytical lens for excavating hidden economic sub-structures -- such as the cohesion of commodity regimes -- that have long transcended the rigid boundaries of the market's formal sectors. 2026-06-05T16:54:11Z 12 pages, 6 figures Muhammad Aldy Hassan Hokky Situngkir http://arxiv.org/abs/2605.27887v2 PortBench: A Correlation-Aware, Full-Pipeline Benchmark for LLM-Driven Portfolio Management 2026-06-04T13:29:42Z Large language models (LLMs) have shown strong performance across diverse financial tasks, yet portfolio management (PM), a critical financial decision-making task, remains poorly benchmarked. Existing benchmarks exhibit two main gaps: they ignore cross-asset correlation structures, thereby failing to distinguish genuinely diversified portfolios from concentrated ones, and fail to evaluate the complete PM decision pipeline in real-world scenarios. We introduce PortBench, a benchmark spanning six heterogeneous asset classes over ten years. PortBench consists of two complementary layers: a static QA dataset of 6,269 correlation-based questions across seven task templates, and a dynamic five-stage allocation pipeline that mirrors the full PM decision cycle. To evaluate these layers, we introduce two dedicated metrics: a dual-layer correlation score that measures whether proposed portfolios exploit inter-class hedging and avoid intra-class concentration, and CEPS, a metric that quantifies how reasoning errors compound across pipeline stages. We further assess strategy robustness and investor alignment under three historical stress regimes and risk profiles. Evaluating ten frontier LLMs, we find that despite strong performance on static financial QA, 90\% of model-profile combinations fail to outperform a basic equal-weight allocation, and models that satisfy every procedural constraint still suffer catastrophic drawdowns under stress. Our source code is available at \href{https://github.com/AgenticFinLab/portbench}{this https URL}. 2026-05-27T03:08:23Z Project page: https://portbench.github.io/ Yuxuan Zhao Sijia Chen Ningxin Su http://arxiv.org/abs/2606.04258v1 Anticipatory Portfolio Optimization 2026-06-02T22:21:19Z A portfolio is \emph{anticipatory} when its optimizer acts on a richer model than the myopic, price-taking estimator used to calibrate it. Enrichment may be informational, via enlarged filtrations; dynamic, via horizon forecasts; or performative, via the deployment law induced by market impact. We give a decision-theoretic definition for all three cases and measure anticipation by the realized control gap between enriched controller and restricted estimator. The same quadratic geometry separates information, planning value, impact correction, and overfitting. For log utility under initial enlargement, value is the information-drift energy $\frac12\mathbb{E} \int_0^Tα_t^2\,dt$, equivalently mutual information or relative entropy. In mean-variance form, signal value is $\frac{1}{2γ}{\rm tr}(Σ^{-1}Ω)$. Dynamic forecast anticipation gives a finite-horizon quadratic premium in the forecast stack, while permanent impact changes the price-taking allocation $θ_{\rm na} =(Λ+γΣ)^{-1}μ$ into $θ_{\rm an} = (2Λ+γΣ)^{-1}μ$ and reveals a spectral phase transition for naive recalibration. The main result is a stacked finite-horizon LQG decomposition: information, forecast, and impact combine into an information trace plus one inverse-precision norm, whose expansion yields the impact term, forecast term, and signed forecast-impact interaction. Sharp angle bounds and an orthogonal nonnegative projection identity resolve the signed term. The stationary extension endogenizes information covariance as Kalman error reduction and carries impact anticipation to an infinite-horizon Lyapunov trace with transaction costs. Finally, the penalty $\frac{1}{2}{\rm tr}(H^{-1}Σ_\varepsilon)$ shows that correctly specified anticipation creates value, vacuous anticipation has zero value, and misspecified anticipation is harmful when estimated structure is optimized as true. 2026-06-02T22:21:19Z Miquel Noguer i Alonso http://arxiv.org/abs/2509.22088v3 Factor-Based Conditional Diffusion Model for Contextual Portfolio Optimization 2026-06-02T05:33:05Z We propose a novel conditional diffusion model for contextual portfolio optimization that learns the cross-sectional distribution of next-day stock returns conditioned on high-dimensional asset-specific factors. Our model leverages a Diffusion Transformer architecture with token-wise conditioning, which enables linking each asset's return to its own factor vector while capturing complex cross-asset dependencies. By drawing generative samples from the learned conditional return distribution, we perform daily mean-variance and mean-CVaR optimization, incorporating transaction costs and realistic constraints. Using data from the Chinese A-share market, we demonstrate that our approach consistently outperforms various standard benchmarks across multiple risk-adjusted performance metrics. Furthermore, we establish a 2-Wasserstein error bound for the conditional diffusion model and quantify how its distributional approximation errors propagate to the downstream portfolio optimization task. Our results demonstrate the potential of generative diffusion models for high-dimensional, risk-sensitive contextual stochastic optimization and financial decision making. 2025-09-26T09:11:08Z Xuefeng Gao Mengying He Xuedong He Jiale Zha http://arxiv.org/abs/2606.03158v1 Portfolio Choice with Competing Precautionary and Accumulation Goals 2026-06-02T05:09:36Z We study optimal portfolio choice for a household simultaneously managing a random-deadline goal, such as a medical emergency or job loss, and a fixed-deadline goal such as retirement or college tuition. Under a forced funding rule, in which each goal is paid in full whenever affordable, the household maximizes a weighted sum of the probabilities of fully funding both goals in a Black--Scholes market. We identify two novel effects absent from single-goal models: a growth crowding-out effect, in which precautionary saving for the random goal distorts investment toward the fixed goal, and a deadline pressure effect, in which a compressed saving horizon forces excess risk-taking. A striking implication is that the value function need not be monotone in wealth: a household just above the random-goal threshold is forced to pay it when the shock arrives, depleting its wealth for the fixed goal, and ends up worse off than a slightly poorer household that missed the random goal but kept its wealth intact. This non-monotonicity is absent from all single-goal benchmarks and arises purely from the interaction between the two goal types under forced funding. We further study an optional funding variant in which the household may decline the fixed-deadline goal at time $T$ rather than being required to fund it. We characterize the ex ante option value, i.e., the full time-$0$ value of this flexibility and the terminal option value, i.e., its value at the funding decision node. We find that both options are most valuable at intermediate wealth levels where paying the fixed-deadline goal would substantially reduce the continuation value of the random-deadline problem. 2026-06-02T05:09:36Z Steven Campbell Agostino Capponi Ananya Parashar http://arxiv.org/abs/2606.02945v1 Infinite Horizon Optimal Consumption: Intertemporal Hedging under Epstein-Zin Preferences 2026-06-01T22:53:18Z We study an infinite-horizon optimal consumption-investment problem for an investor with Epstein-Zin stochastic differential utility with stochastic investment opportunities in an incomplete market. Risk aversion and intertemporal substitution are separated, and we work in the regime $θ\in(0,1)$, where there exists a unique generalised utility process for arbitrary non-negative progressively measurable consumption streams. Our main contribution is a variational characterisation of the value function. We show that the value function is the unique minimiser of a functional whose Euler-Lagrange equation coincides with the Hamilton-Jacobi-Bellman equation. Although the functional may be non-convex, the direct method yields existence, and we prove every minimiser is strictly positive, bounded, and classical. A verification theorem identifies any minimiser with the value function and gives feedback representations for optimal consumption and investment policies. The proof combines a change of measure to the myopic probability with uniqueness results for Epstein-Zin BSDEs and a perturbation argument for optimality. Examples with stochastic volatility, Gaussian excess returns, and fat-tailed excess returns illustrate the scope of the framework and its implications for intertemporal hedging. 2026-06-01T22:53:18Z 27 pages Erhan Bayraktar Emmet Lawless http://arxiv.org/abs/2212.07944v4 Variable Clustering via Distributionally Robust Nodewise Regression 2026-06-01T03:22:24Z We study a multi-factor block model for variable clustering and connect it to regularized subspace clustering through a distributionally robust version of nodewise regression. To solve the latter problem, we derive a convex relaxation, provide a data-driven approach for selecting the size of the robust region, and develop an ADMM algorithm for efficient implementation. We validate our method in extensive numerical studies and demonstrate its superior performance. 2022-12-15T16:23:25Z ICML 2026 Kaizheng Wang Xiao Xu Xun Yu Zhou http://arxiv.org/abs/2606.13697v1 On Reference-Regulated Multiperiod Mean-Variance Portfolio Optimization in High Dimensions 2026-05-31T14:05:52Z The multiperiod mean-variance (MV) portfolio optimization serves as a vital expansion of Markowitz's static MV portfolio selection framework. Just like its static counterpart, the multiperiod MV portfolio remains susceptible to estimation errors. We propose a reference-regulated multiperiod mean-variance (RRMV) framework that penalizes deviations from a reference policy. Therefore, this new optimization successfully combines the advantages of dynamic strategies and reference portfolios. A key contribution of this paper is the characterization of the out-of-sample Sharpe ratio under high-dimensional asymptotics with estimation errors in both the mean vector and the covariance matrix. We show how the reference penalty and the investment horizon jointly affect the optimized portfolio performance, and how regularization operates differently from the single-period portfolio optimization. Extensive simulation and real data studies demonstrate that the proposed framework improves the stability and out-of-sample Sharpe ratios of multiperiod policies significantly. 2026-05-31T14:05:52Z Yutao Deng Jianjun Gao Weichen Wang