https://arxiv.org/api/glqtcY7/WChcDvbgwgn5QzzVY60 2026-07-22T23:01:55Z 2414 150 15 http://arxiv.org/abs/2404.00825v2 Using Machine Learning to Forecast Market Direction with Efficient Frontier Coefficients 2026-04-05T03:39:21Z We propose a novel method to improve estimation of asset returns for portfolio optimization. This approach first performs a monthly directional market forecast using an online decision tree. The decision tree is trained on a novel set of features engineered from portfolio theory: the efficient frontier functional coefficients. Efficient frontiers can be decomposed to their functional form, a square-root second-order polynomial, and the coefficients of this function captures the information of all the constituents that compose the market in the current time period. To make these forecasts actionable, these directional forecasts are integrated to a portfolio optimization framework using expected returns conditional on the market forecast as an estimate for the return vector. This conditional expectation is calculated using the inverse Mills ratio, and the Capital Asset Pricing Model is used to translate the market forecast to individual asset forecasts. This novel method outperforms baseline portfolios, as well as other feature sets including technical indicators and the Fama-French factors. To empirically validate the proposed model, we employ a set of market sector ETFs. 2024-03-31T23:32:34Z Code: https://github.com/nolanalexander/efficient-frontier-coefficients Nolan Alexander William Scherer 10.3905/jfds.2023.1.128 http://arxiv.org/abs/2604.14206v1 Portfolio Optimization Proxies under Label Scarcity and Regime Shifts via Bayesian and Deterministic Students under Semi-Supervised Sandwich Training 2026-04-04T06:42:38Z This paper proposes a machine learning assisted portfolio optimization framework designed for low data environments and regime uncertainty. We construct a teacher student learning pipeline in which a Conditional Value at Risk (CVaR) optimizer generates supervisory labels, and neural models (Bayesian and deterministic) are trained using both real and synthetically augmented data. The synthetic data is generated using a factor based model with t copula residuals, enabling training beyond the limited real sample of 104 labeled observations. We evaluate four student models under a structured experimental framework comprising (i) controlled synthetic experiments (3 x 5 seed grid), (ii) in-distribution real market evaluation (C2A) and (iii) cross-universe generalization (D2A). In real-market settings, models are deployed using a rolling evaluation protocol where a frozen pretrained model is periodically fine tuned on recent observations and reset to its base state, ensuring stability while allowing limited adaptation. Results show that student models can match or outperform the CVaR teacher in several settings, while achieving improved robustness under regime shifts and reduced turnover. These findings suggest that hybrid optimization learning approaches can enhance portfolio construction in data constrained environments 2026-04-04T06:42:38Z 18 pages of main text. 10 pages of appendices. 35 references. Around 13 figures Adhiraj Chattopadhyay http://arxiv.org/abs/2604.02279v1 The Self Driving Portfolio: Agentic Architecture for Institutional Asset Management 2026-04-02T17:13:04Z Agentic AI shifts the investor's role from analytical execution to oversight. We present an agentic strategic asset allocation pipeline in which approximately 50 specialized agents produce capital market assumptions, construct portfolios using over 20 competing methods, and critique and vote on each other's output. A researcher agent proposes new portfolio construction methods not yet represented, and a meta-agent compares past forecasts against realized returns and rewrites agent code and prompts to improve future performance. The entire pipeline is governed by the Investment Policy Statement--the same document that guides human portfolio managers can now constrain and direct autonomous agents. 2026-04-02T17:13:04Z 31 pages, 11 exhibits Andrew Ang Nazym Azimbayev Andrey Kim http://arxiv.org/abs/2603.21672v3 Mislearning of Factor Risk Premia under Structural Breaks: A Misspecified Bayesian Learning Framework 2026-03-31T23:16:16Z While asset-pricing models increasingly recognize that factor risk premia are subject to structural change, existing literature typically assumes that investors correctly account for such instability. This paper studies how investors instead learn under a misspecified model that underestimates structural breaks. We propose a minimal Bayesian framework in which this misspecification generates persistent prediction errors and pricing distortions, and we introduce an empirically tractable measure of mislearning intensity $(Δ_t)$ based on predictive likelihood ratios. The empirical results yield three main findings. First, in benchmark factor systems, elevated mislearning does not forecast a deterministic short-run collapse in performance; instead, it is associated with stronger long-horizon returns and Sharpe ratios, consistent with an equilibrium premium for acute model uncertainty. Second, in a broader anomaly universe, this pricing relation does not generalize uniformly: mislearning is more strongly associated with future drawdowns, downside semivolatility, and other measures of instability, with substantial heterogeneity across anomaly families. Third, the cross-sectional relation between instability and mislearning is inherently conditional: while a monotonic link between break-proneness and average mislearning does not hold in the full cross-section, it re-emerges in low-friction (low-IVOL) environments where break-state severity is more comparable across assets. 2026-03-23T07:54:15Z Yimeng Qiu http://arxiv.org/abs/2603.29994v1 Bridging Stochastic Control and Deep Hedging: Structural Priors for No-Transaction Band Networks 2026-03-31T16:56:17Z This paper studies the problem of hedging and pricing a European call option under proportional transaction costs, from two complementary perspectives. We first derive the optimal hedging strategy under CARA utility, following the stochastic control framework of Davis et al. (1993), characterising the no-transaction band via the Hamilton-Jacobi-Bellman Quasi-Variational Inequality (HJBQVI) and the Whalley-Wilmott asymptotic approximation. We then adopt a deep hedging approach, proposing two architectures that build on the No-Transaction Band Network of Imaki et al. (2023): NTBN-Delta, which makes delta-centring explicit, and WW-NTBN, which incorporates the Whalley-Wilmott formula as a structural prior on the bandwidth and replaces the hard clamp with a differentiable soft clamp. Numerical experiments show that WW-NTBN converges faster, matches the stochastic control no-transaction bands more closely, and generalises well across transaction cost regimes. We further apply both frameworks to the bull call spread, documenting the breakdown of price linearity under transaction costs. 2026-03-31T16:56:17Z Jules Arzel Noureddine Lehdili http://arxiv.org/abs/2603.29751v1 Common Risk Factors in Decentralized AI Subnets 2026-03-31T13:52:14Z I derive a size premium from the constant-product automated market maker used to price Bittensor subnet tokens and test the prediction using daily data on 128 subnets. A small-minus-big factor earns 1.01% daily (Newey-West t = 3.28). The December 2025 halving of token emissions, which the theory predicts should halve the premium, reduces it from 1.17% to 0.51% (p = 0.044). Exact slippage calculations show the premium is implementable only below \$10K in assets under management; at \$100K, transaction costs exceed gross returns. 2026-03-31T13:52:14Z 40 pages, 11 figures, 7 tables Philip Z. Maymin http://arxiv.org/abs/2603.01298v2 Single-Asset Adaptive Leveraged Volatility Control 2026-03-30T18:55:12Z This paper introduces a methodology for constructing a market index composed of a liquid risky asset and a liquid risk-free asset that achieves a fixed target volatility. Existing volatility-targeting strategies typically scale portfolio exposure inversely with a variance forecast, but such open-loop approaches suffer from high turnover, leverage spikes, and sensitivity to estimation error -- issues that limit practical adoption in index construction. We propose a proportional-control approach for setting the index weights that explicitly corrects tracking error through feedback. The method requires only a few interpretable parameters, making it transparent and practical for index construction. We demonstrate in simulation that this approach is more effective at consistently achieving the target volatility than the open-loop alternative. 2026-03-01T22:03:27Z 16 pages, 7 figures Nikhil Devanathan Dylan Rueter Stephen Boyd Emmanuel Candès Trevor Hastie Mykel J. Kochenderfer Arpit Apoorv David Soronow Igor Zamkovsky http://arxiv.org/abs/2511.07014v2 Diffolio: A Diffusion Model for Multivariate Probabilistic Financial Time-Series Forecasting and Portfolio Construction 2026-03-29T05:49:35Z Probabilistic forecasting is crucial in multivariate financial time-series for constructing efficient portfolios that account for complex cross-sectional dependencies. In this paper, we propose Diffolio, a diffusion model designed for multivariate financial time-series forecasting and portfolio construction. Diffolio employs a denoising network with a hierarchical attention architecture, comprising both asset-level and market-level layers. Furthermore, to better reflect cross-sectional correlations, we introduce a correlation-guided regularizer informed by a stable estimate of the target correlation matrix. This structure effectively extracts salient features not only from historical returns but also from asset-specific and systematic covariates, significantly enhancing the performance of forecasts and portfolios. Experimental results on the daily excess returns of 12 industry portfolios show that Diffolio outperforms various probabilistic forecasting baselines in multivariate forecasting accuracy and portfolio performance. Moreover, in portfolio experiments, portfolios constructed from Diffolio's forecasts show consistently robust performance, thereby outperforming those from benchmarks by achieving higher Sharpe ratios for the mean-variance tangency portfolio and higher certainty equivalents for the growth-optimal portfolio. These results demonstrate the superiority of our proposed Diffolio in terms of not only statistical accuracy but also economic significance. 2025-11-10T12:05:32Z 41 pages, 11 figures. Replacement to match the version accepted for publication in Information Fusion (Vol. 133, 104286, 2026). Significant updates have been made from the initial draft to reflect the final accepted manuscript (AAM) Information Fusion, Vol. 133, 104286 (2026) So-Yoon Cho Jin-Young Kim Kayoung Ban Hyeng Keun Koo Hyun-Gyoon Kim 10.1016/j.inffus.2026.104286 http://arxiv.org/abs/2412.16175v3 Mean--Variance Portfolio Selection by Continuous-Time Reinforcement Learning: Algorithms, Regret Analysis, and Empirical Study 2026-03-28T02:52:23Z We study continuous-time mean--variance portfolio selection in markets where stock prices are diffusion processes driven by observable factors that are also diffusion processes, yet the coefficients of these processes are unknown. Based on the recently developed reinforcement learning (RL) theory for diffusion processes, we present a general data-driven RL approach that learns the pre-committed investment strategy directly without attempting to learn or estimate the market coefficients. For multi-stock Black--Scholes markets without factors, we further devise an algorithm and prove its performance guarantee by deriving a sublinear regret bound in terms of the Sharpe ratio. We then carry out an extensive empirical study implementing this algorithm to compare its performance and trading characteristics, evaluated under a host of common metrics, with a large number of widely employed portfolio allocation strategies on S\&P 500 constituents. The results demonstrate that the proposed continuous-time RL strategy is consistently among the best, especially in a volatile bear market, and decisively outperforms the model-based continuous-time counterparts by significant margins. 2024-12-08T15:31:10Z 94 pages, 8 figures, 18 tables Yilie Huang Yanwei Jia Xun Yu Zhou http://arxiv.org/abs/2603.26620v1 Optimal Parlay Wagering and Whitrow Asymptotics: A State-Price and Implicit-Cash Treatment 2026-03-27T17:17:21Z For independent multi-outcome events under multiplicative parlay pricing, we give a short exact proof of the optimal Kelly strategy using the implicit-cash viewpoint. The proof is entirely eventwise. One first solves each event in isolation. The full simultaneous optimizer over the entire menu of singles, doubles, triples, and higher parlays is then obtained by taking the outer product of the one-event Kelly strategies. Equivalently, the optimal terminal wealth factorizes across events. This yields an immediate active-leg criterion: a parlay is active if and only if each of its legs is active in the corresponding one-event problem. The result recovers, in a more transparent state-price form, the log-utility equivalence between simultaneous multibetting and sequential Kelly betting. We then study what is lost when one forbids parlays and allows only singles. In a low-edge regime and on a fixed active support, the exact parlay optimizer supplies the natural reference point. The singles-only problem is a first-order truncation of the factorized wealth formula. A perturbative expansion shows that the growth-rate loss from forbidding parlays is $\OO(\eps^4)$, while the optimal singles stakes deviate from the isolated one-event Kelly stakes only at cubic order. This yields a clean explanation of Whitrow's empirical near-proportionality phenomenon: the simultaneous singles-only optimizer is obtained from the isolated eventwise optimizer by an event-specific cubic shrinkage, so the portfolios agree through second order and differ only by a small blockwise drag. 2026-03-27T17:17:21Z 10 pages, 0 figures Christopher D. Long http://arxiv.org/abs/2603.24154v1 The Geometry of Risk: Path-Dependent Regulation and Anticipatory Hedging via the SigSwap 2026-03-25T10:24:11Z This paper introduces a transformative framework for managing path-dependent financial risk by shifting from traditional distribution-centric models to a geometry-based approach. We propose the SigSwap as a new regulatory instrument that allows market participants to decompose complex risk into terminal price law and the underlying texture of the price path. By utilising the mathematical properties of the path-signature, we demonstrate how previously unmodellable risks, such as lead-lag dynamics and flash-crash spiralling, can be converted into transparent and linear risk factors. Central to this framework is the introduction of Signature Expected Shortfall, a risk metric designed to capture toxic path geometries that traditional methods often overlook. We also present a proactive monitoring system based on the Temporal Exposure Profile, which utilises anticipatory learning to detect potential liquidity traps and geometric decoupling before they manifest as realised volatility. The proposed methodology offers a rigorous alignment with global regulatory mandates, specifically the Fundamental Review of the Trading Book (FRTB), by providing a consistent bridge between physical stress-testing and risk-neutral hedging. Finally, we show that this algebraic approach significantly reduces computational complexity, enabling real-time, high-frequency risk reporting and capital optimisation for the modern financial ecosystem. 2026-03-25T10:24:11Z Daniel Bloch http://arxiv.org/abs/2603.24064v1 Utility-Invariant Support Selection and Eventwise Decoupling for Simultaneous Independent Multi-Outcome Bets 2026-03-25T08:18:33Z For simultaneous independent events with finitely many outcomes, consider the expected-utility problem with nonnegative wagers and an endogenous cash position. We prove a short support theorem for a broad class of strictly increasing strictly concave utilities. On any fixed support family and at any optimal portfolio with positive cash, summing the active first-order conditions and comparing that sum with cash stationarity yields the exact identity \[ \fracλ{K_{\ell}^{(U)}}=\frac{1-P_{\ell,A}}{1-Q_{\ell,A}}, \] where $P_{\ell,A}$ and $Q_{\ell,A}$ are the active probability and price masses of event $\ell$, $λ$ is the budget multiplier, and $K_{\ell}^{(U)}$ is the continuation factor seen by inactive outcomes of that event. Consequently, after sorting each event by the edge ratio $p_{\ell i}/π_{\ell i}$, the exact active support is the eventwise union of the single-event supports, and this support is independent of the utility function. The single-event utility-invariant support theorem is already explicit in the free-exposure pari-mutuel setting in Smoczynski and Miles; the point of the present note is that the simultaneous independent-events analogue follows from the same state-price geometry once the right continuation factor is identified. 2026-03-25T08:18:33Z 7 pages, no figures Christopher D. Long http://arxiv.org/abs/2603.23300v1 Designing Agentic AI-Based Screening for Portfolio Investment 2026-03-24T15:03:40Z We introduce a new agentic artificial intelligence (AI) platform for portfolio management. Our architecture consists of three layers. First, two large language model (LLM) agents are assigned specialized tasks: one agent screens for firms with desirable fundamentals, while a sentiment analysis agent screens for firms with desirable news. Second, these agents deliberate to generate and agree upon buy and sell signals from a large portfolio, substantially narrowing the pool of candidate assets. Finally, we apply a high-dimensional precision matrix estimation procedure to determine optimal portfolio weights. A defining theoretical feature of our framework is that the number of assets in the portfolio is itself a random variable, realized through the screening process. We introduce the concept of sensible screening and establish that, under mild screening errors, the squared Sharpe ratio of the screened portfolio consistently estimates its target. Empirically, our method achieves superior Sharpe ratios relative to an unscreened baseline portfolio and to conventional screening approaches, evaluated on S&P 500 data over the period 2020--2024. 2026-03-24T15:03:40Z Mehmet Caner Agostino Capponi Nathan Sun Jonathan Y. Tan http://arxiv.org/abs/2511.19186v2 Carbon-Penalised Portfolio Insurance Strategies in a Stochastic Factor Model with Partial Information 2026-03-24T08:48:39Z Given the increasing importance of environmental, social and governance (ESG) factors, particularly carbon emissions, we investigate optimal proportional portfolio insurance (PPI) strategies accounting for carbon footprint reduction. PPI strategies enable investors to mitigate downside risk while retaining the potential for upside gains. This paper aims to determine the multiplier of the PPI strategy to maximise the expected utility of the terminal cushion, where the terminal cushion is penalised proportionally to the realised volatility of stocks issued by firms operating in carbon-intensive sectors. We model the risky assets' dynamics using geometric Brownian motions whose drift rates are modulated by an unobservable common stochastic factor to capture market-specific or economy-wide state variables that are typically not directly observable. Using classical stochastic filtering theory, we formulate a suitable optimization problem and solve it for CRRA utility function. We characterise optimal carbon penalised PPI strategies and optimal value functions under full and partial information and quantify the loss of utility due incomplete information. Finally, we carry a numerical analysis showing that the proposed strategy reduces carbon emission intensity without compromising financial performance. 2025-11-24T14:53:55Z Katia Colaneri Federico D'Amario Daniele Mancinelli http://arxiv.org/abs/2603.22880v1 Portfolio Optimization under Recursive Utility via Reinforcement Learning 2026-03-24T07:25:30Z We study whether a risk-sensitive objective from asset-pricing theory -- recursive utility -- improves reinforcement learning for portfolio allocation. The Bellman equation under recursive utility involves a certainty equivalent (CE) of future value that has no closed form under observed returns; we approximate it by $K$-sample Monte Carlo and train actor-critic (PPO, A2C) on the resulting value target and an approximate advantage estimate (AAE) that generalizes the Bellman residual to multi-step with state-dependent weights. This formulation applies only to critic-based algorithms. On 10 chronological train/test splits of South Korean ETF data, the recursive-utility agent improves on the discounted (naive) baseline in Sharpe ratio, max drawdown, and cumulative return. Derivations, world model and metrics, and full result tables are in the appendices. 2026-03-24T07:25:30Z Minkey Chang