https://arxiv.org/api/OUJd/M+Fe5pBxyD4p3omafpVwHg2026-07-21T08:48:51Z24124515http://arxiv.org/abs/2205.08614v4Well Posedness of Utility Maximization Problems Under Partial Information in a Market with Gaussian Drift2026-06-21T10:17:27ZThis paper investigates well posedness of utility maximization problems for financial markets where stock returns depend on a hidden Gaussian mean-reverting drift process. Since that process is potentially unbounded, well posedness cannot be guaranteed for utility functions which are not bounded from above. For power utility with relative risk aversion smaller than that of log-utility this leads to restrictions on the choice of model parameters such as the investment horizon and parameters controlling the variance of the asset price and drift processes. We derive sufficient conditions to the model parameters leading to bounded maximum expected utility of terminal wealth for models with full and partial information.2022-05-17T20:14:04Z26 pages, 1 figureAbdelali GabihHakam KondakjiRalf Wunderlichhttp://arxiv.org/abs/2606.20903v1Reinforcement Learning for Risk-Sensitive Investment Management: a Free Energy--Entropy Duality Approach2026-06-18T19:58:15ZThis paper develops a reinforcement-learning approach to continuous-time risk-sensitive benchmarked asset allocation in a partly model-based setting. The benchmarked problem does not directly fit the standard Markovian stochastic-control template: the state is uncontrolled, whereas the terminal reward contains a controlled Itô integral. We use free energy-entropy duality to reformulate the problem as a linear-quadratic-Gaussian stochastic differential game under an equivalent probability measure, yielding explicit finite- and infinite-horizon saddle-point solutions. This structure guides a continuous-time $q$-learning actor-critic method: the quadratic value function motivates the critic, while the affine saddle-point controls motivate deterministic actors for the portfolio allocation and adversarial control. The learned allocation admits an economic interpretation through fractional Kelly decompositions. A proof-of-concept implementation calibrated to U.S. equity data shows that the actors learn the optimal policy with high accuracy and reveals a favorable asymmetry: the portfolio actor receives a cleaner learning signal than the auxiliary adversarial actor.2026-06-18T19:58:15ZSebastien LleoWolfgang Runggaldierhttp://arxiv.org/abs/2111.14631v2Model Risk in Credit Portfolio Models2026-06-16T08:16:51ZModel risk in credit portfolio models is a serious issue for banks but has so far not been tackled comprehensively. We will demonstrate how to deal with uncertainty in all model parameters in an all-embracing, yet easy-to-implement way.2021-11-23T13:12:47Z12 pages, 2 figures. This version: minor corrections, updates, and commentsChristian Meyerhttp://arxiv.org/abs/2606.17032v1Sharpe Ratio and Return-VaR Ratio Maximization for Option Portfolios with Skew-Elliptical $t$ Underlying Returns2026-06-15T17:52:47ZWe provide a formulation for optimal option portfolios under Sharpe Ratio maximization when the underlying returns follow a skew-elliptical t-distribution. This departs from the traditional normal returns setting in the context of Sharpe ratio maximization by allowing the modelling of heavy-tailed and skewed dynamics. The novelty of this paper and our main result is to provide explicit formulas for the portfolio weights when maximizing the Sharpe ratio and return-to-Value-at-Risk (VaR) ratio in the skew-elliptical setting. Numerical experiments reveal that the optimal portfolios for the two ratios are different.2026-06-15T17:52:47Z14 pagesKyle SungTraian A. Pirvuhttp://arxiv.org/abs/2410.20060v3Constrained portfolio optimization in a life-cycle model: A deep pricing kernel approach2026-06-15T09:52:52ZThis paper considers the constrained portfolio optimization in a generalized life-cycle model. The individual with a stochastic income manages a portfolio consisting of stocks, a bond, and life insurance to maximize their consumption level, death benefit, and terminal wealth. Meanwhile, the individual faces a convex-set trading constraint, with the non-tradeable asset constraint, no short-selling constraint, and no borrowing constraint as special cases. We build the artificial markets to solve this problem by manipulating the compensated drift terms of the underlying assets to meet the trading constraints. By dual transform, we propose a deep pricing kernel approach to compute tight lower and upper bounds for the primal problem, which can be used when the value function lacks an explicit solution due to the pricing kernel's conditional expectation. Finally, we conclude that when considering the trading constraints, the individual will reduce their consumption, demand for life insurance and annuities, and wealth levels due to the restricted market.2024-10-26T03:43:16ZWenyuan LiPengyu Weihttp://arxiv.org/abs/2606.09025v2Continuous Cash-Overlay Filters for a Static Growth--Defensive Risk Sleeve: Slow-Tail Compensation, V-Shape Crash Brakes, Walk-Forward Validation, and Max-Cash Combination2026-06-14T13:59:39ZThis paper studies a modular cash-overlay rule for allocating between a fixed growth-defensive risky sleeve R and interest-bearing cash C. The risky sleeve is a static 50/50 combination of equal-weight growth/technology and defensive income/value ETF baskets; the target is future R-C return, with the cash leg earning the contemporaneous cash rate. Two independent filters are tested. The slow-tail filter maps continuous compensation, rate-headwind, risk-premium-compression, and rate-path-stress states into a cash weight with a 30% material-trade gate. The V-shape filter is a fast crash brake based on continuous VIX, rate, credit, drawdown, and re-entry states. A fixed max-cash layer then uses the larger cash weight requested by either filter each day. On the 2017-2026 common window, the selected max-cash combination earns an 18.83% CAGR versus 16.62% for 100% R and reduces maximum drawdown from -33.59% to -18.05%. In the main walk-forward OOS window, the expanding combination earns 19.35% versus 17.59% for 100% R, with maximum drawdown of -22.05% versus -33.59%; the rolling version earns 18.50% with the same -22.05% drawdown. Post-2022 tests show lower drawdown but lower CAGR during a strong risky-sleeve rebound. The results support modular cash overlays as drawdown-control tools rather than standalone return-enhancement claims; fully real-time variable re-screening and multiple-testing-adjusted inference remain future work.2026-06-08T04:47:43Zdynamic asset allocation; cash overlay; crash protection; VIX; interest rates; credit stress; walk-forward validation; drawdown controlZheli Xionghttp://arxiv.org/abs/2606.14386v1Discovery under Hypothesis Redundancy: A Geometric Theory of Discovery Bottlenecks2026-06-12T12:21:09ZScientific discovery saturates when new hypotheses cease to provide independent information, even if the nominal hypothesis space remains large. We study hybrid discovery systems that combine structured local search with LLM-generated non-local proposals and pose the Search Compression Hypothesis: non-local exploration helps only when three geometric conditions co-occur: spectral compression, orthogonal escape from the explored span, and residual signal alignment with the target. We formalize these conditions, derive necessary conditions for hybrid advantage, and test the mechanism in controlled synthetic environments, large-scale A-share factor discovery, and symbolic-regression benchmarks; a public tabular operational sanity check tests the associated budget-allocation implication. Signal-planting and directed-versus-random experiments show that novelty alone is insufficient: random orthogonal jumps expand coverage but do not improve yield without predictive alignment. Across compression sweeps, real factor archives, and LLM-SRBench tasks, hybrid gains concentrate in weakly represented but target-bearing directions and vanish as the hypothesis space approaches full rank. The framework turns LLM-guided discovery from generic novelty search into a diagnostic procedure for deciding when directed non-local exploration is warranted.2026-06-12T12:21:09Z23 pages, 1 figure, 27 tablesLi XiaBaoxun Wanghttp://arxiv.org/abs/2606.14050v1Battery Bidding under Price Uncertainty in Wholesale Electricity Markets2026-06-12T02:51:55ZGrid-scale batteries increasingly influence outcomes in wholesale electricity markets, but their observed bid patterns remain difficult to interpret. In particular, bids that appear to reflect strategic withholding may instead arise from rational operations under price uncertainty and risk management. We develop an asset-level model of a price-taking battery that submits stepwise buy and sell bid curves in the day-ahead market under a finite set of price scenarios. The battery chooses quantity--price pairs to maximize a mean--CVaR objective subject to physical and market constraints. A direct formulation is a mixed-integer linear program, but we show that its integer decisions can be removed, yielding an exact linear programming reformulation suitable for empirical analysis. Our empirical results deliver three insights. First, withholding behavior can arise even without market power, because scarce stored energy and uncertain future prices increase the value of holding energy. Second, the effect of uncertainty depends on the state of charge: when stored energy is scarce, greater uncertainty raises sell bid prices, whereas when stored energy is abundant it can lower them. Third, risk management reshapes bid curves into layered structures that secure profitable execution across a broad set of scenarios while preserving some exposure to rare but valuable price spikes.2026-06-12T02:51:55ZVincent Yinjun-WangMadeleine Udellhttp://arxiv.org/abs/2606.13618v1A Declining CVaR Glidepath Framework for Target-Date Fund Design with an Application to the Chilean Pension System2026-06-11T17:31:27ZWe propose a framework for designing Target-Date Funds (TDFs) around an explicit return objective while controlling risk directly at the portfolio level through a declining Conditional Value-at-Risk (CVaR) constraint. In this approach, the regulator or sponsor specifies a CVaR glidepath that gives the portfolio manager enough flexibility to reach a target return with a reasonably high probability. The target return is determined exogenously from pension-design inputs such as retirement age, contribution rate, working years, life expectancy, and replacement-rate goals. This differs from conventional TDF design, where age-dependent asset-class limits are set without an explicit link to a required return.
A key feature of the method is that it does not assume the manager selects an optimal portfolio each period. Instead, each month the manager draws an allocation from the set of portfolios satisfying the CVaR constraint. This yields a conservative evaluation of each glidepath: success probabilities are averages over admissible allocations, rather than best-case outcomes. We introduce two figures of merit: the probability of meeting the target return and the cumulative risk assumed over the life of the TDF.
As a proof of concept, we apply the framework to Chile's 2025 pension reform using nine Chilean and global asset classes and a 40-year accumulation horizon. The results show that the transition age at which risk starts to decline is the most consequential design parameter, and that contribution density acts as a hard constraint: below a critical threshold, portfolio design alone cannot compensate for structurally low contributions. The framework is general and can be applied to any TDF designed around an explicit return objective.2026-06-11T17:31:27Z29 pages, 3 figuresIsrael MuñozFernando SuárezOmar LarréArturo Cifuenteshttp://arxiv.org/abs/2606.14798v1Two Sides of Schur Damping: High-Dimensional Pseudo-Likelihoods and Portfolio Allocation2026-06-11T16:44:50ZTwo communities that rarely cite each other -- spatial statisticians fitting high-dimensional weather fields, and quantitative investors building portfolios -- have independently arrived at the same mathematical object: a Schur complement, damped by one interpretable parameter. In spatial modeling the Schur complement is the conditional covariance that makes a Gaussian (Vecchia) pseudo-likelihood estimable at scale, and recent work regularizes it by shrinking toward a base model. In allocation it is the residual risk of a bet net of its hedge, and the same parameter interpolates hierarchical risk parity and the minimum-variance portfolio. We show these are one operation -- reliability shrinkage of a conditional Gaussian -- so that the damping a weather model needs to remain estimable when stations outnumber observations is, term for term, the damping a portfolio needs to remain stable when assets outnumber returns. The optimal amount is a closed-form reliability, a James-Stein shrinkage that is simultaneously a Ledoit-Wolf intensity. The shrinkage machinery is classical, but the identity appears to be new: to our knowledge neither literature has noted that the conditional shrinkage a spatial model fits and the diversification-variance tilt a portfolio chooses are one and the same quantity. We make the correspondence precise, note that the two literatures have each supplied what the other lacks, and report a small experiment on the one genuinely open choice -- how to set the damping -- suggesting the spatial community's fitted intensity is, if anything, the better recipe.2026-06-11T16:44:50ZPeter Cottonhttp://arxiv.org/abs/2510.25740v2A mathematical study of the excess growth rate2026-06-10T21:54:30ZThe excess growth rate, defined as the gap in Jensen's inequality for the logarithm, is a fundamental functional in portfolio theory. In this paper, we present a mathematical study motivated by information theory. We begin by establishing its properties and showing that it has rich connections with information theoretic concepts such as the Helmholtz free energy, L. Campbell's measure of average code length and large deviations. Our main results consist of three axiomatic characterization theorems of the excess growth rate, in terms of (i) the relative entropy, (ii) the gap in Jensen's inequality, and (iii) the logarithmic divergence that generalizes the Bregman divergence. Furthermore, we study maximization of the excess growth rate and compare it with the growth optimal portfolio. Our results not only provide theoretical justifications of the significance of the excess growth rate, but also establish new connections between information theory and quantitative finance.2025-10-29T17:43:40Z54 pages, 2 figuresSteven CampbellTing-Kam Leonard Wonghttp://arxiv.org/abs/2606.12612v1The Mathematics of Heuristic Portfolio Optimization (HPO)2026-06-10T19:13:53ZPractitioners allocate capital with forecast-light rules such as equal weight, inverse volatility, risk parity, HRP, and return-adjusted HRP (RA-HRP). This paper develops \emph{Heuristic Portfolio Optimization} (HPO): an information-restricted projection of the Markowitz/tangency solution onto a stable rule class. The implied-return principle, $\mathbf{w}$ is maximum-Sharpe iff $\mathbfμ_e \propto \mathbfΣ\mathbf{w}$, gives closed-form optimality sets for leading heuristics and exposes the Schur-complement substitutions behind HRP. For RA-HRP, we introduce fixed-tree cluster-Sharpe recursion, unit-free HRP--RA-HRP interpolation, tangency conditions, conditional-risk splits, and pathwise/KL decompositions of weight distortion. First-order Sharpe calculus expresses the marginal value of return information as nodewise alphas against HRP and yields a linear KL trust budget. We formalize generic HPO maps, define the implied-return defect, prove that it equals squared Sharpe inefficiency, characterize tree-HPO coincidence by nodewise mass ratios, and give a bias--variance decomposition for estimated rules. Finally, HPO is embedded into Reinforcement Learning Portfolio Optimization (RLPO): every HPO map induces a deterministic stationary policy; static HPO is the $γ=0$ no-friction face of the Bellman problem; RA-HRP supplies a hierarchical policy prior; and dynamic improvement is warranted when continuation value exceeds myopic HPO defect plus frictions. A performance-difference identity prices the myopic value gap, gives an $\varepsilon/(1-γ)$ myopia bound, and identifies nodewise alphas as policy-gradient coordinates of the hierarchical actor. Thus HPO is the static optimality layer and RLPO the dynamic control layer. The conditions are GRS-testable, extend to mean--CVaR and expected utility under ellipticity, and become Kelly-growth conditions in diffusion limits.2026-06-10T19:13:53ZMiquel Noguer i Alonsohttp://arxiv.org/abs/2411.13579v2Optimal portfolio under ratio-type periodic evaluation in stochastic factor models under convex trading constraints2026-06-10T06:48:02ZThis paper studies a type of periodic utility maximization problem for portfolio management in incomplete stochastic factor models with convex trading constraints. The portfolio performance is periodically evaluated on the relative ratio of two adjacent wealth levels over an infinite horizon, featuring the dynamic adjustments in portfolio decision according to past achievements. Under power utility, we transform the original infinite horizon optimal control problem into an auxiliary terminal wealth optimization problem under a modified utility function. To cope with the convex trading constraints, we further introduce an auxiliary unconstrained optimization problem in a modified market model and develop the martingale duality approach to establish the existence of the dual minimizer such that the optimal unconstrained wealth process can be obtained using the dual representation. With the help of the duality results in the auxiliary problems, the relationship between the constrained and unconstrained models as well as some fixed point arguments, we derive and verify the optimal constrained portfolio process for the original problem over an infinite horizon.2024-11-15T13:19:31ZKeywords: Periodic evaluation, relative portfolio performance, incomplete market, stochastic factor model, convex trading constraints, convex duality approach. This manuscript combines two previous preprints arXiv:2311.12517 and arXiv:2401.14672 into one paper with more general and improved resultsWenyuan WangKaixin YanXiang Yuhttp://arxiv.org/abs/1911.04090v3A post hoc test on the Sharpe ratio2026-06-10T01:16:51ZWe describe a post hoc test for the Sharpe ratio, analogous to Tukey's test for pairwise equality of means. The test can be applied after rejection of the hypothesis that all population Signal-Noise ratios are equal. The test is applicable under a simple correlation structure among asset returns. Simulations indicate the test maintains nominal type I rate under a wide range of conditions and is moderately powerful under reasonable alternatives.2019-11-11T05:47:30ZSteven E. Pavhttp://arxiv.org/abs/2606.01650v2Post Selection Estimation of Sharpe Ratios2026-06-10T01:00:57ZWe consider the problem of estimating the true Sharpe ratio of an asset selected for having the highest observed in-sample Sharpe ratio among many assets. We discuss estimators based on the polyhedral lemma, James Stein shrinkage, debiasing the expected maximum Sharpe ratio, thresholding and empirical Bayes. We test these estimators in simulations, computing bias and root mean square error across different values of sample size, number of assets, and spread and shape of population Sharpe ratios. We also compute rank correlation of the estimators against the underlying quantity, simulating how these estimators might be used to compare or rank the output of different teams which perform this selection process. We find that the James Stein estimator provides the best performance across many different realistic values of the relevant parameters, followed by the GMLEB estimator of Jiang and Zhang. These results are fairly robust to correlation of asset returns, with some caveats.2026-06-01T03:58:35ZSteven E. Pav