https://arxiv.org/api/qw3E7U0fpjjp9r+1wBVp67UOx902026-07-22T23:48:13Z298516515http://arxiv.org/abs/2604.27447v1Sampler-Robust Optimization under Generative Models2026-04-30T05:33:15ZModern stochastic optimization pipelines increasingly rely on learned generative models to represent uncertainty, while downstream decisions are evaluated almost entirely through Monte Carlo scenarios. This shifts the operational object of uncertainty from an explicit probability law to the sampler induced by the learned generator. Reliability therefore depends on two errors: sampler misspecification and finite-simulation error. We propose Sampler-Robust Optimization (SRO), which optimizes decisions against the worst-case sampler induced by perturbing the learned generator. This sampler-first formulation aligns with simulation-based decision pipelines and admits a sharpness-aware interpretation: it favors decisions whose performance is stable under generator perturbations, rather than merely under the nominal sampler. Under a coverage assumption, we show that the empirical worst-case objective provides a high-probability upper certificate for the true population objective, with finite-simulation error partially absorbed by the robustification used to guard against sampler misspecification. The framework accommodates generative models with or without explicit densities and admits efficient minimax procedures. Portfolio-optimization experiments show that SRO produces more stable decisions and improves out-of-sample performance under distribution shift.2026-04-30T05:33:15ZZiwei ZhangJonathan Yu-Meng Lihttp://arxiv.org/abs/2604.25406v1A Motif-Based Framework for Decomposing Risk Spillovers2026-04-28T09:17:31ZConnectedness measures quantify aggregate risk spillovers but obscure the local interaction patterns that generate systemic risk. We develop a motif-based framework that first extracts multiscale backbones from quantile connectedness networks and then identifies directed triadic motifs whose frequencies exceed randomization baselines. To distinguish how assets' sectoral identities shape local spillover structures, we introduce colored motifs under sector partitions of increasing granularity. Using orbit positions that capture each node's structural role within directed triadic motifs, we construct portfolio strategies that exploit an asset's place in the spillover architecture. Applying the framework to 39 commodity and equity futures across lower, median, and upper conditional quantiles, we find that motif-based portfolios outperform minimum correlation and minimum connectedness benchmarks on risk-adjusted returns. We further show that in tail networks, assets with greater orbit-position diversity tend to act as net spillover transmitters rather than receivers, establishing positional diversity as a tail-specific marker of systemic influence. These findings demonstrate that local triadic topology carries portfolio-relevant information that aggregate connectedness measures miss.2026-04-28T09:17:31Z53 pages, 19 figuresYing-Hui ShaoYan-Hong YangYun Zhanghttp://arxiv.org/abs/2604.23983v1A Geometric Witness Framework for Signed Multivariate Tail-Dependence Compatibility: Asymptotic Structure and Finite-Threshold Synthesis2026-04-27T02:54:32ZWe study multivariate tail-dependence compatibility for complete and partial signed tail families, treating lower-tail, upper-tail, and mixed configurations in one geometric witness representation indexed by active coordinate sets and sign patterns. For a complete signed tail family, witness generator weights w = (w_{I,sigma}) give a linear incidence parametrization and are recovered by explicit triangular inversion. Excluding the geometric scale p0, the complete case uses 3^d - 1 generator weights, matching the number of complete signed tail coefficients; for partial specifications, only selected target coefficients need be prescribed. At a fixed threshold p0 in (0, 1/2), the inversion identifies the normalized noncentral ternary cell masses of any realizing copula. Hence finite-threshold compatibility is characterized by nonnegative recovered generator weights, singleton normalization, and the residual central-mass constraint.
This yields a complete Moebius-type synthesis within the witness framework. If the recovered increments are nonnegative and singleton normalization holds, then S(w) = sum(w) determines the admissible finite-scale range, and every admissible p0 gives an exact witness realization. In the canonical ray geometry, such a realization preserves the same complete signed tail family throughout 0 < p <= p0. Thus the primary object is the complete signed tail family lambda: it is realized at every admissible finite scale and can be carried along families of witness copulas with p0 decreasing to 0.
Partial, noisy, or inconsistent specifications are treated through linear-feasibility and weighted-l1 recovery problems in the same parametrization. The representation separates the p0-free incidence/Moebius layer from finite-threshold realization and provides tools for realization, simulation, calibration, completion, repair, and scenario design.2026-04-27T02:54:32Z47 pages, 4 figures, 3 tables; includes a Python implementation appendixJanusz Milekhttp://arxiv.org/abs/2511.11364v3Assessment of loan losses after default2026-04-26T08:57:51ZThe paper shows how to determine the loss on an LGD borrower's loan after default, with or without preparation of a separate model. LGD after default is estimated taking into account the average repayment period of the defaulted loan, knowledge of volumes, moments of default and repayments, the rate or other parameters in the vector of determinants. The calculation of the average repayment period for overdue loans is given in the article. A Bayesian scheme is used to estimate repayable debts, considering the percentage of repayment. A general recovery model was used for the LGD segment recovery process. Only this type of model allows you to set LGD less than or equal to 1, which is required for further estimates.2025-11-14T14:47:11Z14 pages, 4 figures (text on Russian)Pomazanov Mikhailhttp://arxiv.org/abs/2604.23315v1Multiplicative Contractions, Additive Recoveries: Functional-Form Restrictions on Risk Exposure Dynamics2026-04-25T14:09:21ZWe test a regime-conditional functional-form restriction on aggregate risk-exposure dynamics implied by VaR-constrained intermediary models: exposures contract multiplicatively when capital constraints bind and grow additively (level-independent) when slack. The contraction half follows from binding VaR constraints (Brunnermeier and Pedersen 2009; Adrian and Shin 2010; He and Krishnamurthy 2013). The additive-rebuild prediction is derived under constant-rate capital replenishment; we test the joint restriction on FINRA monthly margin debt (1997-2026).
Two findings. First, regime-interacted regression of detrended margin growth on lagged level (T=350 months) yields calm slope -0.040 (p=0.082, additive) and stress slope -0.205 (p<0.001, multiplicative); Wald test on regime x level interaction rejects equal dependence (p=0.0016). Second, the restriction implies drawdown-recovery duration ratio increases with crash depth. On 73 S&P 500 episodes (1950-2026), Cox model gives depth coefficient -13.75 (p<10^{-7}): 75% lower recovery hazard per 10pp deeper drawdown. Continuous-depth regression yields beta=1.22 (p=0.047); beta=1.59 (p<0.001) excluding 1980-82 Volcker. Median duration ratio for crashes >30% is 3.1x; replicates across eight other equity indices. Calibrated Heston, Markov-switching, and block bootstrap nulls match price-level duration asymmetry but lack an exposure state variable, so cannot speak to the regime-conditional flip on direct exposures.
We do not claim the exposure test identifies the intermediary mechanism: FINRA margin debt is a noisy proxy. We claim only that the regime-conditional functional form is a sharper target than return-level moments alone, and confirming it on margin debt is consistent with -- not proof of -- the constrained-intermediary mechanism. A companion test on CFTC weekly speculative positioning is left for future work (Sections 5.2 and F).2026-04-25T14:09:21ZLiang Chenhttp://arxiv.org/abs/2511.00764v2Further Developments on Stochastic Dominance for Convex Combinations of Infinite-Mean Random Variables2026-04-25T12:28:45ZIn recent years, stochastic dominance for independent and identically distributed (iid) infinite-mean random variables has received considerable attention. The literature has identified several classes of distributions of nonnegative random variables that encompass many common heavy-tailed distributions. A key result demonstrates that the weighted sum of iid random variables from these classes is stochastically larger than any individual random variable in the sense of the first-order stochastic dominance. This paper systematically investigates the properties and inclusion relationships among these distribution classes, and extends some existing results to more practical scenarios. Furthermore, we analyze the case where each random variable follows a compound binomial distribution, establishing necessary and sufficient conditions for the preservation of the aforementioned stochastic dominance relation.2025-11-02T01:57:58ZKeyi ZengZhenfeng ZouYuting SuTaizhong Huhttp://arxiv.org/abs/2604.23087v1Beyond Picking Winners: Correlation-Driven Tail Risk in Venture Capital Portfolio Construction2026-04-25T00:45:40ZWe propose a Gaussian-copula-based framework that learns deal-level dependence directly from observed joint success frequencies across founder, geography, and market attributes. Holding marginal deal success probabilities fixed, deal-level correlation preserves expected portfolio outcomes but shifts the portfolio distribution toward heavier right tails and higher kurtosis. In portfolio simulations, correlation reduces the probability of modest success counts while sharply amplifying extreme upside outcomes, especially in structurally concentrated portfolios. Our findings suggest that extreme venture capital outcomes may partly reflect correlation-induced tail amplification rather than solely higher average deal quality, with potential implications for portfolio construction and risk management. We note that the observed dataset reflects selected deals with observable outcomes, which inflates apparent success rates relative to the true population base rate; however, the core finding that correlation reshapes the distributional shape while leaving the mean unchanged is structurally robust to the level of marginal success probabilities.2026-04-25T00:45:40Z20 pages, 9 figuresYunqi LiangHasan Ugur KoyluogluFuat AlicanYigit Ihlamurhttp://arxiv.org/abs/2604.21893v1Revealing Geography-Driven Signals in Zone-Level Claim Frequency Models: An Empirical Study using Environmental and Visual Predictors2026-04-23T17:44:52ZGeographic context is often consider relevant to motor insurance risk, yet public actuarial datasets provide limited location identifiers, constraining how this information can be incorporated and evaluated in claim-frequency models. This study examines how geographic information from alternative data sources can be incorporated into actuarial models for Motor Third Party Liability (MTPL) claim prediction under such constraints.
Using the BeMTPL97 dataset, we adopt a zone-level modeling framework and evaluate predictive performance on unseen postcodes. Geographic information is introduced through two channels: environmental indicators from OpenStreetMap and CORINE Land Cover, and orthoimagery released by the Belgian National Geographic Institute for academic use. We evaluate the predictive contribution of coordinates, environmental features, and image embeddings across three baseline models: generalized linear models (GLMs), regularized GLMs, and gradient-boosted trees, while raw imagery is modeled using convolutional neural networks.
Our results show that augmenting actuarial variables with constructed geographic information improves accuracy. Across experiments, both linear and tree-based models benefit most from combining coordinates with environmental features extracted at 5 km scale, while smaller neighborhoods also improve baseline specifications. Generally, image embeddings do not improve performance when environmental features are available; however, when such features are absent, pretrained vision-transformer embeddings enhance accuracy and stability for regularized GLMs. Our results show that the predictive value of geographic information in zone-level MTPL frequency models depends less on model complexity than on how geography is represented, and illustrate that geographic context can be incorporated despite limited individual-level spatial information.2026-04-23T17:44:52Z35 pages, 8 figuresSherly Alfonso-SánchezCristián BravoKristina G. Stankovahttp://arxiv.org/abs/2604.21734v1Modeling dependency between operational risk losses and macroeconomic variables using Hidden Markov Models2026-04-23T14:38:51ZPredicting future operational risk losses gives rise to a significant challenge due to the heterogeneous and time-dependent structures present in real-world data. Furthermore, stress test exercises require examining the relationship with operational losses. To capture such relationship, we propose to use an extension of Hidden Markov Models to multivariate observations. This model introduces a third auxiliary variable designed to accommodate the economic covariates in the time-series data. We detail the unique aspects of operational risk data and describe how model calibration is achieved via the Expectation-Maximization (EM) algorithm. Additionally, we provide the calibration results for the various risk-event types and analyze the relevance of the inclusion of the macroeconomic covariates.2026-04-23T14:38:51ZNikeethan SelvaratnamDorinel BastideClément FernandesWojciech Pieczynskihttp://arxiv.org/abs/2604.21297v1Identifying dynamical network markers of financial market instability2026-04-23T05:25:35ZMarket instability has been extensively studied using mathematical approaches to characterize complex trading dynamics and detect structural change points. This study explores the potential for early warning of market instability by applying the Dynamical Network Marker (DNM) theory to order placement and execution data from the Tokyo Stock Exchange. DNM theory identifies indicators associated with critical slowing down -- a precursor to critical transitions -- in high-dimensional systems of many interacting elements. In this study, market participants are identified using virtual server IDs from the trading system, and multivariate time series representing their trading activities are constructed. This framework treats each participant as an interacting element, thereby enabling the application of DNM theory to the resulting time series. The results suggest that early warning signals of large price movements can be detected on a daily time scale. These findings highlight the potential to develop practical DNM-based early-warning systems for large price movements by further refining forecasting horizons and integrating multiple time series capturing different aspects of trading behavior.2026-04-23T05:25:35Z94 pages (33 pages main text + 61 pages Supplementary Information)Mariko I. ItoHiroyuki HasadaYudai HonmaTakaaki OhnishiTsutomu WatanabeKazuyuki Aiharahttp://arxiv.org/abs/2604.02832v2Transfer Learning for Loan Recovery Prediction under Distribution Shifts with Heterogeneous Feature Spaces2026-04-23T04:48:29ZAccurate forecasting of recovery rates (RR) is central to credit risk management and regulatory capital determination. In many loan portfolios, however, RR modeling is constrained by data scarcity arising from infrequent default events. Transfer learning (TL) offers a promising avenue to mitigate this challenge by exploiting information from related but richer source domains, yet its effectiveness critically depends on the presence and strength of distributional shifts, and on potential heterogeneity between source and target feature spaces.
This paper introduces FT-MDN-Transformer, a mixture-density tabular Transformer architecture specifically designed for TL in RR forecasting across heterogeneous feature sets. The model produces both loan-level point estimates and portfolio-level predictive distributions, thereby supporting a wide range of practical RR forecasting applications. We evaluate the proposed approach in a controlled Monte Carlo simulation that facilitates systematic variation of covariate, conditional, and label shifts, as well as in a real-world transfer setting using the Global Credit Data (GCD) loan dataset as source and a novel bonds dataset as target.
Our results show that FT-MDN-Transformer outperforms baseline models when target-domain data are limited, with particularly pronounced gains under covariate and conditional shifts, while label shift remains challenging. We also observe its probabilistic forecasts to closely track empirical recovery distributions, providing richer information than conventional point-prediction metrics alone. Overall, the findings highlight the potential of distribution-aware TL architectures to improve RR forecasting in data-scarce credit portfolios and offer practical insights for risk managers operating under heterogeneous data environments.2026-04-03T07:54:49Z35 pages, 14 figures. Christopher Gerling had previously withdrawn his submission due to NDA restrictions, and that matter was resolved. We are authorized to publish the preprint nowChristopher GerlingHanqiu PengYing ChenStefan Lessmannhttp://arxiv.org/abs/2604.20406v1Bond Market Making with a Hit-Ratio Target2026-04-22T10:19:21ZWe study OTC bond market making on a size ladder with quadratic inventory penalty and a running target on the dealer's size-weighted hit ratio within a stochastic optimal control approach. We demonstrate that the corresponding reduced Hamilton-Jacobi-Bellman (HJB) equation remains separable by dualizing the hit ratio target term and provides the exact optimal controls through the inverse of the fill-probability function and the Hamiltonian derivative. We then focus on the quadratic approximation á la Bergault et al., which yields a Riccati equation for the inventory curvature while retaining the exact quote map. In its linearized form, this approximation produces explicit quote decompositions into riskless spread, inventory-risk correction, and hit-ratio correction. The formulation is general and applies to multi-bond, multi-client-tier scenarios, with special cases obtained by restricting the targeted tiers, their bond coverage, and their associated targets.2026-04-22T10:19:21Z16 pages, 9 figuresAlexander BarzykinAxel Cicerihttp://arxiv.org/abs/2605.06678v1A Wasserstein GAN-based climate scenario generator for risk management and insurance: the case of soil subsidence2026-04-22T08:30:53ZAccording to the United Nations Office for Disaster Risk Reduction (2025), the average annual cost of natural catastrophes increased from 70--80 billion USD between 1970 and 2000 to 180--200 billion USD between 2001 and 2020. Reports from organizations such as the IFOA and the WWF highlight the need for the insurance sector to adapt to this rapidly evolving context by developing medium- to long-term strategies that go beyond the one-year horizon of prudential regulations such as Solvency II. This paper introduces an artificial intelligence framework based on Conditional Generative Adversarial Networks (Conditional GANs) to generate future spatio-temporal trajectories of climatic indices. The approach focuses on the Soil Wetness Index (SWI), a key indicator used in France to assess drought severity. Drought accounts for approximately 30% of the indemnities paid under the French natural catastrophe insurance scheme. The proposed model, SwiGAN, simulates plausible drought propagation patterns up to 2050 for a region of France particularly exposed to this hazard. By generating realistic sequences of SWI maps, SwiGAN provides insights into drought dynamics under climate change scenarios and supports the design of adaptive risk management and insurance strategies. The methodology is also generalizable to other climate-related perils and actuarial applications such as economic scenario generation.2026-04-22T08:30:53ZAntoine HeranvalBioSPOlivier LopezCRESTDidier NgatchaCRESTDaniel NkameniCRESThttp://arxiv.org/abs/2603.17463v2Multivariate GARCH and portfolio variance prediction: A forecast reconciliation perspective2026-04-21T17:22:33ZWe assess the advantage of combining univariate and multivariate portfolio risk forecasts with the aid of forecast reconciliation techniques. In our analyzes, we assume knowledge of portfolio weights, a standard for portfolio risk management applications. With an extensive simulation experiment, we show that, if the true covariance is known, forecast reconciliation improves over a standard multivariate approach, in particular when the adopted multivariate model is misspecified. However, if noisy proxies are used, correctly specified models and the misspecified ones (for instance, neglecting spillovers) turn out to be, in several cases, indistinguishable, with forecast reconciliation still providing improvements. The noise in the covariance proxy plays a crucial role in determining the improvement of both the forecast reconciliation and the correct model specification. An empirical analysis shows how forecast reconciliation can be adopted with real data to improve traditional GARCH-based portfolio variance forecasts.2026-03-18T08:10:41ZMassimiliano CaporinDaniele GirolimettoEmanuele Lopetusohttp://arxiv.org/abs/2605.00862v1Replication-Consistent Liquidity Forecasting for Derivatives -- Forward Funding Sensitivities and a Liquidity Valuation Adjustment for Settlement Lags2026-04-21T14:43:28ZWe study cash-flow forecasting for derivatives used in liquidity management and clarify its relation to risk-neutral valuation and replication. While it is well known that expectations under different measures (e.g., $\mathbb{P}$ vs. $\mathbb{Q}$) can yield different undiscounted cash-flows, further inconsistencies arise when payment times are stochastic. We show that using discounting sensitivities (funding-curve hedge ratios) instead of "expected cash-flows" aligns forecasting with the self-financing replication strategy and avoids measure-mixing/aggregation issues. We then illustrate how a standard valuation model delivers pathwise funding requirements and propose a simple liquidity valuation adjustment to capture settlement lags and related timing frictions. The note provides implementation hints (American Monte Carlo with adjoint differentiation) and clarifies when "expected cash-flows" are informative and when sensitivities should be used instead.2026-04-21T14:43:28Z34 pagesChristian P. Fries