https://arxiv.org/api/GOeNMwcMj7cuf0bhco3RPzUemss 2026-07-22T19:05:10Z 2985 75 15 http://arxiv.org/abs/2111.14631v2 Model Risk in Credit Portfolio Models 2026-06-16T08:16:51Z Model risk in credit portfolio models is a serious issue for banks but has so far not been tackled comprehensively. We will demonstrate how to deal with uncertainty in all model parameters in an all-embracing, yet easy-to-implement way. 2021-11-23T13:12:47Z 12 pages, 2 figures. This version: minor corrections, updates, and comments Christian Meyer http://arxiv.org/abs/2412.00607v4 On a risk model with tree-structured Poisson Markov random field frequency, with application to rainfall events 2026-06-16T03:49:28Z In many insurance contexts, dependence between risks of a portfolio may arise from their frequencies. We investigate a dependent risk model in which we assume the vector of count variables to be a tree-structured Markov random field with Poisson marginals. The tree structure translates into a wide variety of dependence schemes. We study the global risk of the portfolio and the risk allocation to all its constituents. We provide asymptotic results for portfolios defined on infinitely growing trees. To illustrate its flexibility and computational scalability to higher dimensions, we calibrate the risk model on real-world extreme rainfall data and perform a risk analysis. 2024-11-30T22:53:37Z 40 pages Hélène Cossette Benjamin Côté Alexandre Dubeau Etienne Marceau http://arxiv.org/abs/2606.17383v1 Model Validation of Agentic AI Systems: A POMDP-Based Framework for Belief-State, Forecast, and Policy Validation 2026-06-16T00:40:55Z Agentic artificial intelligence systems introduce a new class of model risk. Unlike traditional predictive models, autonomous agents continuously acquire information, form beliefs regarding latent states of the environment, generate forecasts, select actions, and adapt their behavior over time. Existing validation methodologies focus primarily on predictive accuracy and therefore provide limited insight into the quality of the underlying decision process. This paper proposes a model validation framework for agentic AI based on Partially Observable Markov Decision Processes (POMDPs). The framework decomposes autonomous decision making into information, beliefs, forecasts, actions, and utility, allowing each component to be validated independently. Large language models (LLMs) are formalized as approximate Bayesian filtering operators, and a model-risk taxonomy is developed encompassing state-space, filtering, forecast, policy, utility-specification, and parameter risks. The model risk validation methodology is demonstrated through a portfolio-management case study in which an agent infers latent market regimes from market and macroeconomic information, generates belief-conditioned forecasts, and constructs portfolios using a Black--Litterman framework. Empirical validation combines performance analysis, belief calibration diagnostics, coverage tests, ablation studies, and parameter-sensitivity analysis. The results indicate that latent-state inference contributes independently to decision quality and that the principal conclusions remain robust across a broad range of parameter values. The principal contribution of the paper is a practical framework for extending established model risk management concepts to autonomous AI systems and providing a rigorous foundation for their validation, governance, and monitoring. 2026-06-16T00:40:55Z 28 pages, 3 figures, 6 tables. Source code available from https://github.com/mfrdixon/agentic-AI-as-POMDP Matthew Francis Dixon http://arxiv.org/abs/2407.06619v2 CAESar: Conditional Autoregressive Expected Shortfall 2026-06-15T02:51:05Z In financial risk management, Value at Risk (VaR) estimates potential portfolio losses but fails to account for losses beyond a certain threshold. Expected Shortfall (ES) addresses this limitation by providing the conditional expectation of such exceedances, providing a better measure of tail risk. However, ES is not elicitable on its own, meaning that it cannot be estimated by minimizing some scoring function, although its joint elicitability with VaR allows for combined estimation. Building on this property, we propose the Conditional Autoregressive Expected Shortfall (CAESar) model, which flexibly handles dynamic patterns and heteroskedasticity, without making distributional assumptions on price returns. The optimization of CAESar coefficients involves three steps: fitting the VaR component via CAViaR regression, formulating ES as an autoregressive process, and jointly estimating VaR and ES coefficients while ensuring a monotonicity constraint to avoid crossing quantiles. Through extensive backtesting, CAESar outperforms existing methods, proving highly effective for risk forecasting. 2024-07-09T07:48:38Z Federico Gatta Fabrizio Lillo Piero Mazzarisi http://arxiv.org/abs/2606.15755v1 A Multiplex Network Hawkes Model for Systemic Risk Measurement 2026-06-14T11:32:20Z We introduce the Multiplex Network Hawkes model, which extends the network Hawkes framework of Linderman & Adams (2014) by allowing multiple excitation layers whose weights depend on observed edge and node covariates. We use the model to investigate how contagion in financial networks is affected by different transmission channels. The multiplex structure separates channel-specific contributions within a single inferred transmission network, allowing candidate propagation mechanisms to be compared directly rather than being absorbed into one homogeneous excitation layer. Covariate-dependent excitation allows us to investigate sources of transmission. We make posterior inference about the inferred directed network and its excitation dynamics using an MCMC sampler. The application uses a broad cross-industry credit default swap (CDS) dataset of 99 North American and European firms, including banks, insurers and non-financial firms over 2004-2022. We evaluate three candidate contagion channels associated with asset similarity, solvency and profitability. The results indicate sparse contagion pathways, with systemic-risk transmission concentrated in outward flows from a small number of influential institutions rather than in mutual feedback between institutions. The channel results show that industry similarity is the most consistently supported asset-similarity effect, while aggregate layer contributions indicate that asset-similarity, solvency and profitability channels all contribute to inferred excitation. 2026-06-14T11:32:20Z Mante Zelvyte Jim E. Griffin http://arxiv.org/abs/2606.15473v1 Belief at Risk: Quantifying Agentic AI Model Risk with LLM-Inferred Bayesian State Filters 2026-06-13T21:08:59Z Agentic AI systems create model risk because uncertain beliefs are coupled to autonomous actions. This paper develops a mathematical framework for quantifying agentic AI risk by representing the system as a partially observed Markov decision process with latent states, Bayesian belief updates, control-dependent losses, and tail-risk functionals. The main methodological contribution is to treat a large language model as an uncertain semantic observation model: the LLM maps high-dimensional evidence into a probability vector over latent regimes, while a Bayesian filter imposes temporal coherence and produces auditable posterior beliefs. The resulting framework separates uncertainty quantification from risk measurement. Uncertainty is represented by posterior entropy, belief drift, and calibration error; risk is represented by the distribution of losses induced by decisions taken under those beliefs. The paper connects this construction to model risk management, coherent risk measures, Bayesian filtering, POMDP theory, robust control, and quantitative portfolio risk. An empirical case study using adjusted daily equity returns from Massive.com illustrates how LLM-inferred belief states can be combined with Bayesian filtering to produce regime probabilities, uncertainty diagnostics, calibration statistics, and VaR/CVaR-style risk measures. The framework is intended as a rigorous foundation for validating agentic AI in financial and other regulated decision environments. 2026-06-13T21:08:59Z 15 pages, 3 figures Matthew Francis Dixon http://arxiv.org/abs/2606.15452v1 PHINN: Persistent Homology Inspired Neural Network for Rare-Event Time Series Generation 2026-06-13T19:56:57Z Rare events in time series are critical to model but hard to learn due to data scarcity. Current generative models struggle with extreme values. We observe that rare events leave distinct topological fingerprints - transitions in Betti numbers from point-cloud embeddings - that are more stable and discriminative than statistical moments. We introduce PHINN, a flow-matching framework using dynamic Betti curves as conditioning signals and a persistence landscape loss for homology consistency. It scales to multivariate data, includes a natural-language interface to set Betti targets, supports cross-domain meta-learning and few-shot generation, and provides certified adversarial robustness. On financial, epidemiological, and multi-modal benchmarks, PHINN outperforms statistical and diffusion baselines in topological fidelity (beta-RMSE down 41-63%, transition accuracy up 84%) and matches jump-diffusion models in tail coverage while exceeding them in shape fidelity. All results have 95% confidence intervals. 2026-06-13T19:56:57Z 15 pages, 4 figures Emre Yusuf Ren Takahashi Jayabrata Bhaduri http://arxiv.org/abs/2508.20225v5 Optimal Quoting under Adverse Selection and Price Reading 2026-06-13T15:50:03Z Over the past decade, many dealers have implemented algorithmic models to automatically respond to RFQs and manage flows originating from their electronic platforms. In parallel, building on the foundational work of Ho and Stoll, and later Avellaneda and Stoikov, the academic literature on market making has expanded to address trade size distributions, client tiering, complex price dynamics, alpha signals, and the internalization versus externalization dilemma in markets with dealer-to-client and interdealer-broker segments. In this paper, we tackle two critical dimensions: adverse selection, arising from the presence of informed traders, and price reading, whereby the market maker's own quotes inadvertently reveal the direction of their inventory. These risks are well known to practitioners, who routinely face informed flows and algorithms capable of extracting signals from quoting behaviour. Yet they have received limited attention in the quantitative finance literature, beyond stylized toy models with limited actionability. Extending the existing literature, we propose a tractable framework that enables market makers to adjust their quotes with greater awareness of informational risk. 2025-08-27T19:04:52Z Alexander Barzykin Philippe Bergault Olivier Guéant Malo Lemmel http://arxiv.org/abs/2504.11775v3 Discrimination-free Insurance Pricing with Privatized Sensitive Attributes 2026-06-12T19:13:00Z Fairness has become an important concern in insurance pricing as insurers increasingly rely on machine learning models to predict expected losses. At the same time, regulatory and privacy constraints often restrict insurers' ability to access or use sensitive attributes such as gender or race. Recent actuarial research addresses fairness in this context through the concept of the discrimination-free premium, which removes both the direct and indirect effects of sensitive attributes while preserving actuarial consistency. However, implementing this approach typically requires access to the sensitive attributes themselves, which may not be available in practice. This paper studies the estimation of discrimination-free insurance premiums when sensitive attributes are observed only in privatized or noise-perturbed form. We consider a multi-party data setting in which insurers observe non-sensitive attributes and outcomes, while a trusted third party holds privatized sensitive attributes generated through a privacy mechanism. Within this framework, we develop statistical methods for estimating discrimination-free premiums using only the privatized attributes. We study two settings of practical relevance: when the privacy mechanism is known and when its noise level is unknown. For both cases, we establish theoretical guarantees for the proposed estimators. Numerical experiments and empirical applications demonstrate that the proposed approach enables fair insurance pricing while respecting privacy and regulatory constraints. 2025-04-16T05:29:11Z Tianhe Zhang Suhan Liu Peng Shi http://arxiv.org/abs/2606.14830v1 Pricing Excess-of-Loss Reinsurance and CAT Bonds under Climate Uncertainty: A Cox Process Framework with Temperature-Dependent Stochastic Intensity 2026-06-12T14:52:13Z This paper develops a climate-aware pricing framework for excess-of-loss (XL) reinsurance contracts and catastrophe (CAT) bonds under non-stationary catastrophe risk. Catastrophe arrivals are modeled as a Cox process whose stochastic intensity depends exponentially on a temperature-related climate index. To represent climate dynamics, the index is modeled as a mean-reverting Ornstein--Uhlenbeck process around a time-dependent warming trend. Within this setting, aggregate losses follow a compound Cox structure with lognormal severities. Pricing is performed under a reduced-form risk-adjusted measure, which provides a tractable valuation approach for XL reinsurance layers and binary zero-coupon CAT bond payoffs in an incomplete market setting. Because catastrophe losses are not dynamically replicable, the framework emphasizes scenario-based valuation rather than model-independent no-arbitrage bounds. A Monte Carlo valuation scheme is implemented to quantify the economic implications of climate-dependent catastrophe intensity. The numerical results show that climate dependence materially changes the loss-generation mechanism and affects the valuation of catastrophe-linked contracts. In the baseline calibration, the climate-aware model increases the excess-of-loss reinsurance premium and lowers the CAT bond price relative to the stationary benchmark. Furthermore, our analysis of the 99.5\% Tail Value-at-Risk (TVaR) indicates that stationary benchmarks may underestimate economic capital requirements by approximately 13.7\% compared to the climate-aware framework, highlighting the potential regulatory relevance of the proposed model. This finding highlights that benchmark design is critical for interpreting climate-pricing effects. 2026-06-12T14:52:13Z Nader Karimi Foad Shokrollahi http://arxiv.org/abs/2606.14484v1 Quantum Horizon: An evaluation of quantum computing as a threat to Bitcoin and Ethereum 2026-06-12T14:21:58Z Quantum computing poses a real, broad-based, but bounded and substantially mitigable threat to Bitcoin and Ethereum. We separate the two quantum algorithms that public discussion routinely conflates: Shor's algorithm breaks the elliptic-curve signatures (ECDSA over secp256k1, BLS over BLS12-381) that authorize spending, whereas Grover's algorithm does not meaningfully threaten proof-of-work mining, which is protected by a merely quadratic speedup, fault-tolerant per-operation costs, a square-root parallelization wall, and difficulty adjustment. Folding hardware scaling, the falling resource requirement, a fault-tolerance readiness lag, and expert surveys into a single Monte-Carlo forecast yields a wide, bimodal arrival distribution for a cryptographically relevant quantum computer: about a one-in-six chance by 2035, near 30% by 2040, and about 60% by 2050. Exposure is concentrated and mostly migratable: of Bitcoin's roughly six million quantum-exposed coins only about 2.3 million are irreducibly at risk, while 50 to 65% of Ether sits at key-revealed accounts that can adopt post-quantum signatures. A timely migration beats even an optimistic 2035 machine, so the binding constraint is governance, not technology. A survey of the top twenty cryptocurrencies finds none fully post-quantum. Reproducible models accompany every quantitative claim. 2026-06-12T14:21:58Z 21 pages, 5 figures, 3 tables. Reproducible model code, data, and figures: https://github.com/imgcode/quantum-horizon Iosif M. Gershteyn Jacob A. Alber http://arxiv.org/abs/2605.18784v2 The Insurability Frontier of AI Risk: Mapping Threats to Affirmative Coverage, Silent Exposures, and Exclusions 2026-06-12T11:34:44Z The rapid diffusion of agentic AI has created a new coverage problem for commercial insurance: some AI-mediated losses are now affirmatively insured, some create silent-AI exposure under legacy cyber, technology errors-and-omissions (E&O), directors-and-officers (D&O), employment practices liability (EPLI), crime, and media policies, and others are being actively excluded. This paper maps that emerging boundary by coding 55 AI threat classes against 26 insurance products, endorsements, and exclusion regimes using public carrier materials and OWASP/MITRE threat catalogs. We identify a four-tier insurability frontier: affirmatively insured perils, silent-AI exposures, actively excluded perils, and perils outside conventional private insurance structures. Our coding measures publicly claimed positioning rather than executed contract wording; the headline statistics describe what carriers publicly state about coverage, not what would be paid in any specific claim. Three patterns emerge. First, affirmative AI coverage is beginning to differentiate by primary risk emphasis: public materials often position Munich Re around model performance and drift, Armilla and parts of the Lloyd's market around hallucination and broader AI liability, Tokio Marine Kiln and CFC around IP and technology E&O concerns, Apollo ibott around emerging autonomous system liability, and Coalition around deepfake and AI-enabled cyber response. Second, legacy lines retain silent-AI exposure where AI is an instrumentality rather than the legal cause of loss. Third, foundation model concentration is the clearest genuinely novel insurability frontier because upstream model failure can correlate losses across many cedents at once; the relevant market design question is which insurability constraint each candidate structure relaxes, not merely which systemic risk template exists. 2026-05-06T13:24:44Z Version 2 Alex Leung Rex Zhang Ervin Ling Kentaroh Toyoda SiewMei Loh http://arxiv.org/abs/2606.14050v1 Battery Bidding under Price Uncertainty in Wholesale Electricity Markets 2026-06-12T02:51:55Z Grid-scale batteries increasingly influence outcomes in wholesale electricity markets, but their observed bid patterns remain difficult to interpret. In particular, bids that appear to reflect strategic withholding may instead arise from rational operations under price uncertainty and risk management. We develop an asset-level model of a price-taking battery that submits stepwise buy and sell bid curves in the day-ahead market under a finite set of price scenarios. The battery chooses quantity--price pairs to maximize a mean--CVaR objective subject to physical and market constraints. A direct formulation is a mixed-integer linear program, but we show that its integer decisions can be removed, yielding an exact linear programming reformulation suitable for empirical analysis. Our empirical results deliver three insights. First, withholding behavior can arise even without market power, because scarce stored energy and uncertain future prices increase the value of holding energy. Second, the effect of uncertainty depends on the state of charge: when stored energy is scarce, greater uncertainty raises sell bid prices, whereas when stored energy is abundant it can lower them. Third, risk management reshapes bid curves into layered structures that secure profitable execution across a broad set of scenarios while preserving some exposure to rare but valuable price spikes. 2026-06-12T02:51:55Z Vincent Yinjun-Wang Madeleine Udell http://arxiv.org/abs/2512.21973v6 When Indemnity Insurance Fails: Parametric Coverage under Binding Budget and Risk Constraints 2026-06-12T01:36:53Z In high-risk environments, traditional indemnity insurance is often unaffordable or ineffective, despite its well-known optimality under expected utility. We compare excess-of-loss indemnity insurance with parametric insurance within a common mean-variance framework, allowing for fixed costs, heterogeneous premium loadings, and binding budget constraints. Motivated by the disaster insurance and risk-sharing literature, we show that, once these realistic frictions are introduced, parametric insurance can yield higher welfare for risk-averse individuals, even under the same utility objective and without relying on behavioral assumptions. The welfare advantage arises precisely when indemnity insurance becomes impractical (particularly when households face binding premium budgets), and disappears once both contracts are unconstrained. Our results help reconcile classical insurance theory with the growing use of parametric risk transfer in high-risk settings, and rationalize the interest in hybrid designs that combine both indemnity and parametric elements. 2025-12-26T10:37:32Z Benjamin Avanzi Debbie Kusch Falden Mogens Steffensen http://arxiv.org/abs/2606.13880v1 A Longitudinal Attribute-Conditioned Neural Network for Modeling Health-State Transition Probabilities in Temporally Irregular Data: The LANTERN Framework 2026-06-11T20:15:14Z Accurate estimation of long-term care transition probabilities is central to disability insurance pricing, reserving, and solvency assessment. Classical actuarial multi-state models commonly rely on Markov, semi-Markov, or proportional-hazard specifications, which provide a direct connection to cohort projection but may be restrictive for irregular longitudinal health data with nonlinear aging patterns and heterogeneous covariate histories. This paper develops a well-calibrated estimator of multi-state transition probabilities for irregular longitudinal health data. The model learns from individual health history, incorporates the time elapsed between observations, and conditions transition probabilities on demographic and socioeconomic attributes. It produces a valid probability distribution over the next observed health state, with four possible states: healthy, mild disability, severe disability, and death. Individual probabilities are aggregated by age group and origin state to form transition matrices compatible with actuarial cohort projection. Using longitudinal data from the Health and Retirement Study, we compare the proposed estimator with logistic regression, gradient-boosted trees, a recurrent neural network, and a last-state persistence benchmark. The evaluation considers probabilistic accuracy, endpoint discrimination and calibration for severe disability and death, risk concentration, and transition matrix error after aggregation. The proposed estimator improves severe disability discrimination relative to logistic regression and gradient-boosted tree benchmarks, maintains strong calibration, and yields the lowest transition matrix error among the evaluated models in the held-out test analysis. Results show that a structured machine learning estimator can support long-term care transition modeling when judged by calibration and projection fidelity, beyond discrimination. 2026-06-11T20:15:14Z 35 pages, 17 figures Bright Kwaku Manu Beckett Sterner Petar Jevtic