https://arxiv.org/api/5hOuerTdRHMuQn6fziunW9RhDIk 2026-07-21T14:57:25Z 5630 75 15 http://arxiv.org/abs/2607.11920v1 Sensitivity to Subjective Expected Utility Maximization: A Methodological Study, with an Illustrative Application to LLM Decision-Making 2026-07-08T00:29:18Z Evaluating decisions made under uncertainty is hard when labeled outcomes are scarce, costly, or confounded with luck. We treat subjective expected utility (SEU) maximization as a stated standard and define a graded measure -- SEU sensitivity -- of an agent's conformity to it. The vehicle is a softmax choice model with a sensitivity parameter $α$ on SEU-valued alternatives; the contribution is a sequence of identifiability results for $α$ and for belief and utility parameters $(β, δ)$, validated in Stan via prior predictive checks, parameter recovery, and simulation-based calibration (SBC), with finite-sample caveats intact. In the uncertain-choice-only model $m_0$, $α$ is identifiable given the expected-utility vector $η$ and sharply recovered, while $(β, δ)$ are only weakly informed: the posterior barely contracts and concentrates on a $β$-$δ$ trade-off. In the extended model $m_1$, $δ$ becomes identifiable in principle via a $β$-free risky block, but its practical recovery gain at realistic sample sizes is negligible (matched-count CI-width reduction under 1%), and that block yields no detected $α$-precision gain at matched choice count. These are two distinct phenomena: for $δ$, identifiability does not imply precise estimability at realistic $n$; for $α$, identifiability is silent about what governs finite-$n$ precision. Marginal SBC passes for both models even where the joint posterior is weakly informed -- a demarcation we make precise. A two-by-two application (GPT-4o and Claude 3.5 Sonnet, each on insurance-claims triage and Ellsberg-style urns, with sampling temperature as the lever) runs end-to-end on real LLM choice data, detecting a structured comparative $α$ effect in two of four cells. 2026-07-08T00:29:18Z 64 pages, 5 figures. Code and data archived at Zenodo: https://doi.org/10.5281/zenodo.21250951 Jeff Helzner http://arxiv.org/abs/2603.24705v3 Amortized Inference for Correlated Discrete Choice Models via Equivariant Neural Networks 2026-07-07T18:36:27Z Discrete choice models are fundamental tools in management science, economics, and marketing for understanding and predicting decision-making. Logit-based models are dominant in applied work, largely due to their convenient closed-form expressions for choice probabilities. However, these models entail restrictive assumptions on the stochastic utility component, constraining our ability to capture realistic and theoretically grounded choice behavior$-$most notably, substitution patterns. In this work, we propose an amortized inference approach using a neural network emulator to approximate choice probabilities for general error distributions, including those with correlated errors. Our proposal includes a specialized neural network architecture and accompanying training procedures designed to respect the invariance properties of discrete choice models. We provide group-theoretic foundations for the architecture, including a proof of universal approximation given a minimal set of invariant features. Once trained, the emulator enables rapid likelihood evaluation and gradient computation. We use Sobolev training, augmenting the likelihood loss with a gradient-matching penalty so that the emulator learns both choice probabilities and their derivatives. We show that emulator-based maximum likelihood estimators are consistent and asymptotically normal under mild approximation conditions, and we provide sandwich standard errors that remain valid even with imperfect likelihood approximation. Simulations show significant gains over the GHK simulator in accuracy and speed. 2026-03-25T18:30:11Z Easton Huch Michael Keane http://arxiv.org/abs/2607.06412v1 A Machine-Learning-Compatible Omnibus Test for Treatment Effect Heterogeneity 2026-07-07T15:42:31Z This study proposes a formal, computationally efficient nonparametric omnibus test for treatment-effect heterogeneity that is compatible with a broad class of estimators, including modern machine-learning methods. The test is designed for settings in which identification can rely on high-dimensional controls while heterogeneity is assessed with respect to a low-dimensional subset of covariates. We derive the test statistic's asymptotic null distribution and develop a bootstrap procedure that is efficient because it avoids re-estimating nuisance parameters in each iteration. The testing approach applies to multiple empirical designs, including randomized experiments, selection-on-observables, difference-in-differences, and instrumental-variables settings. Monte Carlo simulations show that the test attains near-nominal size under the null and exhibits good power against heterogeneous alternatives. We further illustrate the procedure using two empirical applications on retirement savings and trade liberalization. 2026-07-07T15:42:31Z 42 pages, 5 figures, 2 tables Elia Lapenta Anthony Strittmatter Pedro Vergara Merino http://arxiv.org/abs/2607.06368v1 Factor-Augmented Machine Learning Panel Regressions 2026-07-07T15:06:00Z This paper develops the asymptotic theory for high-dimensional panel data regressions in settings with cross-sectionally dependent errors driven by common shocks. We consider a factor-augmented sparse-group LASSO estimator that combines MIDAS aggregation with latent factors. The estimator can take advantage of the mixed-frequency group structure in the time-series dimension. Theory shows that it can outperform the standard LASSO estimator both for prediction and estimation while allowing for cross-sectional dependence. 2026-07-07T15:06:00Z Andrii Babii Luca Barbaglia Eric Ghysels Jonas Striaukas http://arxiv.org/abs/2607.04743v2 Stabilized Higher-Order Influence Functions: Statistical Theory of a Class of Bilinear Forms 2026-07-07T14:49:14Z Higher-order influence functions, introduced in a series of articles (Robins et al., 2008, 2009a; van der Vaart, 2014; Robins et al., 2016, 2023; Liu et al., 2017), are a unified framework for constructing rate-optimal point estimates of a class of statistical functionals under various complexity-reducing assumptions on the posited statistical model that generates the observed data. Although higher-order (influence functions) estimators are theoretically appealing, they have very limited practical uptake compared to their first-order counterparts. The original higher-order estimators proposed in Robins et al. (2008) and Robins et al. (2017) involve nonparametric density estimation of multi-dimensional covariates, a highly nontrivial statistical and computational problem on its own. The density estimator is, in turn, used in the evaluation of the inverse population Gram matrix $Ω$ of a set of $k$-dimensional basis transformations of covariates. There, $k$ is allowed to be as large as $o (n^2)$. To partially address this potential shortcoming, Liu et al. (2017) restrict $k$ to $o (n)$ and instead estimate $Ω$ directly using the inverse sample Gram matrix estimator, but computed from an independent sample often obtained by sample-splitting. Liu et al. (2017) refer to this alternative estimator as the empirical higher-order estimator. Although the empirical higher-order estimator bypasses density estimation, it suffers from numerical instability due to inverting a large-dimensional sample Gram matrix. In this article, for a class of bilinear forms/functionals that often appear in substantive fields, we propose a new stabilized higher-order estimator without sample splitting, which exhibits more stable finite-sample performance compared to the empirical higher-order estimator. We also prove that this new class of higher-order estimators enjoys similar statistical guarantees. 2026-07-06T07:27:01Z This paper improves and supersedes a previous draft by the second and last authors: arXiv:2302.08097 Na Liu Chang Li Yujia Gu Lin Liu http://arxiv.org/abs/2102.04048v4 On global identification in structural vector autoregressions 2026-07-07T11:54:22Z In a landmark contribution to the structural vector autoregression (SVARs) literature, Rubio-Ramirez, Waggoner, and Zha (2010, `Structural Vector Autoregressions: Theory of Identification and Algorithms for Inference,' Review of Economic Studies) shows a necessary and sufficient condition for equality restrictions to globally identify the structural parameters of a SVAR. The simplest form of the necessary and sufficient condition shown in Theorem 7 of Rubio-Ramirez et al (2010) checks the number of zero restrictions and the ranks of particular matrices without requiring knowledge of the true value of the structural or reduced-form parameters. However, this note shows by counterexample that this condition is not sufficient for global identification. Analytical investigation of the counterexample clarifies why their sufficiency claim breaks down. The problem with the rank condition is that it allows for the possibility that restrictions are redundant, in the sense that one or more restrictions may be implied by other restrictions, in which case the implied restriction contains no identifying information. We derive a modified necessary and sufficient condition for SVAR global identification and clarify how it can be assessed in practice. 2021-02-08T08:14:27Z 16 pages, no figures Emanuele Bacchiocchi Toru Kitagawa http://arxiv.org/abs/2509.19911v2 Decomposing Co-Movements in Matrix-Valued Time Series: A Pseudo-Structural Reduced-Rank Approach 2026-07-07T08:46:12Z A pseudo-structural framework is proposed for analyzing contemporaneous co-movements in stationary reduced-rank matrix autoregressive (RRMAR) models. Unlike conventional vector autoregressive (VAR) models that discard the matrix structure, the formulation preserves it, enabling a decomposition of co-movements into three interpretable components: row-specific, column-specific, and joint (row--column) interactions across the matrix-valued time series. The estimator admits standard asymptotic inference and a BIC-type criterion is proposed for the joint selection of the reduced ranks and the autoregressive lag order. The method's finite-sample performance in terms of estimation accuracy, coverage, and rank selection is validated through simulation experiments, including cases of rank misspecification. Practical usefulness is illustrated through an application to labor market data from nine Midwestern U.S. states, revealing distinct row-, column-, and joint co-movement patterns. 2025-09-24T09:10:09Z Alain Hecq Ivan Ricardo Ines Wilms http://arxiv.org/abs/2607.05882v1 Revision Risk in Real-Time Macroeconomic Forecasting 2026-07-07T06:27:45Z Macroeconomic forecasts refer to outcomes that are first released and then revised. A 90 percent interval for the first GDP release, a six-month value, or a latest-value benchmark is not the same uncertainty statement. We ask how revision risk evolves through the release cycle and what can be reported in real time when later-outcome errors are scarce. We decompose later-outcome MSE into preliminary forecast risk, revision risk, and their covariance. In SPF data, first-release to roughly 180-day revisions account for 8.3 percent of later-outcome MSE across real-activity targets, versus 3.6 percent across inflation targets. We show that later-outcome uncertainty is partially identified: released histories give early-error and revision marginals, but not their dependence. This yields a sharp Frechet-Makarov set and motivates direct late calibration, dependence-robust transport, and signed or revision-model transport. Out-of-sample results support method choice rather than a universal transport rule: coverage and stability determine when transport gains are usable. 2026-07-07T06:27:45Z Yizhou Kyle Kuang http://arxiv.org/abs/2607.05878v1 Bolivia and an IMF Extended Fund Facility: Financial Sustainability, Verifiable Social Sustainability, and Net Financing Additionality 2026-07-07T06:19:53Z This paper evaluates stabilization scenarios for Bolivia under external-financing stress by jointly modeling financial sustainability, verifiable social sustainability and the net additionality of multilateral financing. The main correction is that financing already received, contracted or expected from multilateral institutions cannot be counted as a marginal benefit of a hypothetical IMF Extended Fund Facility. Already secured resources are moved to the baseline; only incremental, liquid, timely and non-displaced financing is rewarded. The analysis combines corrected scenario scoring, poverty and inequality measures, maternal-child mortality indicators, public investment multipliers, self-defeating consolidation diagnostics, Monte Carlo uncertainty, leave-one-criterion-out tests and adverse execution assumptions. The results do not support an orthodox IMF-first strategy or front-loaded fiscal consolidation. The strongest designs are concessional multilateral packages with verified net additionality, productive execution, social-health floors and protection against poverty and maternal-child deterioration. An IMF arrangement is valuable only if it adds resources or credibility that were not already secured. The conclusion is that sustainability must be evaluated as a joint financial and social condition: closing cash gaps by weakening growth, poverty, inequality or maternal-child outcomes can be macroeconomically self-defeating. 2026-07-07T06:19:53Z Ricardo Alonzo Fernández Salguero http://arxiv.org/abs/2607.05862v1 A Framework for Transportation and Land Use Integration as a Parallel Constrained Multiple Discrete-Continuous Extreme Value (PC-MDCEV) Home Production Model 2026-07-07T05:41:53Z Integrated urban models (IUM) typically rely on a measure of accessibility or travel time to form the link between the transportation and land use systems. Such integration does not fully capture the trade-offs made by households in how they spend their limited temporal and monetary budgets. We propose a microeconomic foundation for transportation and land use choice model integration based on the theory of home production. A utility function is developed that considers both household monetary expenditure and individual time use. We address several limitations in previous home production functions. First, the introduction of a parallel constrained multiple discrete-continuous extreme value (MDCEV) structure that allows for the inclusion of multi-person households in the model. Second, travel time is defined as the minimum time required to conduct an activity and deducted from the temporal budget. This assumption has several appealing features. It defines the minimum time to complete an activity as a measure of accessibility. An empirical application is provided for the Greater Toronto Area using a validated synthetic dataset. Empirical results demonstrate an economy of scale in time devoted to home production, analogous to scaling exhibit in market production. It was found that the mix of dwelling types (detached, apartment, etc.) has a significant influence on both time use and consumption. Finally, we provide several directions for future research to advance the practice of urban modelling and better capture the complex dynamics of household decision-making. 2026-07-07T05:41:53Z Transportmetrica B: Transport Dynamics (2024) Jason Hawkins Khandker Nurul Habib 10.1080/21680566.2024.2303042 http://arxiv.org/abs/2607.05792v1 Estimating Causal Effects from Data Generated by Stochastic Algorithms 2026-07-07T03:43:04Z Recommendation systems and chatbots present content to users, typically using stochastic algorithms that select the content based on user characteristics or context. Examples of content include chat responses, videos, or items available for purchase. Scientists and application developers are often interested in whether characteristics of content increase outcomes such as user engagement. Estimates of such causal effects may guide content providers to generate content that emphasize desirable features. However, in settings with a large content library or where content is generated uniquely for a given user, it can be difficult to use observational data to learn the causal effect of content features, because the content a user sees is tailored to that user, and because content varies in many dimensions. This paper proposes a new method for estimating the impact of content features using observational data, when the algorithm that determines user exposure incorporates some randomization, and when two additional data elements are logged for each user: $(i)$ the identity of at least one item that could have been exposed to the user, but was not (the unexposed item); $(ii)$ an estimate of the ratio of the probability that the unexposed item would have been shown to the probability that the exposed item was shown. We show that causal effects of features are identified in this setting, even in the presence of unobserved confounders that affect both user preferences and the identity of the considered pair of items (exposed and unexposed). Our estimator differs from prior approaches in terms of what data is used and how the estimator is constructed. 2026-07-07T03:43:04Z 49 pages Susan Athey Guido Imbens Zoe Ji http://arxiv.org/abs/2607.05699v1 Identification, Estimation and Inference Based on Structural Error Projection 2026-07-06T23:37:43Z This paper proposes to project and expand the conditional mean function of the structural error given the regressors in an endogenous regression under consideration. As the projection process is semiparametric, we define this procedure as a semiparametric projection (SP) method to address endogeneity in regression models by internally constructed instrumental variables. The SP method is applicable to many classes of regression models associated with endogeneity, such as linear, nonlinear, and non- and semi-parametric models, and provides a simple and computationally tractable alternative to conventional instrumental variable approaches available from the existing literature. This paper establishes identification conditions and derives the asymptotic properties of the resulting estimators. It then proposes a simple LASSO selection method to examine the finite-sample performance of both the proposed method and the established theory by simulated and real data examples. 2026-07-06T23:37:43Z Chaohua Dong Jiti Gao Oliver Linton Bin Peng http://arxiv.org/abs/2407.21119v5 Potential weights and implicit causal designs in linear regression 2026-07-06T20:11:03Z Applied researchers routinely use linear regression to estimate causal effects, justified by quasi-experimental treatment variation, while leaving assumptions on treatment assignment implicit. We formalize a minimal criterion for quasi-experimental interpretation -- that the regression estimates some contrast of potential outcomes under the true assignment process, regardless of potential outcomes -- and characterize its implications for arbitrary regressions. This criterion implies linear restrictions on the true treatment distribution, whose solutions we call implicit designs. A regression is exactly quasi-experimental if and only if the true design is an implicit design, and approximately so when it is close to one, in a sense we formalize. Our framework unifies existing results and uncovers new ones across many settings. Qualitatively, an AI-assisted census of 1,051 recent papers finds quasi-experimental regression pervasive and often vulnerable to our negative results. Quantitatively, we assess exact and approximate quasi-experimental interpretation in nine studies by computing their implicit designs and estimands. 2024-07-30T18:22:12Z Jiafeng Chen http://arxiv.org/abs/2607.05534v1 Empirical Global Games of Regime Change 2026-07-06T18:14:55Z Global games theory provides a tractable framework for analyzing coordination problems with multiple equilibria, with regime overthrow serving as a canonical application. A large empirical literature on coups d'état examines the relationship between country-level characteristics, coup occurrence, and coup success using reduced-form approaches that leave the underlying coordination problem implicit. Bridging these literatures, we develop an estimable global games model of coups d'état. The model incorporates strategic coordination into the empirical analysis of coups, employing the global games framework as an equilibrium selection device. The model distinguishes between the feasibility and desirability of regime overthrow, allowing observable fundamentals to enter separately into beliefs about regime strength and perceived gains from rebellion. The model therefore provides a theoretical basis for decomposing coup outcomes into feasibility and desirability components under maintained exclusion restrictions and equilibrium assumptions. If information on coup strength is available, the model also allows estimation of agents' uncertainty about regime strength. We show how the model can be estimated using simulated maximum likelihood coupled with a contraction mapping, and demonstrate how observable covariates map into regime strength and the perceived benefits of overthrow. We illustrate how the framework can be applied through counterfactuals varying coup benefits, regime strength, and information quality. 2026-07-06T18:14:55Z Matthew J. Baker Khaled Eltokhy Weichao Guo http://arxiv.org/abs/2607.05350v1 Approximate Minimax Estimation of a Bounded Normal Mean via Stochastic Mirror Ascent 2026-07-06T17:31:35Z This paper presents a computational approach to find an approximately minimax estimator for the classical Bounded Normal Mean problem. The suggested procedure is the Bayes estimator corresponding to an approximately least-favorable distribution obtained from a stochastic mirror ascent routine for concave maximization. The paper shows that both the approximately least-favorable distribution and the approximately minimax estimator are indeed close (in a sense we make precise) to their desired targets. Simulation evidence suggests that the approximately minimax estimator can yield, with a reasonable amount of compute, risk improvements from 6% to almost 18% relative to the minimax linear estimator (which is known to admit a maximal improvement of 20%). The approximately minimax estimator is then applied to the problem of how to best aggregate the information contained in local projections and vector autoregressions to estimate an impulse response coefficient. 2026-07-06T17:31:35Z José Luis Montiel Olea Ekaterina Zubova