https://arxiv.org/api/r0XawnMgX+jGICeofK6pDjp3S+s 2026-09-11T18:44:54Z 5831 15 15 http://arxiv.org/abs/2602.20115v2 Compound decisions and empirical Bayes via Bayesian nonparametrics 2026-09-09T17:38:31Z We study compound decision theory from a nonparametric Bayesian perspective, with particular emphasis on their relationship to empirical Bayes (EB) procedures. Motivated by the sharp risk guarantees available for EB procedures based on the nonparametric maximum likelihood estimator (NPMLE), we investigate whether analogous guarantees can be established for fully Bayesian decision rules. In a class of Gaussian compound decision problems, we show that the fully Bayesian posterior mean achieves near-optimal risk. Moreover, it is admissible as a genuine Bayes rule, whereas the corresponding NPMLE plug-in rule is inadmissible. Simulations illustrate the performance of nonparametric Bayes procedures relative to common alternatives. As an application, we apply our methodology to Census tract-level estimates of economic mobility from the Opportunity Atlas. 2026-02-23T18:33:57Z 69 pages Nikolaos Ignatiadis Sid Kankanala http://arxiv.org/abs/2603.07914v3 Event-Study Designs for Discrete Outcomes with Latent Transition Heterogeneity 2026-09-09T16:38:25Z We develop an identification strategy for average treatment effects on the treated (ATT) in panel data with discrete outcomes. For such outcomes, the parallel trends assumption underlying difference-in-differences (DiD) fails in three distinct ways: mean reversion generates divergent trends when groups differ at baseline, counterfactual probabilities can leave the unit interval, and no single trend is well defined for multi-category outcomes. We replace parallel trends with \textit{transition independence}: absent treatment, transition dynamics conditional on pre-treatment outcomes would be identical between treated and control groups. To accommodate selection on persistent unobserved heterogeneity in transition dynamics, we require transition independence to hold only within latent types. Modeling outcomes as a finite mixture of Markov chains, we identify latent-type and aggregate ATTs from short panels. The framework also yields a flow decomposition of the ATT into inflow and outflow channels. In three empirical applications, our ATT estimates differ substantially from conventional DiD. 2026-03-09T03:14:52Z Young Ahn Hiroyuki Kasahara http://arxiv.org/abs/2501.10675v3 Recovering Unobserved Network Links from Aggregated Relational Data: Bayesian Latent Surface Modeling and Penalized Regression 2026-09-09T16:35:52Z Aggregated relational data (ARD) record counts of ties to attribute-defined groups while leaving individual edges unobserved. We compare latent-geometry and regularized network estimators through a common observation map. The comparison distinguishes the realized adjacency matrix, conditional edge probabilities, and model parameters. We study roster-based ARD with known node-level group memberships, giving both estimators the same roster and aggregate counts. We relate the aggregate means to a Poisson working likelihood and a Huber loss, and give their derivatives. Overlapping groups, shared edges, and reporting error affect the interpretation of these objectives. Geometry restricts the representation of edge probabilities, while regularization selects among candidate fits. Identification depends on the observation map and model restrictions rather than uniqueness of a numerical optimizer. A reproducible synthetic experiment specifies the data-generating process, estimation algorithms, and evaluation targets under matched information. The matrix estimator gives better realized-edge rankings and aggregate fit, while the geometric estimator gives lower error for generating probabilities. The resulting framework organizes ARD reconstruction around the interaction of observation design, structural assumptions, and computation. 2025-01-18T06:51:51Z 14 pages, 2 figures. Substantially revised replacement of the withdrawn version. Clarified observation regime and targets; corrected likelihood and loss calculations; added a reproducible matched-input synthetic experiment. Code and saved results are included as ancillary files Yen-hsuan Tseng http://arxiv.org/abs/2110.10650v5 Attention Overload 2026-09-09T11:57:52Z We introduce an Attention Overload Model (AOM) in which alternatives compete for attention, so each alternative's consideration probability weakly decreases as the choice problem expands. This nonparametric restriction has a search-capacity foundation. We identify exactly which preference rankings are compatible with observed choices and establish sharp bounds on latent attention. We then consider heterogeneous preferences in settings where alternatives are presented in a list, establishing nonparametric identification results for attention and the distribution of preferences. We also develop high-dimensional inference and finite-sample methods for these models. Using the travel-mode experiment of Wang and Zhu (2025), we illustrate the methods and find that recovered pairwise preference shares closely match subjects' self-reported rankings. 2021-10-20T16:46:35Z Matias D. Cattaneo Paul Cheung Xinwei Ma Yusufcan Masatlioglu http://arxiv.org/abs/2609.09544v1 When is statistical evidence strong enough? Using hypothesis tests to value data collection 2026-09-08T23:59:49Z We recast statistical significance as a choice between making an immediate policy recommendation and deferring it until further evidence is collected. We show that the welfare-optimal decision corresponds, under minimax regret, to a statistical test whose level depends on the cost and precision of additional evidence. Inverting this rule, we introduce and recommend reporting the abstention-value (A-value) alongside traditional p-values to determine where additional data collection is most needed. The A-value defines the break-even welfare cost of abstaining and recommending further experimentation given the initial evidence. When experimentation capacity is limited, prioritizing additional data collection where A-values are the largest yields finite-sample welfare guarantees. We illustrate its implications for economic program evaluation. 2026-09-08T23:59:49Z Aristotelis Epanomeritakis Davide Viviano http://arxiv.org/abs/2609.09488v1 Two Margins in Difference-in-Differences with a Continuous Treatment 2026-09-08T22:08:23Z This paper studies difference-in-differences with staggered adoption and a continuous, time-invariant dose. Each cohort-time comparison contains two margins. The level margin is the average treatment effect at realized doses. Under level parallel trends it equals the level contrast between the treated cohort and not-yet-treated controls. The response margin is the within-cohort slope of the outcome change on dose. It uses no controls, and its causal interpretation requires a response parallel trends assumption and a restriction on selection on gains. We show that the continuous-dose OLS coefficient in each cohort-time comparison is a convex combination of the response index and the level contrast per unit of mean dose, with a mixing weight that depends on the not-yet-treated share. We provide estimators of both margins, joint inference across cohort-time comparisons and event-time aggregates, and a covariate-adjusted extension. In an application to hydraulic fracturing, the level leads reject a joint zero restriction, whereas the response-index leads do not. The continuous-dose OLS coefficient draws primarily on the level margin. We report the level and response margin separately. 2026-09-08T22:08:23Z Fangzhou Yu http://arxiv.org/abs/2603.27881v2 A Simple and Powerful Diagnostic Test for Binary Choice Models 2026-09-08T18:13:28Z Conventional binary choice models, such as probit and logit, impose thin-tailed errors, and that tail determines whether the parameters of a binary choice model can be estimated at the regular rate. We test the restriction on observables, asking whether the conditional choice probability decays at a polynomial rate in a covariate. Identification of tail heaviness requires no independence, no linear index, and no homoskedasticity. The test is simple to implement and attains nearly the point-optimal power envelope among invariant tests. An application to firm innovation decisions rejects the thin tail. 2026-03-29T21:37:13Z Ting Ji Laura Liu Yulong Wang Jiahe Xing http://arxiv.org/abs/2609.10617v1 Average Treatment Effect Localization: Projection Methods in Synthetic Control 2026-09-08T17:20:57Z Many real-world policies and business interventions require assessing short-term effects to inform timely decisions, even though most causal inference methods focus on long-term average treatment effects. In this paper, we introduce average treatment effect localization (ATEL), which captures localized, short-term policy impacts in panel data settings with a single treated unit and provides early indicators of policy impact. To accommodate both time-varying and nonlinear effects of observed and unobserved covariates, we propose a nonparametric model for untreated outcome, interpreted as a time-varying factor model via sieve approximation. Estimating the time-varying factor model is challenging due to the boundary bias and identification. Our estimation method based on diversified projection can effectively address these issues. We develop an asymptotic distribution theory to facilitate inference for the ATEL estimator. In an empirical application, we apply our proposed methodology to assess the impact of right-to-carry laws on violent crime rate. 2026-09-08T17:20:57Z Ruei-Chi Lee http://arxiv.org/abs/2609.09039v1 Covariate Adjustment in Randomized Experiments: A Unified Framework for Decision and Practice 2026-09-08T17:03:57Z Should researchers adjust for covariates in randomized experiments, and if so, how? The literature offers three distinct prescriptions: do not adjust because randomization guarantees unbiasedness; adjust for outcome-prognostic covariates to improve precision; or adjust for covariates imbalanced between treatment arms. These competing prescriptions create confusion and uncertainty. We develop a unified framework for decision and practice. Given available information, we show that the optimal correction is what we call ex-post bias. The only relevant criterion for adjustment is prognosticity for ex-post bias; neither raw covariate imbalance nor outcome prognosticity is sufficient by itself. We also show that correcting imbalance and improving precision are two sides of the same decision problem. We develop two estimation approaches, one of which recovers familiar adjustment estimators and provides a new theoretical justification for them. Simulations compare alternative covariate-selection and adjustment strategies. Overall, our framework provides a unified foundation for covariate adjustment in randomized experiments. 2026-09-08T17:03:57Z Jiawei Fu Donald P. Green http://arxiv.org/abs/2609.10615v1 ACT, WAIT, or EXPERIMENT: A Causal Governance Framework for Retail Price Optimization Under Abstentions 2026-09-08T15:42:31Z This paper presents a causal decision-making framework for estimating price elasticity in retail channels, a process typically confounded by promotions, competitor movements, and market frictions. Rather than forcing a calculation when data is ambiguous, the system introduces decision abstention (\textsc{wait}) as an active diagnostic tool rather than an estimation failure. Combining Double Machine Learning and conformal prediction, the tool evaluates whether reliable conditions exist to adjust prices or if pausing the decision is preferable. When the system abstains, it exhaustively classifies the reason for the pause, identifying which products require designed pricing experiments or whether aggregating data to the brand level restores usable estimates. Tested on controlled synthetic data, the model shows that this operational discipline drastically reduces estimation error (lowering RMSE from 0.571 to 0.159) and offers a practical, secure alternative to blind estimation in thin-data retail environments. 2026-09-08T15:42:31Z Pedro Cadahia http://arxiv.org/abs/2403.05850v3 Estimating Causal Effects of Discrete and Continuous Treatments with Binary Instruments 2026-09-08T14:17:09Z We propose an instrumental variable framework for identifying and estimating causal effects of discrete and continuous treatments with binary instruments. The basis of our approach is a local copula representation of the joint distribution of the potential outcomes and unobservables determining treatment assignment. This representation allows us to introduce an identifying assumption, so-called constant local dependence, that restricts the local dependence of the copula with respect to the treatment propensity. We show that constant local dependence identifies treatment effects for the entire population and other subpopulations such as the treated. The identification results are constructive and lead to practical estimation and inference procedures based on distribution regression. An application to estimating the effect of sleep on well-being uncovers interesting patterns of heterogeneity. 2024-03-09T09:20:35Z 66 pages, 4 figures, includes supplemental appendix, major revision with respect to previous version Victor Chernozhukov Iván Fernández-Val Sukjin Han Kaspar Wüthrich http://arxiv.org/abs/2604.12611v6 Ordinal Distributional Change and Conservative Transition Benchmarks: Measurement, Identification, and Inference 2026-09-08T12:37:35Z Repeated cross-sections reveal changes in ordinal distributions but not the transitions producing them. I axiomatically characterize a probability metric for ordinal change based on threshold-crossing geometry. Its optimal-transport representation measures the minimum average number of thresholds crossed and yields conservative transition benchmarks. With missing outcomes, I derive sharp identified sets for the discrepancy and endpoint-conditioned benchmark plans. I develop finite-sample-valid projection inference using randomized Monte Carlo calibration and a convergent global-search procedure for the resulting numerical projections. Applied to Arab Barometer data, the framework documents a robust shift toward broader and more regular remittance receipt in Lebanon and a strictly positive amount of minimum ordinal restructuring after allowing for item nonresponse and sampling uncertainty. The conservative benchmarks provide strong numerical evidence that least-displacement restructuring excludes movement toward less frequent receipt and requires some reassignment from nonreceipt to recurrent receipt. 2026-04-14T11:37:24Z Substantially revised version. The asymptotic Wilks calibration has been removed, and the computational analysis now includes a persistent global-search procedure with an almost-sure convergence guarantee for the projection extrema Rami V. Tabri http://arxiv.org/abs/2609.08411v1 Policy Gains or Household Need? The Allocation Logic of China's Dibao Program 2026-09-08T08:18:39Z When social assistance is scarce, should it prioritize households in greatest need or those expected to benefit most? Using panel data from the China Household Finance Survey, we distinguish allocation principles by combining predicted policy gains with entry into China's Minimum Living Standard Guarantee (Dibao). We estimate heterogeneous predicted gains in consumption and education and then recover the conditional priorities revealed by recipient selection. Predicted gains explain little of allocation: a Shapley decomposition attributes 96.9\% of the improvement in allocation fit to household priorities and 3.1\% to predicted gains. Lower income, lower education of the household head, and elderly presence consistently predict higher priority across supported outcome-value specifications. Holding local program capacity fixed and removing household-priority differences while retaining the full model's estimated outcome values, the resulting ranking overlaps with recipients by only 10.9\%, implying 89.1\% recipient churn. These findings show that Dibao allocation is more closely aligned with household circumstances associated with poverty and vulnerability than with policy gains predictable from the outcomes and information observed in our data, underscoring the distinction between distributional and impact targeting in social assistance. 2026-09-08T08:18:39Z Haojie Liu Jiyuan Ling Zihan Lin http://arxiv.org/abs/2609.08335v1 Designing Spatial Treatments 2026-09-08T07:06:05Z Spatial treatments are interventions assigned to locations potentially distinct from those of the responding units. We study their optimal design under a general model in which a unit's response diminishes with distance to a treated site. Our estimand of interest is an ``uncontaminated'' effect equal to the average impact of a single intervention site over all hypothetical sites. We propose a novel design based on a Matérn point process which separates treatments by a distance of at least $r$. A larger choice of $r$ reduces bias by separating interventions but increases variance by reducing their numerosity. We choose $r$ to maximize the rate of convergence of a Horvitz-Thompson estimator and prove that this is minimax rate-optimal. We provide weak conditions under which the estimator is asymptotically normal and propose a variance estimator. 2026-09-08T07:06:05Z Stefan Faridani Michael P. Leung http://arxiv.org/abs/2603.24705v4 Amortized Inference for Correlated Discrete Choice Models via Equivariant Neural Networks 2026-09-07T20:17:28Z Discrete choice models are fundamental tools in management science, economics, and marketing for understanding and predicting decision-making. Logit-based models are dominant in applied work, largely due to their convenient closed-form expressions for choice probabilities. However, they impose restrictive assumptions on the stochastic utility component, constraining our ability to capture realistic substitution patterns. We propose an amortized inference approach that relies on a neural network emulator to approximate choice probabilities for general error distributions, including those with correlated errors. We develop a specialized neural network architecture designed to respect the invariance properties of discrete choice models. We provide group-theoretic foundations for the architecture, including a proof of universal approximation given a minimal set of invariant features. Once trained, the emulator enables rapid likelihood evaluation and gradient computation. We use Sobolev training, augmenting the likelihood loss with a gradient-matching penalty, so that the emulator learns both choice probabilities and their derivatives. We show that emulator-based maximum likelihood estimators are consistent and asymptotically normal under mild approximation conditions, and we provide sandwich standard errors that remain valid for a psuedo-true parameter even with imperfect likelihood approximation. Simulations show significant gains over the GHK simulator in accuracy and speed. 2026-03-25T18:30:11Z Easton Huch Michael Keane