https://arxiv.org/api/fHFpxRRMlUDXKInkbChK+5dGqdo 2026-09-11T17:47:31Z 5831 0 15 http://arxiv.org/abs/2609.11915v1 Generative Marketing Mix Modeling: A Causal Inference Framework Linking GEO and GEM to Business Impact 2026-09-10T17:57:28Z Generative artificial intelligence changes how firms reach customers, but standard marketing data do not record how often users see and notice a firm's name in generated answers. We develop Generative Marketing Mix Modeling (GMMM) to estimate the causal effects of Generative Engine Optimization (GEO) and Generative Engine Marketing (GEM). For GEO, GMMM combines repeated generated answers with question counts, shares of use across generative systems, and notice probabilities. For GEM, it combines records of sponsored placements with notice probabilities. GMMM compares expected business responses under alternative treatment sequences and establishes sufficient conditions for identifying the resulting effects. We investigate the empirical performance of the proposed method using simulated answers to product recommendation in English and Japanese. 2026-09-10T17:57:28Z Masahiro Kato Daiki Honma Taka Kato http://arxiv.org/abs/2608.23508v2 Testing selection on observables in parametric models with refreshment samples 2026-09-10T17:35:45Z In panels with sample selection (that may occur due to attrition, nonresponse, etc.), the assumption of selection on observables (missing at random, MAR) is commonly imposed despite often being implausible. However, this assumption becomes testable when a refreshment sample is available. We develop a statistical test of MAR based on a distance between two estimated distributions: one obtained using the standard inverse probability weighting (IPW) that is valid under MAR and the other obtained using an alternative weighting that is valid under a weaker assumption of additive nonignorability of Hirano et al. (2001). This test implicitly compares the distribution of the IPW-weighted sample in the attrition period with the distribution of the refreshment sample, which coincide if the MAR assumption holds. We establish that, when the input distributions are parametric, our test statistic converges to the generalized chi-squared distribution under the null of MAR. This limit distribution can be estimated using the recursive formulas derived by Franguridi et al. (2026). We illustrate the performance of our test in Monte Carlo simulations. Finally, we apply our test to an empirical example using a subsample of the Understanding America Study (UAS) dataset. 2026-08-24T17:12:10Z Grigory Franguridi Arie Kapteyn http://arxiv.org/abs/2509.13492v2 Generalized Covariance Estimator under Misspecification 2026-09-10T15:44:21Z This paper investigates the properties of the Generalized Covariance (GCov) estimator under misspecification with application to processes with local explosive patterns, such as causal-noncausal processes. We show that GCov is consistent and has an asymptotically Normal distribution under misspecification. Then, we construct GCov-based Wald-type and score-type tests to test one specification against the other, all of which follow a $χ^2$ distribution. We validate the finite-sample performance of the proposed estimators and tests in the context of causal-noncausal models. Finally, we provide applications of the noncausal model to the final energy demand commodity index. 2025-09-16T19:49:47Z Aryan Manafi Neyazi http://arxiv.org/abs/2609.11575v1 Market-Informed Networks for Modeling and Forecast Evaluation of Financial Extremes 2026-09-10T14:10:43Z Modeling the joint distribution of extreme values in high-dimensional financial time series is challenging because extremes are sparse and locally extreme observations are not necessarily extreme relative to their full marginal distribution. To address this, we introduce a time-dependent network Hüsler-Reiss model in which market-informed adjacency matrices determine how strongly observations contribute to the estimation. We propose binary and weighted specifications, including the Joint Extremes Adjacency Matrix (JEAM) which combines information about individual extremeness with historical patterns of joint extreme movements. In the forecasting evaluation part, covering one-minute stock returns from three sectors of the S&P 100, JEAM achieves the best out-of-sample log scores for both tail directions; improving scores by 12.5-13.6% in the lower tail and 11.4-14.9% in the upper tail. The results show that incorporating market-informed network structures in the estimation, improves forecast evaluation of extremes across time series. 2026-09-10T14:10:43Z Ayla Jungbluth Johannes Lederer Simon Trimborn http://arxiv.org/abs/2606.14887v2 Estimating Sloppy Directions via KDE: The Case of Kirman's Ants 2026-09-10T14:10:35Z Models whose predictions depend on only a handful of well-constrained parameter combinations, termed sloppy models, are ubiquitous in nonlinear stochastic systems. The information-geometric approach to sloppiness advocates using the symmetrized Kullback--Leibler divergence and its associated Hessian, the Fisher Information Matrix (FIM), as the natural loss function. However, prior applications have relied on analytically known or parametrically fitted distributions. In practice, for general agent-based or stochastic models the distribution must be estimated from simulation data. I demonstrate, using Kirman's ant recruitment model as a worked example, that a standard kernel density estimate (KDE) converges to the analytical FIM eigenvectors and eigenvalues with simulation budgets accessible in practice. I derive the analytical Hessian in closed form, show numerical convergence of the KDE-based estimate as a function of simulation data, and demonstrate how the stiff direction enables efficient phase exploration across the model's unimodal and bimodal regimes. 2026-06-12T18:48:22Z Karl Naumann-Woleske http://arxiv.org/abs/2401.00618v5 Changes-in-Changes for Ordered Choice Models with Underreporting 2026-09-10T09:37:20Z We develop a Difference-in-Differences framework for discrete, ordered outcomes subject to underreporting. Such outcomes commonly arise in self-reported surveys on socially undesirable or stigmatized behaviors, where respondents may conceal their true behavior. For a discrete Changes-in-Changes model that is shown to admit an equivalent threshold-crossing representation, we derive nonparametric bounds for the counterfactual and factual outcome distributions as well as for the associated quantile treatment effects when outcomes are underreported. These bounds are shown to be sharp uniformly across outcome levels under additional support conditions, and we propose suitable estimation and bootstrap inference procedures. In an extension, we also consider a semiparametric underreporting model that allows to point identify and estimate distributional treatment effects. As an application, we investigate the impact of recreational marijuana legalization on the consumption behavior of 8th-grade students in several U.S. states. 2024-01-01T00:12:56Z Daniel Gutknecht Cenchen Liu http://arxiv.org/abs/2609.11222v1 Optimal Covariate Adjustment beyond the Average Treatment Effect: Treated-Population and Overlap-Weighted Estimands 2026-09-10T08:23:14Z Graphical causal inference supplies a complete theory of efficient covariate adjustment for the average treatment effect: one adjustment set, computable from the graph, is optimal under every compatible distribution. We show that this is a property of the average treatment effect's inverse-prevalence weights, not of causal estimands in general. For the average treatment effect on the treated we index the efficiency bound by the adjustment set and derive exact identities for its change under treatment-side and outcome-side extensions of a valid set. Covariates that predict only the treated-arm outcome are exactly efficiency-neutral, and covariates that predict the control-arm outcome can strictly increase the bound when the propensity is below one half -- a reversal of the supplementation lemma whose source is an arithmetic-geometric-mean inequality that holds for the average treatment effect and fails for the treated-population estimand. A construction with two faithful distributions on one graph proves that no graphical optimality criterion exists for the treated-population estimand; under no effect modification the ATE-optimal set is nonetheless optimal among the graphically valid sets, with an exact expression for its advantage. The results extend to weighted average treatment effects with propensity-dependent weights, yielding symmetric thresholds for overlap weights, an estimand-drift phenomenon under instrument adjustment, and a characterization of constant weights as the only smooth positive weights for which outcome-side supplementation never increases the bound. Simulations and the LaLonde data provide illustrations. 2026-09-10T08:23:14Z 45 pages, 1 figure, 3 tables; proofs and numerical verification in the appendices. Replication code and data: https://github.com/sokubo/paper-estimand-adjustment-replication Shoki Okubo http://arxiv.org/abs/2609.10971v1 Experimental Design for Policy Choice 2026-09-10T01:44:37Z We show how to optimally design experiments when the resulting data will be used to choose a welfare-maximizing policy subject to constraints. A decision maker seeks to maximize Bayes expected welfare by choosing a policy whose effects depend on an unknown finite-dimensional parameter. The decision maker has access to a first wave of experimental data with a fixed design but may choose the design of a second wave that will be collected before choosing the policy. The resulting experimental design--policy choice problem is a very high-dimensional dynamic program that is generally intractable in finite samples. We propose a tractable approximation based on the limit experiment and show it is asymptotically optimal using a new asymptotic representation theorem for adaptive experiments with continuous treatments. We apply the method to a conditional cash transfer experiment and demonstrate the potential for large gains from tailoring the experiment to the policy choice. 2026-09-10T01:44:37Z Samuel D. Higbee http://arxiv.org/abs/2502.19620v3 Triple Difference Designs with Heterogeneous Treatment Effects 2026-09-10T01:11:04Z Triple difference designs have become increasingly popular in empirical economics. The advantage of a triple difference design is that, within a treatment group, it allows another subgroup of the population -- potentially less impacted by the treatment -- to serve as a comparison for the subgroup of interest. While literature on difference-in-differences has discussed heterogeneity in treatment effects between treated and control groups or over time, relatively little attention has been given to triple difference designs and the implications of heterogeneity in treatment effects in this setting. In this paper, I show that the parameter identified under common triple difference assumptions does not allow for causal interpretation of differences between subgroups when subgroups may differ in their underlying (unobserved) treatment effects. I propose a new parameter of interest, the controlled difference in average treatment effects on the treated, which allows for causal comparisons between subgroups. I then propose identification assumptions and doubly-robust estimators for this parameter. I use a simulation study to highlight the desirable finite-sample properties of these estimators, as well as to show the difference between the two parameters. An empirical application shows the importance of considering treatment effect heterogeneity in practical applications. 2025-02-26T23:18:41Z Laura Caron http://arxiv.org/abs/2609.10925v1 Identification in Linear Quantile Panel Models 2026-09-10T00:17:42Z This paper studies identification in linear quantile panel models with unrestricted individual heterogeneity when the number of time periods is fixed and small. We impose strict exogeneity, whereby the conditional quantile restriction holds given the individual's complete regressor history and latent individual effect, but otherwise allow the disturbances to be arbitrarily dependent over time. 2026-09-10T00:17:42Z Shakeeb Khan Elie Tamer http://arxiv.org/abs/2205.10310v5 Partial Identification from Bunching at Kinks and Notches 2026-09-09T20:30:39Z This paper proposes a partial identification approach to using choices around a kink or notch as a means to overcome endogeneity in a general nonparametric choice model. I show that observed choices are informative about the joint distribution of two counterfactual choices, which leads naturally to analyzing bunching using tools from causal inference. I define as the parameter of interest an average treatment effect among bunchers, which nests the traditional elasticity parameter while allowing for heterogeneity in responsiveness. Identification of the buncher ATE (and hence the elasticity) requires the researcher to extrapolate the distribution of each counterfactual choice beyond where it is observed. I introduce a flexible family of nonparametric shape restrictions to obtain partial identification, and propose a method for empirical validation of this approach using changes in the location of the kink/notch or using the distribution from another comparison group. I find that a log-concavity version of the distributional assumption is widely---though not universally---supported across the empirical literature, and leads to informative bounds in a canonical tax kink application. The approach accommodates diffuse bunching, at the expense of wider bounds. 2022-05-20T17:22:57Z Leonard Goff http://arxiv.org/abs/2604.07131v3 When Is GMM Actually LATE? Weighting Matrices and Causal Interpretation in Overidentified IV 2026-09-09T18:56:54Z Under heterogeneous treatment effects, the weighting matrix of overidentified IV-GMM selects the estimand, not just its precision. We characterize the selection exactly: for any parameter-free weighting-matrix map, the GMM estimand is a sum-to-one combination of instrument-specific Wald estimands, with closed-form weights and an exact non-negativity condition; efficient weighting adds a heterogeneity penalty. Continuously updated GMM exits this class through a variance-score remainder. Under positive regression dependence each Wald estimand is a convex combination of compliance-type LATEs, and under maintained validity a $J$-rejection indicates unequal Wald estimands rather than invalid instruments. We propose Representativeness Targeting (RT), which estimates a researcher-specified convex combination of the Wald estimands without imposing a common coefficient across moments; RT weights compliance types nonnegatively, attains the local asymptotic minimax bound for its target, and extends to unreachable policy targets via projection with identification-gap bounds. In Tennessee STAR, we find the $J$-test rejects the Wald-estimand equality while the heterogeneity penalty pulls the efficient-GMM estimate substantially below 2SLS; in a patent-leniency design, RT delivers a policy-relevant surrogate that standard GMM weightings miss. 2026-04-08T14:24:30Z Chun Pang Chow Hiroyuki Kasahara http://arxiv.org/abs/2211.13610v8 Dynamic Innovation Transmission Through Networks: Theory, Large $T$-Inference, and the Role of Input-Output Conversion in Business Cycles 2026-09-09T18:15:36Z I develop an econometric framework that rationalizes the dynamics of a cross-sectional variable by lagged transmissions of innovations along bilateral links between units. The NVAR I propose is parameterized by $α\in \mathbb{R}^p$, $p \in \mathbb{N}$ -- showing the time profile of transmission along a direct link -- and $q \in \mathbb{N}$ -- showing the relative frequency of network interactions to observation. While nesting the Spatial Autoregression and Spatial Error Model in the limit as $q \to \infty$ and producing equivalent impulse-responses in the long run for any finite $q$, it can accommodate general transmission patterns over time and yields ``networked'' transition dynamics distinct from those implied by autocorrelated innovations. For a given network, $α$ is identified at least up to alternating sign and its Gaussian Maximum Likelihood estimator is consistent and asymptotically Normal under mild assumptions. I then estimate an NVAR for monthly industrial production among 23 US manufacturing sectors, as derived under a Real Business Cycle economy with lagged input-output conversion (IOC), and I quantify the extent to which business cycles can be endogenized by the lagged transmission of productivity shocks along supply chains. Compared to an economy with contemporaneous IOC, the preferred lagged-IOC specification reduces the estimated shock-variances on average by 73\% and accounts for around 85\% of the persistence in aggregate output growth. In this environment, a single common productivity shock explains 90\% of aggregate fluctuations, leaving a negligible role for sector-specific shocks once sectoral heterogeneity in the temporal exposure to common shocks is accounted for. 2022-11-24T13:53:15Z Marko Mlikota http://arxiv.org/abs/2608.13224v2 Parameter Identification and Inference in Discretely Sampled or Temporally Aggregated Autoregressions 2026-09-09T17:48:54Z I consider an AR($p$) process that is observed every $q$ periods, either as a snapshot (stock variable) or as a sum over the sampling interval (flow variable). I first characterize the resulting ARMA process followed by observables. Under fairly mild assumptions, I then derive the identified set for general lag lengths $p \in \mathbb{N}$ and sampling frequencies $q \in \mathbb{N}$, I bound its cardinality, and I provide an algorithm to compute all candidate points and determine their membership in the identified set. My exact but implicit characterization supports the following conjecture that I prove in some settings and verify numerically more broadly: (i) the error term-variance is point-identified, (ii) under temporal aggregation, the autoregressive parameters are point-identified, and (iii) under discrete sampling they are point-identified for odd $q$ and identified up to alternating sign for even $q$. My analysis supplements existing inference results that show consistency and asymptotic Normality of the Gaussian Maximum Likelihood estimator conditional on point-identification. Holding the number of observations fixed, I show that its precision does not necessarily decrease with $q$. 2026-08-13T13:27:30Z Marko Mlikota http://arxiv.org/abs/2501.19394v5 Fixed-Population Causal Inference for Models of Equilibrium 2026-09-09T17:45:32Z In contrast to problems of interference in (exogenous) treatments, models of interference in unit-specific (endogenous) outcomes do not usually produce a reduced-form representation where outcomes depend on other units' treatment status only at a short network distance, or only through a known exposure mapping. This remains true if the structural mechanism depends on outcomes of peers only at a short network distance, or through a known exposure mapping. In this paper, we first define causal estimands that are identified and estimable from a single experiment on the network under minimal assumptions on the structure of interference, and which represent average partial causal responses which generally vary with other global features of the realized assignment. Under a fixed-population, design-based approach, we show unbiasedness and consistency for inverse-probability weighting (IPW) estimators for those causal parameters from a randomized experiment on a single network. We also analyze more closely the case of marginal interventions in a model of equilibrium with smooth response functions where we can recover LATE-type weighted averages of derivatives of those response functions. Under additional structural assumptions, these ``agnostic" causal estimands can be combined to recover model parameters, but also retain their less restrictive causal interpretation. 2025-01-31T18:48:12Z Konrad Menzel