https://arxiv.org/api/hsGaC23ElSiQoD9V5wX/asU/fAg 2026-07-21T10:48:22Z 5630 60 15 http://arxiv.org/abs/2607.10043v1 The Projection Solution to the Incidental Parameter Problem 2026-07-10T23:48:55Z This paper introduces a new approach to econometric analysis of nonlinear panel data models when the number of observations per observational unit is small. In such models the presence of variables that are constant within, while varying across, units results in an incidental parameter problem. The approach taken in this paper removes these incidental parameters via projection, which produces a correspondence specifying all combinations of observed variables and within-unit-varying unobserved heterogeneity that are achievable by choice of some value of the unit-specific incidental parameters. With unit-specific variables removed, there is no need for assumptions concerning their joint distribution with other variables. The result is an incomplete model which is typically partially identifying. Identified sets are characterized via moment inequalities using tools of random set theory. Examples of application to static and dynamic models with discrete or continuous outcomes using distribution-free restrictions on within-unit-varying unobserved heterogeneity are presented. 2026-07-10T23:48:55Z Andrew Chesher Adam M. Rosen Yuanqi Zhang http://arxiv.org/abs/2607.09608v1 Media Measurement and the Assisted Own Goal: Attribution, Marketing-Mix Models, and Individual-Level Incrementality 2026-07-10T17:05:59Z We use the assisted own goal hypothesis as a lens into media measurement. A demand-generating (upper-funnel) advertising platform such as a short-video social network can cause an incremental purchase, yet see that purchase booked on -- and credited to -- a downstream trusted marketplace, because consumers who discover a product on the platform complete the transaction elsewhere, for example because of distrust of the generating platform as a psychological mechanism. Under attribution-based return-on-ad-spend (ROAS) measurement, the diverted conversions are invisible to the originating platform. Marketing-mix models (MMMs) do not know which channel to credit with the outcome, and channel-by-week aggregation denies the audience-level granularity that budget decisions require. We develop an incrementality-based measurement model with two ingredients: ambient audience-level randomization -- each activated audience carries its own intent-to-treat (ITT) experiment -- and an individual-level extension of Predicted Incrementality by Experimentation (PIE), which learns a mapping from individual features to experiment-identified incremental outcomes. Because ITT contrasts are computed on channel-complete outcomes, the estimator is unbiased and the own goal disappears 2026-07-10T17:05:59Z Tobias Konitzer GrowthLoop http://arxiv.org/abs/2607.09536v1 Misspecified regressions with mixed regressors: robust inference and causal interpretation 2026-07-10T15:41:29Z For analytic convenience, existing statistical frameworks either assume random or fixed regressors. However, it is a little awkward that they do not cover the practical case of estimating the average treatment effect in experiments with randomized treatments and non-randomized, fixed pretreatment covariates. We unify the literature by providing the theory for regressions with mixed regressors that contain both random and fixed components. Importantly, our theory allows for misspecification of the regression functions. We first establish general results for estimating equations with both random and fixed components and then use it to analyze misspecified linear regression, with applications to completely randomized experiments. We focus on the causal interpretation of the regression coefficients and standard errors even when the models are wrong. We start with the theory for independent data and then extend the discussion to clustered data. 2026-07-10T15:41:29Z Mengsi Gao Peng Ding http://arxiv.org/abs/2607.09461v1 Deep Learning for Dynamic Programming with Recursive Utility Using First-order Conditions 2026-07-10T14:36:28Z This paper proposes the certainty-equivalent first-order learning (CEFOL) algorithm, a deep learning algorithm for solving discrete-time dynamic programming problems with recursive utility. Dynamic programming with recursive utility is challenging because nonlinear certainty equivalent appears in the Bellman equation and the first-order optimality conditions but is difficult to evaluate. By introducing a separate neural network to represent the certainty equivalent, CEFOL enables the exploitation of the Bellman and model-specific first-order optimality conditions. In addition to certainty equivalent, CEFOL also uses neural networks to learn the value functions, policy functions, and Lagrange multipliers by using model-specific first-order conditions to construct residuals for minimization. By using first-order and KKT residuals to learn the policy, CEFOL directly accommodates general equality and inequality constraints on the controls, including occasionally binding constraints, without requiring penalty functions or problem-specific reformulations. We apply the algorithm to risk-sensitive and Epstein--Zin consumption-saving problems, a small-noise robust-control problem, and a DSGE model with recursive preferences and stochastic volatility. Across these applications, out-of-sample Bellman diagnostics and model-specific optimality residuals, including Euler or first-order residuals where applicable, are generally of order 1.0e-4 to 1.0e-3 over the relevant state regions, with larger values mainly near binding constraints, and the learned value and policy functions closely match VFI benchmarks when available. The CEFOL algorithm also works for dynamic programming problems with expected utility, as expected utility is a special case of recursive utility. 2026-07-10T14:36:28Z 86 pages, 44 figures Xianhua Peng Wu Guo Songyan Wang Jianfei Zhu http://arxiv.org/abs/2603.20388v2 From Cross-Validation to SURE: Asymptotic Risk of Tuned Regularized Estimators 2026-07-10T13:18:31Z We derive the asymptotic risk function of regularized empirical risk minimization (ERM) estimators tuned by $n$-fold cross-validation (CV). The out-of-sample prediction loss of such estimators converges in distribution to the squared-error loss (risk function) of shrinkage estimators in the normal means model, tuned by Stein's unbiased risk estimate (SURE). This risk function provides a more fine-grained picture of predictive performance than uniform bounds on worst-case regret, which are common in learning theory: it quantifies how risk varies with the true parameter. As key intermediate steps, we show that (i) $n$-fold CV converges uniformly to SURE, and (ii) while SURE typically has multiple local minima, its global minimum is generically well separated. Well-separation ensures that uniform convergence of CV to SURE translates into convergence of the tuning parameter chosen by CV to that chosen by SURE. 2026-03-20T18:05:39Z Karun Adusumilli Maximilian Kasy Ashia Wilson http://arxiv.org/abs/2507.14621v3 Testing Clustered Equal Predictive Ability with Unknown Clusters 2026-07-09T16:33:59Z We develop tests of clustered equal predictive ability (C-EPA) in panels where the clusters are unknown and estimated by the Panel Kmeans algorithm. To address the challenge of testing hypotheses that depend on data-driven clusters, we adopt a selective conditional inference framework. Specifically, we first derive a Wald-type test for pairwise equality and show that the limiting distribution of its square root conditional on the estimated clusters is that of a truncated $χ$ variable. We characterize the associated truncation set by quadratic inequalities in the data space. Then, for the C-EPA hypothesis, we propose a $p$-value combination method by aggregating the evidence against the pairwise equality and overall EPA null hypotheses. The Monte Carlo results show accurate size control and good finite-sample power of the proposed tests. An empirical application to exchange-rate forecasting, using both traditional time-series models and machine-learning methods, illustrates the practical relevance of our procedure. 2025-07-19T13:38:05Z Oguzhan Akgun Alain Pirotte Giovanni Urga Zhenlin Yang http://arxiv.org/abs/2407.20386v3 On the power properties of inference for parameters with interval identified sets 2026-07-09T16:15:08Z This paper studies the power properties of confidence intervals (CIs) for a partially-identified parameter of interest with an interval identified set. We assume the researcher has bounds estimators needed to construct the CIs proposed by Imbens and Manski (2004), Stoye (2009), and Stoye (2020), denoted by CI_alpha^1, CI_alpha^2, CI_alpha^3, and CI_alpha^4. We also assume these bounds estimators are ``ordered'': the lower bound estimator is less than or equal to the upper bound estimator. This setup arises in economic applications involving missing data and treatment effects. Under these conditions, we establish two results. First, we show that CI_alpha^1 and CI_alpha^2 are equally powerful, and both dominate CI_alpha^3 and CI_alpha^4. Second, we consider a favorable situation in which there are two possible bounds estimators to construct these CIs, and one is more efficient than the other. One would expect that the more efficient bounds estimator yields more powerful inference. We prove that this desirable result holds for CI_alpha^1 and CI_alpha^2, but not necessarily for CI_alpha^3 or CI_alpha^4. In summary, within the class of models considered, CI_alpha^1 and CI_alpha^2 have identical power properties, and both compare favorably to CI_alpha^3 or CI_alpha^4. 2024-07-29T19:22:48Z 60 pages, 50 pages of main text and 10 of online supplement Federico A. Bugni Mengsi Gao Filip Obradovic Amilcar Velez http://arxiv.org/abs/2607.08324v1 Finite-Population Inference for Heterogeneity in Many-Group Synthetic Difference-in-Differences 2026-07-09T10:08:00Z Synthetic difference-in-differences is widely used to estimate treatment effects for many treated groups against a common donor pool. When the same donors are reused across groups, the group-specific estimates are cross-sectionally dependent, and plug-in second moments overstate effect heterogeneity. We develop finite-population inference for heterogeneity in many-group synthetic difference-in-differences: the projection of realized group effects on observed group covariates, the projected group-effect curve, the between-group variance, and the explained share. The theory combines a modular first-stage representation, a joint covariance kernel for donor sharing and block dependence, analytic and leave-out corrections for second moments, and calibrated omnibus and directed tests under explicit exchangeability or fit-matching conditions. In an American Community Survey application to the Affordable Care Act Medicaid expansion, whose estimand is the incremental effect of expansion status, pre-expansion uninsured rates explain much of the state-level effect variation on the percentage-point scale, household split-samples validate the decomposition, and donor sharing materially increases the standard error for the average effect. In a county-level Clean Air Act application, groupwise estimates are noisy, but a pre-specified projection on baseline fine-particulate pollution reveals a sign-stable directed component under state and division block covariance; placebo analyses attribute part of the raw gradient to regional convergence. 2026-07-09T10:08:00Z Takahiro Hoshino Makoto Nakakita http://arxiv.org/abs/2607.03933v2 Rational Bubbles at the Spectral Edge: An Operator-Spectral Theory of Fragility, Identification and Finite-Sample Certification 2026-07-09T08:07:43Z When markets move more and more in lockstep, are they drifting towards the point where a price bubble becomes possible, and can that drift be measured before the crossing? This paper joins two long-separate ideas, that a rational bubble is a price outgrowing its dividends and that a crisis threshold can be read off the strength of a market's single dominant factor, onto one object recovered from the data: a summary of how asset returns move together, paired with a discount rate. We call this crossing point the fragility edge and show it plays three roles at once. A stated discipline says what the data support: the edge firmly, with a margin of error; whether a bubble exists, only roughly; which asset carries it, not at all. Across eighteen global equity indices from 2004 to 2024, that dominant factor strengthens in every documented crisis, the market collapsing from about six to about four independent factors; once the discount is set so that calm markets sit at the edge, this strength crosses it in crisis. These readings coincide with crises, not forecasts. 2026-07-04T15:56:28Z JEL classification: C62, D58, D80, E10; Keywords: rational bubbles, dependence operator, spectral radius, transversality, systemic fragility, partial identification, certification Avishek Bhandari http://arxiv.org/abs/2607.11922v1 Modeling the Dynamic Relationship Between Brent Crude Oil Prices and the Nepal Stock Exchange: An Integrated Econometric and Explainable Machine Learning Approach 2026-07-09T02:59:00Z This study examines the dynamic relationship between the global oil prices and Nepal Stock Exchange (NEPSE) using an integrated approach which combines traditional econometric techniques with machine learning and explainable AI techniques. For this, Daily data of International Oil prices and NEPSE index is analyzed from approximately thirteen years (June 2013 to June 2026) using Granger causality, EGARCH(1,1), and DCC-GARCH models to examine different properties like predictive relationships, asymmetric volatility behaviour, and time-varying correlations. To further supplement the econometric analysis, Machine Learning Models like Random Forest, LightGBM, and XGBoost algorithms were used to capture nonlinear relationships, along with explainable artificial intelligence techniques like SHAP values, Partial Dependence Plots, and Individual Conditional Expectation plots to further interpret the results of the model. The results from the econometric analysis showed a statistically significant unidirectional Granger causality from Brent crude oil to NEPSE with a four-day lag, high volatility persistence in both markets, and weak yet highly time-varying conditional correlations. Among the machine learning models, XGBoost achieves the best performance, and explainability analysis reveals that NEPSE own momentum and short-term volatility mainly influence its own behaviour and oil-related information serves as a minor, method-dependent contributor. The findings demonstrate that econometric and explainable machine learning approaches provide insights into the oil and equity market relationship in a way that each approach complements the result of one another. 2026-07-09T02:59:00Z Anamol Khadka Milan Arjel Ayush Lataula Aayam Dhakal Prajun Trital Mingmar Sherpa Biman Rimal http://arxiv.org/abs/2505.18077v3 Bayesian Deep Learning for Discrete Choice 2026-07-08T18:07:33Z Discrete choice models (DCMs) are used to analyze individual decision-making in contexts such as transportation choices, political elections, and consumer preferences. DCMs play a central role in applied econometrics by enabling inference on key economic variables, such as marginal rates of substitution, rather than focusing solely on predicting choices on new unlabeled data. However, while traditional DCMs offer high interpretability and support for point and interval estimation of economic quantities, these models often underperform in predictive tasks compared to deep learning (DL) models. Despite their predictive advantages, DL models remain largely underutilized in discrete choice due to concerns about their lack of interpretability, unstable parameter estimates, and the absence of established methods for uncertainty quantification. Here, we introduce a deep learning model architecture specifically designed to integrate with approximate Bayesian inference methods, such as Stochastic Gradient Langevin Dynamics (SGLD). Our proposed model collapses to behaviorally informed hypotheses when data is limited, mitigating overfitting and instability in underspecified settings while retaining the flexibility to capture complex nonlinear relationships when sufficient data is available. We demonstrate our approach using SGLD through a Monte Carlo simulation study, evaluating both predictive metrics--such as out-of-sample balanced accuracy--and inferential metrics--such as empirical coverage for marginal rates of substitution interval estimates. Additionally, we present results from two empirical case studies: one using revealed mode choice data in NYC, and the other based on the widely used Swiss train choice stated preference data. 2025-05-23T16:33:47Z Daniel F. Villarraga Ricardo A. Daziano http://arxiv.org/abs/2607.07524v1 Robust Inference for Weighted Estimands 2026-07-08T15:18:54Z Researchers often conduct inference on weighted estimands, defined as weighted averages of group-level effects. Example settings include event studies with cohort-level effects and experiments with site-level effects. Under heterogeneous effects, different weighting schemes yield estimands with distinct empirical and policy interpretations, leading to ambiguity and disagreement over the choice of weights. I establish bounds on differences between weighted estimands and confidence bounds on effect heterogeneity, which I use to construct estimators that minimize worst-case bias and confidence intervals that are uniformly valid over classes of weighted estimands. I apply these methods to an event study in Lakdawala, Nakasone, and Kho (2023), which studies the effects of school-based internet access on test scores. I find that results are robust to broad classes of weights. I then apply the methods to Tennessee's Project STAR experiment and find that results are sensitive to small departures from baseline weights. 2026-07-08T15:18:54Z Vod Vilfort http://arxiv.org/abs/2603.10999v2 Double Machine Learning for Time Series 2026-07-08T10:04:47Z We modify the Double Machine Learning estimator to broaden its applicability to macroeconomic time-series settings. A deterministic cross-fitting step, termed Reverse Cross-Fitting, leverages the time-reversibility of stationary series to improve sample utilization and efficiency. We detail and prove the conditions under which the estimator is asymptotically valid. We then demonstrate, through simulations, that its performance remains valid in realistic finite samples and is robust to model misspecification and violations of assumptions, such as heteroskedasticity. In high dimensions, predictive metrics for tuning nuisance learners do not generally minimize bias in the causal score. We propose a calibration rule targeting a "Goldilocks zone", a region of tuning parameters that delivers stable, partialled-out signals and reduced small-sample bias. Finally, we apply our procedure to residualized Local Projections to estimate the dynamic effects of a rise in Tier 1 regulatory capital. The results underscore the usefulness of the methodology for inference in macroeconomic applications. 2026-03-11T17:22:57Z The Econometrics Journal, utag019, 2026 Milos Ciganovic Federico D'Amario Massimiliano Tancioni 10.1093/ectj/utag019 http://arxiv.org/abs/2407.07988v2 Production function estimation using subjective expectations data 2026-07-08T09:59:07Z Standard proxy methods for estimating production functions in the \Olley and Pakes (1996) tradition require assumptions on input choices. We introduce a new method that exploits (increasingly available) data on firms' expectations of their future output and inputs that allows us to obtain consistent production function parameter estimates while relaxing these input demand assumptions. In contrast to both proxy and dynamic panel methods like Blundell and Bond (2000), our proposed estimator can be implemented on a single cross-section of data and Monte Carlo simulations show it outperforms alternative estimators when firms' material input choices are subject to optimization error. Implementing a range of production function estimators on UK panel data, we find our proposed estimator yields results that are either similar to or more credible than commonly-used alternatives. These differences are larger in industries where material inputs appear harder to optimize. We show that the share of cross-firm TFP dispersion accounted for by persistent productivity differences is substantially larger when calculated using parameter estimates from our proposed estimator. 2024-07-10T18:40:24Z Agnes Norris Keiller Aureo de Paula John Van Reenen http://arxiv.org/abs/2507.20550v2 Policy Learning under Unobserved Confounding: A Robust and Efficient Approach 2026-07-08T04:44:18Z This paper develops a robust and efficient method for policy learning from observational data in the presence of unobserved confounding, complementing existing instrumental variable (IV) based approaches. We employ the marginal sensitivity model (MSM) to relax the commonly used yet restrictive unconfoundedness assumption by introducing a sensitivity parameter that captures the extent of selection bias induced by unobserved confounders. Building on this framework, we consider two distributionally robust welfare criteria, defined as the worst-case welfare and policy improvement functions, evaluated over an uncertainty set of counterfactual distributions characterized by the MSM. Closed-form expressions for both welfare criteria are derived. Leveraging these identification results, we construct doubly robust scores and estimate the robust policies by maximizing the proposed criteria. Our approach accommodates flexible machine learning methods for estimating nuisance components, even when these converge at moderately slow rates. We establish asymptotic regret bounds for the resulting policies, providing a robust guarantee against the most adversarial confounding scenario. The proposed method is evaluated through extensive simulation studies and empirical applications to the JTPA study and Head Start program. 2025-07-28T06:29:49Z Zequn Jin Gaoqian Xu Xi Zheng Yahong Zhou