https://arxiv.org/api/UQ0qY4KvRGXWpLgEU9GEJiY9IKI2026-07-21T17:47:26Z56309015http://arxiv.org/abs/2607.05291v1Forecasting Realized Volatility with Time Series Foundation Models: A Comparison with Econometric Benchmarks2026-07-06T16:30:41ZWe ask whether pretrained time series foundation models (TSFMs) improve on established econometric benchmarks for forecasting realized volatility. Using the VOLARE dataset, we conduct the first systematic comparison of nine zero-shot TSFMs against eight econometric specifications, including the Heterogeneous Autoregressive (HAR) family, across 50 assets in equities, foreign exchange, and futures, and three forecast horizons, with formal pairwise and multi-model forecast-comparison tests. Foundation models do not deliver a uniform gain. Pooled losses favor them, but the advantage is concentrated in a few outlier assets; averaging each asset's loss ratio to a well-specified Log-HAR benchmark, so that no single asset dominates, only one small model, Tiny Time Mixers (TTM), beats the benchmark at every horizon, and by a narrow margin. The other foundation models do not improve on Log-HAR, and the econometric benchmarks remain competitive throughout. A Mincer--Zarnowitz recalibration, which removes level and scale bias from every forecast, shows that much of the short-horizon advantage reflects better-scaled forecasts rather than better prediction of volatility dynamics, and only at the monthly horizon does a genuine informational gain remain. Because this edge is thin and even TTM is not best on every asset, a simple equal-weight average of TTM and Log-HAR matches the best single model and enters the Model Confidence Set for 98 to 100\% of assets, more often than either component alone, so a forecaster need not identify the best model for each asset in advance. Our most durable finding is that performance varies so much across foundation-model architectures that choosing the right architecture matters more than the broader choice between foundation and econometric models.2026-07-06T16:30:41Z5 figures, 41 pagesAlessio Brinihttp://arxiv.org/abs/2607.05215v1Variance Estimation for Saturated Fixed-Effect Specifications2026-07-06T15:32:08ZWe characterize the asymptotic behavior of conventional variance estimators in linear regression with high-dimensional fixed effects under a drift in which both the proportional fixed-effect dimension $ρ_n = d_{K_n}/n \to ρ\in [0,1)$ and the residual treatment variance $τ_n^2 = nQ_{K_n} \to τ^2 \in (0, \infty]$ are non-degenerate. Three findings emerge. First, under strict exogeneity and conditional homoskedasticity, the Cattaneo--Jansson--Newey-corrected $t$-statistic is asymptotically exact for any $τ^2 > 0$: there is no Stock--Yogo-style threshold in $τ^2$. Second, the Eicker--White HC0 estimator is biased downward by a fixed factor $(1-ρ)$, producing over-rejection that grows with saturation. Third, HC3 over-corrects in the opposite direction by a factor $1/(1-ρ)$. The leave-one-out estimator (HC2) removes the first-order leverage distortion and is asymptotically exact under homoskedasticity or design-balanced heteroskedasticity; under general heteroskedasticity with non-uniform leverage, HC2 retains an additional bias of order $ρ|μ- ω^2|$ that we characterize. An empirical application to Piotroski F-Score returns in CEE markets illustrates the predicted variance hierarchy in real data.2026-07-06T15:32:08ZSubmitted to The Econometrics Journal for considerationStanisław M. S. Halkiewiczhttp://arxiv.org/abs/2511.00944v3Empirical Characteristic Function Method for Leverage Effect and Volatility of Volatility: Estimation and Feasible Inference2026-07-06T12:23:54ZWe develop jump-robust estimators of the leverage effect and volatility of volatility using high-frequency data. Our construction begins with a spot volatility estimator based on the empirical characteristic function of high-frequency increments. This method can mitigate the contamination from jumps, which can be of infinite variation. We then construct estimators of the leverage effect and volatility of volatility and correct for the bias induced by spot volatility estimation. We establish consistency and central limit theorems under conditions that allow greater jump activity than existing methods. We also develop consistent estimators of the asymptotic variances, making the limiting results feasible for statistical inference. Simulation studies demonstrate the improved finite-sample performance of the proposed estimators, particularly in the presence of infinite variation jumps. An empirical application provides evidence of nonzero leverage effect and volatility of volatility, when the jump activity is intensive.2025-11-02T14:05:48ZQiang LiuZhi LiuGuangren YangWang Zhouhttp://arxiv.org/abs/2607.04885v1Geometric Control of Decisions' Affordability2026-07-06T10:09:34ZThis paper studies the performance of data-driven decisions from a geometric perspective. A policymaker learns from an innovated donor population to decide whether to innovate groups in a distinct target population, and must compensate for any mistake. I introduce certification: an estimator yields certified decisions when it controls the probability of a mistake, whenever intervention effects are sufficiently large in magnitude. First, I show that certification implies a bound on worst-case compensation. Then, I study matching estimators with positive weights and show that, in a large-sample regime, affordability by certification becomes a purely geometric problem. I prove that a Delaunay interpolant, whose properties are well-known from results in computational geometry, delivers the best affordability guarantee. Finally, I show how this result can be leveraged to guide donor-data collection plans to bring worst-case compensation cost below a target level. I illustrate the gains of adopting this geometric point of view in targeting and collection plans with a semi-synthetic empirical application in development economics.2026-07-06T10:09:34ZGiacomo Opocherhttp://arxiv.org/abs/2312.00590v7Inference on common trends in functional time series2026-07-06T08:20:31ZWe study statistical inference on unit roots and cointegration for time series in a Hilbert space. We develop statistical inference on the number of common stochastic trends embedded in the time series, i.e., the dimension of the nonstationary subspace. We also consider tests of hypotheses on the nonstationary and stationary subspaces themselves. The Hilbert space can be of an arbitrarily large dimension, and our methods remain asymptotically valid even when the time series of interest takes values in a subspace of possibly unknown dimension. This has wide applicability in practice; for example, to cointegrated vector time series that are either high-dimensional or of finite dimension, to high-dimensional factor models that include a finite number of nonstationary factors, to cointegrated curve-valued (or function-valued) time series, and to nonstationary dynamic functional factor models. To illustrate our methods, we include two empirical examples.2023-12-01T13:55:12ZMorten Ørregaard NielsenWon-Ki SeoDakyung Seonghttp://arxiv.org/abs/2606.21224v2Uniform Confidence Bands for Infinite-Dimensional Partially Identified Parameters2026-07-06T02:05:20ZInfinite-dimensional parameters are ubiquitous in empirical economics. This paper develops an Imbens--Manski--Stoye type confidence band for infinite-dimensional partially identified parameters. In particular, we propose multiplier bootstrap-based construction of a uniform confidence band. By employing approximation theorems for suprema of non-centered empirical processes indexed by possibly non-Donsker classes \citep{chernozhukov2016empirical}, we confirm the uniform validity of the proposed procedure.2026-06-19T08:42:17ZShunsuke ImaiYuta Okamotohttp://arxiv.org/abs/2607.04567v1Causal Overlap Effects: A Cumulative Fixed Effect Approach2026-07-06T00:34:36ZSocial scientists often ask about the effect of increasing one's duration of exposure to a social context on one's outcomes, i.e. the overlap effect. Past studies adopted a unidimensional treatment effect framework to estimate the effect of overlap, imposing important restrictions. In this paper, we propose a new causal framework of multidimensional treatments where the overlap effects include both the duration and the content of overlap, under which, for instance, the grandparent overlap effect is defined as the union of all causal effects of a grandparent's observed and unobserved characteristics (i.e., the content) on the grandchild across their shared life course (i.e., the duration). The multidimensional framework allows for a more flexible and context rich approach to effect heterogeneity, where unobserved contextual characteristics play two roles as unobserved confounders and as integral components of overlap effects -- overlap effects in this framework are not easily estimated with conventional fixed effects estimation. Hence, we develop a new cumulative fixed effects (CFE) approach that can estimate a range of interesting heterogeneous causal overlap effects from three-wave individual panel data. We show that the CFE approach is unbiased even in highly non-linear simulations, and we discuss assumptions and extensions.2026-07-06T00:34:36ZJingying HeFelix Elwerthttp://arxiv.org/abs/2507.12693v2Placebo Discontinuity Design2026-07-06T00:09:57ZStandard regression discontinuity design (RDD) models rely on the continuity of expected potential outcomes at the cutoff. The standard continuity assumption can be violated by strategic manipulation of the running variable, which is realistic when the cutoff is widely known and when the treatment of interest is a social program or government benefit. In this work, we identify the treatment effect despite such a violation, by leveraging a placebo treatment and a placebo outcome. We introduce a local instrumental variable estimator. Our estimator decomposes into two terms: the standard RDD estimator of the target outcome's discontinuity, and a new adjustment term based on the placebo outcome's discontinuity. We show that our estimator is consistent, and we justify a robust bias-corrected inference procedure. Our method expands the applicability of RDD to settings with strategic behavior around the cutoff, which commonly arise in social science.2025-07-16T23:59:24ZRahul SinghMoses Stewarthttp://arxiv.org/abs/2606.06253v2When the Scaffold Stays On: AI, Practice Style, and Screening in Elite Skill Formation2026-07-05T21:54:49ZGenerative AI raises short-term productivity by completing tasks that learners would otherwise practice on their own. Whether this exchange erodes frontier skill depends on the mode of use: substitute-users let AI stand in for deliberate practice and fail to develop skill, while complement-users use it to accelerate skill development. For institutions that train and certify talent, the design question is not whether to allow AI but how to govern the mode of its use. We ask whether AI-prohibited evaluation gates can separate the two modes. In elite competitive programming, the International Collegiate Programming Contest (ICPC) and the International Olympiad in Informatics (IOI) prohibit AI under in-person proctoring, with qualification-round entry, whereas Codeforces (CF) practice is unproctored and open to all. From CF submission histories we build an AI-prompt signature, more first-attempt acceptances, fewer attempts, fewer debugging retries, consistent with AI-assisted practice. CF practice has shifted toward this signature across entry cohorts spanning two AI rollouts. In CF contests, a stronger signature predicts smaller rating gains for users with no ICPC-IOI affiliation, but not for those who qualified. Inside the AI-prohibited ICPC environment, a shift toward AI-style practice predicts higher non-AI-aided scores for AI-era entrants. The same signature carries opposite signs across the two environments, exactly the pattern a type-separating gate predicts. The message is constructive: AI-style practice is compatible with frontier skill; the erosion risk links to the substitute mode; and that mode is separable by gates standard at credential boundaries, from medical and legal boards to professional certification.2026-06-04T14:54:44Z61 pages, 4 figuresSong Yaohttp://arxiv.org/abs/2504.03228v11Weak instrumental variables due to ignored nonlinearities in panel data: A Super Learner Control Function estimator2026-07-05T20:52:25ZA triangular structural panel data model with additive separable individual-specific effects is used to model the causal effect of a covariate on an outcome variable when there are unobservable confounders with some of them time-invariant. In this setup, a linear specification for the reduced-form equation might be problematic when the conditional mean of the endogenous covariate and the instrumental variables is nonlinear in the population. The reason is that ignoring the nonlinearity could lead to weak instruments (instruments are weakly correlated with the endogenous covariate) due to misspecification as shown using a generalized concentration parameter for panel data. As a solution, we propose a triangular simultaneous equation model for panel data with additive separable individual-specific fixed effects composed of a linear structural equation with a nonlinear reduced form equation. The parameter of interest is the structural parameter of the endogenous variable. The identification of this parameter is obtained under the assumption of available exclusion restrictions and using a control function approach. We provide an estimator that we call Super Learner Control Function estimator (SLCFE). The estimation procedure is composed of two main steps and cross-fitting. First, we estimate the control function using a super learner. In the following step, we use the estimated control function to control for endogeneity in the structural equation. Cross-fitting is done across the individual dimension. The estimator is consistent and asymptotically normal achieving a parametric rate of convergence. We show that the SLCF estimator differs from both the plug-in IV estimator and a naive plug-in 2SLS estimator, with the former not being consistent without cross-fitting, and the latter not being consistent even with cross-fitting.2025-04-04T07:22:18ZMonika Avila-Marquezhttp://arxiv.org/abs/2607.04468v1IMF Programs and Growth: A Source-Informed Robustness Reanalysis2026-07-05T19:28:14ZThis article reassesses the meta-analytic evidence on the effect of International Monetary Fund programs on economic growth. The point of departure is the influential meta-analysis by Balima and Sokolova, which assembles 994 estimates from 36 studies and reports a positive average effect with substantial heterogeneity. The reanalysis presented here imposes four stricter requirements. It treats the study, rather than the reported estimate, as the primary inferential unit; it uses a source-informed classification of causal credibility; it models within-study dependence through study aggregation, correlated-effects sensitivity, multilevel CR2 inference, and robust variance estimation; and it evaluates publication-bias sensitivity through Egger, PET, PEESE, WAAP-like top-precision analysis, p-curve diagnostics, trim-and-fill, and exploratory selection models. The central result is not that IMF programs reduce growth everywhere, nor that the true effect is exactly zero. The result is narrower and stronger: the positive average effect in the aggregate literature is not robust once dependence, publication selection, heterogeneity, and credibility of identification are treated as first-order concerns. In the most defensible specifications, the average effect is statistically indistinguishable from zero, while equivalence to a substantively negligible effect is only partially supported and depends on the chosen equivalence bound.2026-07-05T19:28:14ZRicardo Alonzo Fernández Salguerohttp://arxiv.org/abs/2606.16230v2Semiparametric Dynamic Logit Model with Endogenous Networks2026-07-05T19:10:18ZThis paper develops identification and estimation methods for a semiparametric dynamic logit model in which a binary outcome depends on observed covariates, the lagged outcome, and an unknown function of a latent social characteristic that also governs the formation of social ties. The unobserved characteristic is allowed to vary across agents and over time, and the network formation process is left completely unspecified. Identification combines three elements: conditional likelihood arguments that exploit the logistic structure, network-type matching that eliminates the unknown social influence function by comparing agents whose observed linking behavior reveals identical latent characteristics, and local temporal smoothing that handles the interaction between dynamics and time-varying unobserved heterogeneity. A kernel-weighted conditional maximum likelihood estimator is proposed, and its consistency and asymptotic normality are established at the $\sqrt{n}$ rate. Monte Carlo simulations show that the estimator substantially reduces the bias present in naive and control-function approaches across a range of network formation models and achieves close to nominal coverage at moderate sample sizes. The method is applied to longitudinal data on adolescent smoking and friendship networks from the Glasgow Teenage Friends and Lifestyle Study. An extension to ordered outcomes is developed using composite conditional maximum likelihood.2026-06-15T05:24:34ZBrice Romuald Gueyap Koungahttp://arxiv.org/abs/2512.19824v4Regret in Treatment Choice when Welfare Varies with an Uncertain Event: The Prediction-Threshold Problem2026-07-05T19:02:13ZWe study maximum regret (MR) of binary treatment choice in a population with observed covariates x, when welfare varies with an uncertain binary event. We consider decision making with plug-in probabilistic predictions of the event and pre-specified decision thresholds, which we term the prediction-threshold problem. The optimal treatment for persons with covariate value x is B if the conditional probability P(y=1|x) of a binary outcome y exceeds a particular x-specific threshold and is A otherwise. This structure is common in medical decision making and other contexts. Plug-in prediction uses data to estimate P(y|x) and acts as if the estimate is accurate. However, plug-in prediction is often performed with misspecified prediction models and conventional x-invariant thresholds. We use a combination of algebraic and computational analysis of limit and finite-sample MR to demonstrate how MR depends on the prediction model, the state space, and the thresholds used to choose treatments.2025-12-22T19:30:53ZJeff DominitzCharles F. Manskihttp://arxiv.org/abs/2607.04380v1Properties of the Conditional Likelihood Ratio Test under Discrete Approximation2026-07-05T16:12:38ZThe conditional likelihood ratio (CLR) test is a valuable tool for inference under weak identification, with appealing theoretical properties in both linear and non-linear settings. Its implementation nevertheless requires minimizing a non-convex objective function, a difficulty long recognized even in the linear IV setting. While grid-based methods that provide a practical approximation may perform well in particular designs, such procedures do not guarantee that the resulting test preserves the theoretical properties of the CLR test uniformly across a class of data-generating processes. This paper examines the implementation challenges and their consequences for test size and power. In the linear IV settings, we contrast the grid-based method with the polynomial approach of Moreira, Newey, and Sharifvaghefi(2024), which guarantees global minimization and aligns computation with the theoretical properties of the CLR test.2026-07-05T16:12:38ZMarcelo J. MoreiraMahrad Sharifvaghefihttp://arxiv.org/abs/2607.04278v1Deep Learning for Dynamic Programming with Recursive Utility2026-07-05T12:43:03ZWe propose the first deep learning algorithm, the Certainty Equivalent Learning (CEL) algorithm, for solving high-dimensional discrete-time dynamic programming problems with recursive utility. Dynamic programming with recursive utility is numerically challenging because the recursive utility does not have an explicit representation and the Bellman equation contains a certainty equivalent that is difficult to evaluate. The CEL algorithm learns this certainty-equivalent value directly with neural networks and jointly approximates value functions, policy functions, and certainty-equivalent functions. The CEL algorithm is mesh-free and simulation-based, allowing high-dimensional state and control spaces, and does not rely on Euler equations, first-order conditions, or differentiability of the state transition function. The CEL algorithm also works for dynamic programming problems with expected utility as expected utility is a special case of recursive utility. We apply the CEL to discounted linear exponential quadratic Gaussian control, small-noise robust control, Epstein-Zin DSGE, and multivariate strategic asset allocation problems. Compared with closed-form and VFI-based benchmarks, the CEL delivers accurate value and policy approximations, remains effective in high-dimensional problems, achieves accuracy comparable to VFI in the small-noise robust-control case, and produces out-of-sample Bellman errors and Euler or first-order residuals that are in the range from 1.0e-4 to 1.0e-3 for most problems.2026-07-05T12:43:03Z93 pages, 44 figuresXianhua PengWu Guo