https://arxiv.org/api/XnFAReHlGO4LFYnaBLNQZh/cGd42026-09-11T20:44:48Z58314515http://arxiv.org/abs/2609.05883v1Fixed-smoothing Uniform Inference for Quantile Regression2026-09-05T05:11:39ZThis paper develops fixed-smoothing (fixed-b, fixed-K) inference methods for time-series quantile regression that are robust to heteroskedasticity and autocorrelation. Our approach is uniformly valid over quantile levels and accounts for dependence both over time and across quantiles. It enables the construction of uniform confidence bands, Wald, and Sup-t tests for joint hypotheses, and tests of shape restrictions, providing a unified framework for assessing heterogeneity in quantile effects. A key challenge is that, under weak dependence, uniform inference for quantile regression processes is generally non-pivotal because the limiting distributions depend on the long-run covariance structure across quantiles. To address this issue, we develop two complementary approaches. The uniform-in-$τ$ method estimates the covariance structure and simulates the non-pivotal limiting distribution. For certain tests involving a finite collection of quantile levels, the stack-Wald method delivers pivotal fixed-smoothing inference. We establish the asymptotic validity of both approaches. Simulation results show that the proposed methods substantially improve size control relative to existing HAC-based procedures while maintaining good power. An application to predictive quantile regressions for stock returns reveals substantial heterogeneity in predictive effects across both quantiles and forecast horizons.2026-09-05T05:11:39ZKaicheng ChenAntonio F. GalvaoSeunghwa RhoTimothy J. VogelsangJungmo Yoonhttp://arxiv.org/abs/2609.05792v1Scalable Clustered Network Connectedness with Control Variables: Theory and Application to Global Banking2026-09-05T01:06:20ZWe extend the clustered connectedness framework of Buchwalter, Diebold and Yilmaz (2026) in two complementary directions that improve the robustness and interpretability of cross-cluster connectedness. First, we develop a diagnostic for residual ordering sensitivity by characterizing the distribution of cluster-level net connectedness across all admissible identification orderings and, in particular, by pairing first- and last-position orderings while holding fixed the relative ordering of all other clusters. Second, we introduce a dedicated cluster of control variables to absorb variation associated with observed common macro-financial factors while preserving the computational scalability of the clustered framework. The control cluster is fixed first, and bank innovations are residualized with respect to it before the remaining bank clusters are permuted and orthogonalized as usual, leaving the number of admissible bank-cluster identification strategies unchanged. Under the maintained recursive assumption that control-cluster innovations are contemporaneously exogenous to bank-cluster innovations, the remaining cross-cluster connectedness among the bank clusters can be interpreted as bank-to-bank transmission net of those observed common-factor shocks. We apply the methodology to seventy-one global banks grouped into seven regional clusters over 2003--2024. The treatment of common macro-financial factors materially affects both system-wide cross-group connectedness and cluster-level net positions. Placing the controls in a dedicated first cluster also substantially reduces paired first-versus-last ordering sensitivity across all seven bank clusters, with especially large reductions for the United States and the European clusters.2026-09-05T01:06:20ZBastien BuchwalterFrancis X. DieboldKamil Yilmazhttp://arxiv.org/abs/2412.02767v5Endogenous Heteroskedasticity in Linear Models2026-09-04T17:33:33ZLinear regressions with endogeneity are widely used to estimate causal effects. This paper studies a framework that involves two common practical issues: endogeneity of the regressors and heteroskedasticity that depends on endogenous regressors, i.e., endogenous heteroskedasticity. To address the inconsistency of the two-stage least squares estimator in this scenario, and recover the causal parameters of interest, we develop a framework for practical estimation and inference based on the control function approach allowing for discrete and continuous regressors. In particular, we suggest a simple two-step estimation procedure. We establish the limiting properties of the estimator, namely, consistency and asymptotic normality. In addition, we develop practical valid inference methods by proposing an estimator for the asymptotic variance-covariance matrix, and formally establishing its consistency. Monte Carlo simulations provide evidence on the finite-sample performance of the proposed methods and evaluate different implementation strategies. We revisit an empirical application on job training to illustrate the methods.2024-12-03T19:09:48ZJavier AlejoAntonio F. GalvaoJulian Martinez-IriarteGabriel Montes-Rojashttp://arxiv.org/abs/2609.05372v1Does p-Hacking Mitigate or Exacerbate the Effects of Publication Bias?2026-09-04T17:24:40ZThis paper studies the effects of p-hacking on the bias of published estimates when papers with statistically significant results are selectively published. We show that fast p-hacking---actions that lead to large changes in p-values---always exacerbates the bias from selective publication. On the other hand, slow p-hacking---actions that lead to small changes in p-values---exacerbates bias when selection is weak, but mitigates it when selection is strong. In a model featuring both types of p-hacking, we show that a normality assumption identifies the true distribution of effects as well as the counterfactual mean that would obtain under selective publication without p-hacking. Applying the model to meta-analyses on the effects of behavioral nudges and development aid, we find suggestive evidence that both mitigation and exacerbation can arise in practice.2026-09-04T17:24:40ZYong CaiAgathe PernoudBoli Xuhttp://arxiv.org/abs/2601.12896v3Quantitative Methods in Finance2026-09-04T14:41:41ZThese lecture notes provide a comprehensive introduction to Quantitative Methods in Finance (QMF), designed for graduate students in finance and economics with heterogeneous programming backgrounds. The material develops a unified toolkit combining probability theory, statistics, numerical methods, and empirical modeling, with a strong emphasis on implementation in Python. Core topics include random variables and distributions, moments and dependence, simulation and Monte Carlo methods, numerical optimization, root-finding, and time-series models commonly used in finance and macro-finance. Particular attention is paid to translating theoretical concepts into reproducible code, emphasizing vectorization, numerical stability, and interpretation of outputs. The notes progressively bridge theory and practice through worked examples and exercises covering asset pricing intuition, risk measurement, forecasting, and empirical analysis. By focusing on clarity, minimal prerequisites, and hands-on computation, these lecture notes aim to serve both as a pedagogical entry point for non-programmers and as a practical reference for applied researchers seeking transparent and replicable quantitative methods in finance.2026-01-19T09:50:52Z597 pages. v3: new sections on writing a project, Git/GitHub and Python setup; companion primer Mathematics for Finance: A Self-Contained PrimerEric Vansteenberghehttp://arxiv.org/abs/2606.22391v2On the Asymptotic Inadmissibility of Double Machine Learning Estimators Under Structure-Agnostic Models2026-09-04T13:26:12ZStructure-agnostic (SA) models introduced by Balakrishnan et al. (2026) aim to reflect the general lack of knowledge of structural assumptions on data-generating laws such as smoothness or sparsity in practice. Roughly speaking, SA models restrict the observed-data generating law to be in some rn-neighborhood of (black-box machine learning) estimates, treated as given and fixed, where rn encodes the convergence rates of the estimates to the truth. Under SA models, Balakrishnan et al. (2026) show that the popular Double Machine Learning (DML) estimators for three functionals, the quadratic functional in the Gaussian sequence model, the quadratic density integral functional and the expected conditional covariance, are minimax. However, minimax estimators may be inadmissible. In this paper, we show that, for the first two of the three functionals, the DML estimator is asymptotically inadmissible under the SA model. In particular, we show that these two functionals fall into a class of functionals, which we refer to as the monotone bias class. For this class, we exhibit second-order (U-statistic) estimators, which asymptotically dominate DML estimators, under the SA model. These second-order estimators are empirical higher-order influence function (HOIF) estimators introduced in Liu et al. (2017). Furthermore, the empirical HOIF estimator, like the DML estimator, is minimax for the third functional (the expected conditional covariance), although neither asymptotically dominates the other.2026-06-21T08:35:22Z25 pages; The current version of the paper strengthens the pointwise asymptotic (in)admissibility results of the original arXiv version to uniformLin LiuRajarshi MukherjeeJames M Robinshttp://arxiv.org/abs/2609.04994v1Identification and Estimation of Intergenerational Income Mobility Measures2026-09-04T11:04:54ZMeasuring the intergenerational transmission of lifetime economic status is complicated by researchers often only observing snapshots of income at specific ages. Consequently, standard practice estimates intergenerational mobility using income averages, introducing life-cycle bias that compromises reliability and comparability across studies, time, and place. I develop a missing data framework that exploits available income data and observable characteristics to eliminate life-cycle bias. This method combines nonparametric identification with Neyman-orthogonal moments to construct debiased machine learning estimators for intergenerational income mobility measures under plausible missing-at-random and testable independence assumptions. I apply this framework to estimate the intergenerational elasticity for the U.S. using the Panel Study of Income Dynamics across birth cohorts from 1954 to 1977 with rolling 10-year windows. While existing approaches estimate values between 0.41 and 0.54, the proposed method yields substantially higher estimates ranging from 0.6 to 0.7, averaging 0.64. These results align closely with recent evidence using long time averages over mid-career periods, reinforcing high U.S. intergenerational persistence.2026-09-04T11:04:54ZAlejandro Puerta-Cuartashttp://arxiv.org/abs/2502.09740v3High-dimensional censored MIDAS logistic regression for corporate survival forecasting2026-09-04T08:13:23ZThis paper addresses the challenge of forecasting corporate distress, a problem marked by three key statistical hurdles: (i) right censoring, (ii) high-dimensional predictors, and (iii) mixed-frequency data. To overcome these complexities, we introduce a novel high-dimensional censored MIDAS (Mixed Data Sampling) logistic regression. Our approach handles censoring through inverse probability weighting and achieves accurate estimation with numerous mixed-frequency predictors by employing a sparse-group penalty. We establish finite-sample bounds for the estimation error, accounting for censoring, MIDAS approximation error, and heavy tails. For statistical inference, we develop a de-sparsified version of the proposed penalized estimator and establish its asymptotic theory, which enables valid statistical inference in high-dimensional settings with censoring. We show that censoring induces a nonstandard variance structure for the de-sparsified estimator, a feature that, to the best of our knowledge, has not been studied in the existing literature. The superior performance of the method is demonstrated through Monte Carlo simulations. Finally, we present an extensive application of our methodology to predict the financial distress of Chinese-listed firms and to identify covariates that are statistically significant for predicting distress. Our novel procedure is implemented in the R package \texttt{Survivalml}.2025-02-13T19:51:36ZWei MiaoJad BeyhumJonas StriaukasIngrid Van Keilegomhttp://arxiv.org/abs/2607.19925v2Efficient difference-in-differences estimation under partial interference with incremental propensity score policies2026-09-04T04:45:49ZThis paper develops efficient difference-in-differences (DID) estimation under partial interference with a cluster incremental propensity score (CIPS) policy. We define direct and spillover average treatment effects on the treated, establish their identification, and derive their efficient influence functions, from which we construct a cross-fitted estimator. Simulations evaluate its finite-sample performance. An application to China's New Rural Pension Scheme recovers the reduction in farmwork among pension recipients reported by the original county-level analysis, separately estimates a within-household spillover alongside the direct effect, and traces both effects across the policy parameter.2026-07-22T08:58:38ZJunjie LiYukitoshi Matsushitahttp://arxiv.org/abs/2605.05404v3Causal State-Dependent Local Projections2026-09-03T20:53:56ZState-dependent local projections (LPs) are widely used to study how causal effects vary as a function of economic states, but shock exogeneity alone does not identify this response function. We show that identification follows when the underlying conditional mean is linear in the shock with a state-dependent coefficient, a condition satisfied in canonical micro-macro environments, including first-order perturbation solutions of heterogeneous-agent and macro-finance models. Even then, standard linear-interaction LPs generally recover only a projection of the response function, motivating LPs with nonparametric state dependence. We develop a sieve estimator and establish pointwise and uniform inference for micro-macro panels, where a distinctive challenge is that the estimator can converge at different rates across the state space. Applied to firm investment, the method uncovers a hump-shaped response to monetary policy shocks and shows that standard linear-interaction LPs substantially understate the aggregate role of financial heterogeneity.2026-05-06T19:41:03Z43 pages for main paper, 35 pages for the Online Appendix, 9 pages for the Technical Note; 6 figures, 6 tablesJoel M. DavidRaffaella GiacominiXiyu JiaoWeining Wanghttp://arxiv.org/abs/2609.04356v1Blockchain-Enabled Secure Logging for Fiscal Electronic Mechanisms: Evaluation of the Greek eSEND and myDATA Tax Systems2026-09-03T18:19:44ZThis paper analyzes the implementation of blockchain-based integrity mechanisms in Greek Fiscal Electronic Mechanisms (FEMs) and the central tax information system eSEND. The study examines the cryptographic architecture of fiscal devices, including Electronic Cash Registers, Fiscal Printers, Fiscal Signing Machines, and FEMAS devices, which implement double or triple hash-chain structures to ensure transaction immutability. The transmission protocol between fiscal devices and the central database is also evaluated with respect to encryption, sequential validation, and blockchain verification. In contrast, the architecture of Electronic Invoicing Provider Services and the myDATA central platform is analyzed, highlighting the absence of blockchain-based integrity guarantees. The comparison demonstrates that hardware-based fiscal mechanisms provide stronger guarantees for transaction completeness and tamper resistance than purely software-based invoicing infrastructures. The findings highlight architectural weaknesses in the current e-invoicing framework and propose improvements for ensuring transaction integrity in digital tax ecosystems.2026-09-03T18:19:44ZPanagiotis MavridisAnargyros BaklezosChristos Nikolopouloshttp://arxiv.org/abs/2609.04136v1Natural Disasters and the Nonprofit Sector2026-09-03T17:30:52ZWhen natural disasters strike, individuals, communities, and even entire countries can suffer. Researchers have studied the impacts of disasters on various factors of interest, from mental health, to poverty, to economic activity. However, the impact of disasters on the nonprofit sector is understudied despite the nonprofit sector's perhaps surprising role in local or national economies as well as its role in disaster response and recovery. Thus, we study the effect of natural disaster damage on different county-level nonprofit outcomes using a panel dataset spanning 1991 to 2021 and causal inference methods tailored to panel data. Contrary to prior work, which found small but positive associations between disaster damage and nonprofit revenue or assets, we find no evidence of a causal effect.2026-09-03T17:30:52Z47 pages, 10 figures, 5 tablesMayleen Cortez-Rodriguezhttp://arxiv.org/abs/2608.06053v3Breakdown Reliability for Saturated Fixed-Effect Inference2026-09-03T15:42:13ZFixed-effect saturation does not itself distort conventional inference, but classical measurement error does. Under a local noise drift $σ_ν^2=c^2/n$, the FE-OLS $t$-statistic converges to a non-central normal; saturation contributes a common $\sqrt{1-ρ}$ scaling rather than preferentially destroying signal or noise. Inverting the size distortion gives a Stock--Yogo-style critical value. Self-consistency of the within-reliability-corrected pilot yields a breakdown reliability $λ^{\dagger}=|t|/(|t|+η^{\dagger})$ --- the minimum within reliability at which conventional inference retains nominal size within the chosen tolerance --- computable from the reported $t$-statistic alone and algebraically $ρ$-free conditional on it; $η^{\dagger}\approx0.65$ at $5\%$ size and a 5-point tolerance. Replacing $|t|$ by $|t|+z_{1-γ_β}$ gives a certified breakdown reliability; with a lower-reliability bound whose coverage error is $γ_λ$, false certification is at most $γ_β+γ_λ$. Under a checkable projection-compatibility condition, a cluster-level score CLT and consistency of the Arellano variance estimator in the many-fixed-effect regime justify applying the same map to the reported cluster-robust $t$-statistic; clustering can reverse a verdict. In a saturated democracy--growth panel, aggregate V-Dem polyarchy is certified at $γ_β=0.05$, conditional on the supplied measurement model, while its judicial-constraints sub-index is flagged under i.i.d.\ and clustered standard errors. In a twin-pair wage design, the specification is flagged under both independent and correlated reporting-error models, although implied coverage of the nominal-$95\%$ interval ranges from $8\%$ to $68\%$. The diagnostic covers classical error in a continuous regressor, not binary-treatment misclassification.2026-08-06T13:59:49ZStanisław M. S. Halkiewiczhttp://arxiv.org/abs/2609.03227v1Randomization Inference for Matched Pairs with Binary Outcomes2026-09-03T00:03:35ZWe give an exact randomization-based confidence set for the average treatment effect (ATE) in matched-pair studies with a binary outcome, requiring neither monotonicity nor any distributional assumption beyond the within-pair coin flip. At its core is an analytic solution to the worst-case allocation of attributable effects: two binomial-symmetry lemmas identify the pattern hardest to reject as a single boundary corner, so testing null hypotheses needs no integer program and no numerical search. Inverting the test via binary search yields a prediction set for the attributable effect in O(log S) Binomial tail calculations; the Bonferroni proposition of Rigdon and Hudgens (2015) produces the ATE confidence set at the same computational cost. The same corner extends without further machinery to a sensitivity analysis for matched observational studies under Rosenbaum's $Γ$-model. A simple formula for the design sensitivity illuminates when an observational study can hope to provide evidence for an effect.2026-09-03T00:03:35ZBob Wilsonhttp://arxiv.org/abs/2606.06253v4When the Scaffold Stays On: AI, Practice Style, and Screening in Elite Skill Formation2026-09-02T17:03:25ZGenerative AI raises short-term productivity by completing tasks learners would otherwise practice on their own. Whether this exchange erodes frontier skill depends on the mode of use: substitute-users let AI stand in for practice and fail to develop skills, while complement-users use AI to learn faster. The modes look alike in AI-aided output, so organizations screening on that output cannot tell them apart. We ask whether the AI-prohibited evaluation gates organizations already operate can separate the modes. In elite competitive programming, ICPC and IOI contests prohibit AI under in-person proctoring, with qualification-round entry, whereas Codeforces (CF) practice and contests are unproctored and open to all. From CF practice histories we build an AI-prompt signature consistent with AI usage, more first-attempt acceptances, fewer attempts and debugging retries. CF practice has shifted toward this signature across entry cohorts spanning two AI rollouts. On CF, a stronger signature predicts smaller rating gains for users with no ICPC-IOI affiliation, but not for those who qualified. Inside the AI-prohibited ICPC environment, AI-era entrants show no skill erosion, and shifts toward AI-style practice predict higher non-AI-aided scores. One screening mechanism fits both: where the modes mix, a stronger signature flags substitute-users; among those who qualified, a strengthening signature marks adoption of the complement mode. The message is constructive: AI-style practice is compatible with frontier skill; the erosion risk links to the substitute mode; and separating the modes is a design question for the exams organizations regularly administer, from medical and legal boards to professional certification.2026-06-04T14:54:44Z65 pages, 4 figuresSong Yao