REVIEW 2 major objections 5 minor 1 cited by
Efficient Difference-in-Differences and Event Study Estimators
T0 review · 2 major / 5 minor · reviewed 2026-08-15 · deepseek-v4-flash
Pith's one-line read Under parallel trends, the paper derives the first closed-form semiparametric efficiency bounds for multi-period DiD and event-study estimators.
desk verdict The paper's core efficiency claim is not established as written: the proof inverts a singular matrix by including a redundant propensity-score moment. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The central object is the efficient influence function (EIF), derived from an equivalent representation of the DiD model as sequential conditional moment restrictions. For each $ATT(g,t)$, the EIF aggregates influence functions that use different pre-treatment periods and comparison groups, weighting them by the inverse of the conditional covariance matrices $V^*_{gt}(X)$ (single treatment date) or $\Omega^*_{gt}(X)$ (staggered adoption). These weights automatically make the estimator Neyman orthogonal and give it the smallest asymptotic variance among regular estimators. The machinery also shows why equal weighting of pre-treatment periods is generally suboptimal: it is optimal only when outcome changes are conditionally uncorrelated across periods with constant variances.
What would settle it
Simulate a short-panel data-generating process that satisfies PT-Post but violates PT-All in early pre-treatment periods (for example, group-specific linear trends starting two periods before treatment), and compute the empirical coverage of the proposed efficient estimator over many replications. If its coverage drops substantially below the nominal level while a PT-Post estimator maintains coverage, the practical claim that efficiency is attainable without bias under the stated assumptions fails.
Extended reading notes
Core claim
Under parallel trends that hold across all groups and all periods (Assumption PT-All), the paper shows that the multi-period DiD model is nonparametrically overidentified: the observed distribution satisfies more moment restrictions than are needed to pin down the causal parameters. Its central result is a closed-form efficient influence function for each $ATT(g,t)$, equal to a covariance-weighted average of influence functions based on different pre-treatment baselines and comparison groups, with weights given by inverse conditional covariance matrices of outcome changes. The paper then proves that plug-in estimators based on this EIF are consistent, asymptotically normal, and achieve the semiparametric efficiency bound, so no regular estimator under the same assumptions can have smaller asymptotic variance. For event-study parameters the same EIF delivers efficiency through linear aggregation of the $ATT(g,t)$ results.
Load-bearing premise
The load-bearing premise is that parallel trends hold across every group and every pre-treatment period (Assumption PT-All); if pre-treatment outcome trends differ across groups, the efficient estimator that aggressively weights all pre-treatment periods can be biased despite its precision.
Editorial extensions
If this is right
- Existing heterogeneous DiD and event-study estimators, including two-way fixed effects, never-treated and not-yet-treated comparison-group estimators, and imputation estimators, generally fail to attain the semiparametric efficiency bound because they weight pre-treatment periods equally, use only the last pre-treatment period, or impose auxiliary assumptions on serial correlation.
- Under PT-All, achieving efficiency requires non-uniform, covariate-dependent weights across pre-treatment periods and comparison groups; equal weights are optimal only under knife-edge conditions.
- The proposed plug-in estimators are consistent, asymptotically normal, and attain the closed-form efficiency bound, with simulation and empirical evidence of RMSE and confidence-interval length reductions often exceeding 40%.
- The overidentified structure yields a Hausman-type test that compares the efficient PT-All estimator with the just-identified PT-Post estimator, offering a formal check on whether using all pre-treatment periods is warranted.
- Under the weaker PT-Post assumption, the multi-period model collapses to a just-identified two-period DiD, whose efficiency bound reduces to the known two-period result.
Reading between the lines
- The closed-form weights could be repurposed as a diagnostic: plotting them shows which pre-treatment periods and comparison groups carry the most information, which may guide robustness checks or data collection in applied work.
- The same EIF method likely extends to other panel causal designs with sequential moment restrictions, such as instrumented DiD, changes-in-changes, and treatments that turn on and off, where analogous overidentification and non-uniform optimal weights should appear.
- If PT-All holds only approximately, the efficient estimator's aggressive weighting of many pre-treatment periods could amplify bias from mild pre-trend violations; an adaptive estimator that trades off this bias against variance would be a natural next step.
- The paper's efficiency benchmark also provides a way to rank existing estimators directly across applications, since any estimator whose influence function differs from the EIF is provably less efficient under PT-All.
Signed reviews
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. This paper develops semiparametric efficiency theory for difference-in-differences (DiD) and event study (ES) estimators under parallel trends and no-anticipation assumptions, allowing for staggered treatment timing and covariates. The authors provide an equivalent characterization of the DiD potential outcome model through sequential conditional moment restrictions, derive closed-form efficient influence functions (EIFs) for ATT and ES parameters, and propose plug-in estimators that are claimed to attain the semiparametric efficiency bound. The theoretical results are complemented by calibrated simulations and an empirical application to hospitalization and out-of-pocket spending, showing substantial precision gains relative to existing DiD estimators.
Significance. If the theoretical results are correct, this is an important contribution: it appears to be the first semiparametric efficiency analysis for multi-period DiD and ES estimators with heterogeneous treatment effects and no parametric functional-form assumptions. The observational-equivalence characterization in Lemmas 3.1 and 3.2 is a useful conceptual contribution, and the closed-form EIF provides practical guidance on how to weight pre-treatment periods and comparison groups. The simulation study is extensive and the empirical illustration is well executed, giving the paper substantial practical relevance. However, the proofs of the main theorems currently contain a technical gap that must be fixed before the efficiency claims can be regarded as established.
major comments (2)
- The vector ρ2(W,h) in the proof includes both G_g − p_g(X) and G_8 − p_8(X) as separate entries. Since G_8 = 1 − G_g and p_8 = 1 − p_g, these two entries are perfectly collinear, so the conditional covariance matrix Σ2(X) = E[ρ2ρ2'|X] is singular. The proof nevertheless inverts Σ2(X) and, in particular, defines the block Π = [[p_g(1−p_g), −p_g p_8], [−p_g p_8, p_8(1−p_8)]], whose determinant equals p_g(1−p_g)p_8(1−p_8) − p_g^2 p_8^2 = 0. All subsequent derivations that use Π^{-1} (or equivalently Σ2(X)^{-1}) are therefore undefined. This is load-bearing because the closed-form EIF and the variance bound in Theorem 3.1 are obtained from these inverses. The final formula may be correct after one removes the redundant propensity-score moment, but that corrected derivation is not supplied in the manuscript.
- The same singularity appears in the staggered-treatment proof. The bottom-right block Π of Σ2(X) is the generalized propensity-score covariance matrix indexed by all groups, including the never-treated. Because sum_{g∈G} G_g = 1 and sum_{g∈G} p_g(X) = 1, the rows of Π sum to zero, so Π is singular. The proof uses Π^{-1} to conclude L(X)'Σ2(X)^{-1} = −(0, Π^{-1}) and subsequently derives the EIF for π_g; that step is invalid as written. The redundancy can be removed by dropping one group's propensity-score moment, but the paper does not provide the resulting argument. Since Theorem 4.1 establishes efficiency by showing the estimator's influence function equals the EIF from Theorem 3.2, the gap in Theorem 3.2's proof also leaves the efficiency claim in Theorem 4.1 unproven, although the consistency and asymptotic normality parts are argued directly.
minor comments (5)
- In the log-unemployment-rate DGP (row 10), EDiD has a reported bias of 0.73 (×10) at n=50 while TWFE has −0.24, and both TWFE and SDiD have lower RMSE than EDiD. The text states that 'all estimators are (nearly) unbiased when n=50'; this statement is hard to reconcile with row 10 and should be qualified.
- The heatmap color scales differ across panels (Figure 2c uses a different weight range than Figures 2a, 2b, and 2d), which makes cross-panel comparisons of the efficiency weights difficult; consider using a common scale or explicitly noting the change.
- The notation '1g−1' and '02' is used in the matrix blocks of L(X) without being defined at first use; the vectors should be defined explicitly to avoid ambiguity.
- The definition of V*_gt(X) uses Cov(Y_t − Y_j, Y_t − Y_k | G = g, X) and the analogous term for G=8; the text should clarify that these are covariances of outcome changes between the post-treatment period t and the pre-treatment periods j and k, since the notation Y_t, Y_j, Y_k may otherwise be confusing.
- The phrase 'observational equivalent' is used in the statements of Lemmas 3.1 and 3.2 but 'observationally equivalent' is used elsewhere; please standardize the spelling and grammar.
Circularity Check
No significant circularity: the efficiency bound is solved from the model's moment restrictions via standard semiparametric machinery; self-citations are not load-bearing.
full rationale
The paper's central claim—closed-form EIF and efficiency bound for DiD/ES parameters—is derived from the sequential conditional moment restrictions in Lemmas 3.1/3.2 using the general semiparametric efficiency result of Ai and Chen (2012). The target ATT(g,t) appears in the moment vector, but it enters as the parameter to be solved for; the EIF is obtained by an orthogonalization and calculus-of-variations argument, not by assuming the EIF or the estimand. The moment-restriction characterization is shown to be observationally equivalent to the DiD assumptions by an explicit construction of potential outcomes from the observed distribution, so it is not a self-definition. The EIF-based estimands in (3.5) and (3.13) are implications of the EIF having mean zero, and are algebraically the same identification identities as Lemma 2.1; no parameter is fitted and then renamed a prediction. The simulations use DGPs calibrated to external CPS and Compustat data and a real HRS application, so the precision gains are not forced by construction. The self-citations (Ai and Chen 2012; Chen and Santos 2018; Sant'Anna and Zhao 2020) are standard tools or comparators: Ai-Chen supplies a general theorem whose assumptions do not include the paper's efficiency conclusion, Chen-Santos supplies the overidentification/Hausman framework, and Sant'Anna-Zhao is a just-identified special case, not an input to the bounds. The skeptic's singular-Σ2 concern about p_g and p_8 collinearity is a proof-correctness issue (possible division by zero in the displayed inversion), not a circularity: it does not show that the EIF is equivalent to its inputs by construction. No load-bearing self-citation chain or renamed empirical pattern was found, so the derivation is self-contained for circularity purposes.
Assumptions & free parameters
assumptions (8)
- domain assumption Random sampling (Assumption S)
- domain assumption Overlap (Assumption O)
- domain assumption No anticipation (Assumption NA)
- domain assumption Parallel trends (Assumption PT-Post or PT-All)
- domain assumption Technical regularity conditions (Assumption C.1)
- standard math Ai and Chen (2012) efficiency bound theorem for sequential conditional moment restrictions
- standard math Empirical process theory (van der Vaart and Wellner, 1996)
- standard math Matrix identities: Sherman-Morrison, Schur complement, determinants of minors
Cite this review
Pith. "Pith review of Efficient Difference-in-Differences and Event Study Estimators." pith.science (2026). https://pith.science/paper/L3LEMGYS
@misc{pith2026250617729,
author = {Pith},
title = {Pith review of: Efficient Difference-in-Differences and Event Study Estimators},
year = {2026},
howpublished = {\url{https://pith.science/paper/L3LEMGYS}},
note = {Machine review of arXiv:2506.17729}
}
read the original abstract
This paper investigates efficient Difference-in-Differences (DiD) and Event Study (ES) estimation using short panel data sets within the heterogeneous treatment effect framework, free from parametric functional form assumptions and allowing for variation in treatment timing. We provide an equivalent characterization of the DiD potential outcome model using sequential conditional moment restrictions on observables, which shows that the DiD identification assumptions typically imply nonparametric overidentification restrictions. We derive the semiparametric efficient influence function (EIF) in closed form for DiD and ES causal parameters under commonly imposed parallel trends assumptions. The EIF is automatically Neyman orthogonal and yields the smallest variance among all asymptotically normal, regular estimators of the DiD and ES parameters. Leveraging the EIF, we propose simple-to-compute efficient estimators. Our results highlight how to optimally explore different pre-treatment periods and comparison groups to obtain the tightest (asymptotic) confidence intervals, offering practical tools for improving inference in modern DiD and ES applications even in small samples. Calibrated simulations and an empirical application demonstrate substantial precision gains of our efficient estimators in finite samples.
Figures
Forward citations
Cited by 1 Pith paper
-
Robust Inference for Weighted Estimands
The paper constructs minimax-bias estimators and uniformly valid confidence intervals for weighted estimands by bounding differences via parameter heterogeneity and weight distance.
Reference graph
Works this paper leans on
-
[1]
Asymptotic Efficiency of Semiparametric Two-step GMM,
Ackerberg, Daniel, Xiaohong Chen, Jinyong Hahn, and Zhipeng Liao, “Asymptotic Efficiency of Semiparametric Two-step GMM,”The Review of Economic Studies, 2014,81(3), 919–943. Ai, Chunrong and Xiaohong Chen, “Efficient estimation of models with conditional moment restrictions containin unknown functions,”Econometrica, 2003,71(6), 1795–1843. and , “Estimatio...
arXiv 2014
-
[7]
Difference-in-differences with variation in treatment timing,
Goodman-Bacon, Andrew, “Difference-in-differences with variation in treatment timing,”Journal of Econometrics, 2021,225(2), 254–277. 69 Gormley, Todd A. and David A. Matsa, “Growing Out of Trouble? Corporate Responses to Liability Risk,”Review of Financial Studies, 2011,24(8), 2781–2821. Hansen, Casper Worm and Asger Mose Wingender, “National and Global I...
work page 2021
-
[8]
A Simple Sequentially Rejective Multiple Test Procedure,
Holm, Sture, “A Simple Sequentially Rejective Multiple Test Procedure,”Scandinavian Journal of Statistics, 1979,6(2), 65–70. Imbens, Guido W. and Davide Viviano, “Identification and Inference for Synthetic Controls with Confounding,”arXiv:2312.00955,
arXiv 1979
-
[9]
When can we get away with using the two-way fixed effects regression?
Jacobson, Louis S., Robert J. LaLonde, and Daniel G. Sullivan, “Earnings losses of displaced workers,”American Economic Review, 1993,83(4), 685–709. Lal, Apoorva, “When can we get away with using the two-way fixed effects regression?,” arXiv:2503.05125 [econ.EM],
work page Pith review arXiv 1993
-
[10]
Interpreting pre-trends as anticipation: Impact on estimated treatment effects from tort reform,
Available at SSRN: http: //dx.doi.org/10.2139/ssrn.4516518. Malani, Anup and Julian Reif, “Interpreting pre-trends as anticipation: Impact on estimated treatment effects from tort reform,”Journal of Public Economics, 2015,124, 1–17. Marcus, Michelle and Pedro H. C. Sant’Anna, “The role of parallel trends in event study settings: An application to environm...
arXiv 2015
-
[11]
Higher order properties of GMM and generalized empirical likelihood estimators,
Newey, Whitney K. and Richard J. Smith, “Higher order properties of GMM and generalized empirical likelihood estimators,”Econometrica, 2004,72(1), 219–255. 70 Roth, Jonathan, “Pre-test with Caution: Event-study Estimates After Testing for Parallel Trends,” American Economic Review: Insights, 2022,4(3), 305–322. , Pedro H. C. Sant’Anna, Alyssa Bilinski, an...
work page 2004
-
[12]
How Much Should We Trust Modern Difference-in-differences Estimates?,
arXiv:2103.01280 [econ, math, stat]. Weiss, Amanda, “How Much Should We Trust Modern Difference-in-differences Estimates?,”OSF Preprints,
-
[13]
Cluster-Sample Methods in Applied Econometrics,
Wooldridge, Jeffrey M, “Cluster-Sample Methods in Applied Econometrics,”American Economic Review P&P, 2003,93(2), 133–138. , “Two-Way Fixed Effects, the Two-Way Mundlak Regression, and Difference-in-Differences Estimators,”Working Paper, 2021, pp. 1–89. 71
work page 2003
Show all 13 references
-
[2021]
Tracking the Credibility Revolution across Fields,
Goldsmith-Pinkham, Paul, “Tracking the Credibility Revolution across Fields,”arXiv:2405.20604,
-
[2022]
Revisiting Event Study Designs: Robust and Efficient Estimation,
Borusyak, Kirill, Xavier Jaravel, and Jann Spiess, “Revisiting Event Study Designs: Robust and Efficient Estimation,”Review of Economic Studies, 2024,91(6), 3253–3285. Braghieri, Luca, Ro’ee Levy, and Alexey Makarin, “Social Media and Mental Health,”American Economic Review, 2...
2024
-
[2023]
Two-Way Fixed Effects Estimators with Heterogeneous Treatment Effects,
de Chaisemartin, Cl´ ement and Xavier D’Haultfœuille, “Two-Way Fixed Effects Estimators with Heterogeneous Treatment Effects,”American Economic Review, 2020,110(9), 2964–2996. and , “Two-way fixed effects and differences-in-differences with heterogeneous treatment effects: a s...
2020
-
[2024]
Evolution vs. Creationism in the Classroom: The Lasting Effects of Science Education,
67 Arold, Benjamin W, “Evolution vs. Creationism in the Classroom: The Lasting Effects of Science Education,”The Quarterly Journal of Economics, 2024, p. qjae019. Athey, Susan and Guido W. Imbens, “Identification and Inference in Nonlinear Difference-in- Differences Models,”Ec...
2024 arXiv
-
[2025]
How much should we trust staggered difference-in-differences estimates?,
, David F. Larcker, and Charles C. Y. Wang, “How much should we trust staggered difference-in-differences estimates?,”Journal of Financial Economics, 2022,144(2), 370–395. Bertrand, Marianne, Esther Duflo, and Sendhil Mullainathan, “How Much Should We Trust Differences-In-Diff...
2022
Reviewed August 15, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.