With continuous treatment and staggered adoption, the usual DiD slope hides two distinct comparisons: the within-cohort dose–response and the treated–vs-not-yet-treated level contrast, and the pooled OLS coefficient is a convex mixture whose weight depends on dose variance and the share of not-yet-treated units, so researchers should report and test both margins separately.
Citation observations
Cumulative provider counts captured on specific dates; providers are never combined.
No provider observation is available for this paper.
Missing data, not a zero citation count.
This paper studies difference-in-differences with staggered adoption and a continuous, time-invariant dose. Each cohort-time comparison contains two margins. The level margin is the average treatment effect at realized doses. Under level parallel trends it equals the level contrast between the treated cohort and not-yet-treated controls. The response margin is the within-cohort slope of the outcome change on dose. It uses no controls, and its causal interpretation requires a response parallel trends assumption and a restriction on selection on gains. We show that the continuous-dose OLS coefficient in each cohort-time comparison is a convex combination of the response index and the level contrast per unit of mean dose, with a mixing weight that depends on the not-yet-treated share. We provide estimators of both margins, joint inference across cohort-time comparisons and event-time aggregates, and a covariate-adjusted extension. In an application to hydraulic fracturing, the level leads reject a joint zero restriction, whereas the response-index leads do not. The continuous-dose OLS coefficient draws primarily on the level margin. We report the level and response margin separately.
Summary
Main Finding
With staggered adoption and a continuous, time-invariant dose, each cohort-time DiD comparison contains two distinct margins: (i) a level margin — the cohort’s average treatment effect at its realized doses relative to zero treatment (identified under a between-group parallel-trends assumption, PT-L); and (ii) a response margin — the within-cohort OLS slope of outcome changes on dose (identified under a within-treated parallel-trends assumption, PT-R, but still subject to selection-on-gains). The usual continuous-dose TWFE/long-difference OLS coefficient is an exact convex combination of the response index and the level contrast per unit mean dose, with a mixing weight that depends on the treated-dose variance, mean dose and the share of not-yet-treated units. Thus a single pooled OLS slope can blend two different objects and vary as the untreated share changes even when cohort-level distributions are fixed.
Key Points
- Two margins (distinct objects and assumptions)
- Level margin τL,g,t = E[τc(D) | G = g]: average treatment effect at realized doses for cohort g. Observed counterpart (level contrast) δL,g,t = E[Yc | G = g] − E[Yc | G > t]. Identified under PT-L: E[Uc | G = g] = E[Uc | G > t].
- Response margin θR,g,t = Cov(D, Yc | G = g)/Var(D | G = g): within-cohort projection slope of outcome change on dose. Identified (as covariance with τ) under PT-R: E[Uc | D, G = g] = E[Uc | G = g], but may mix causal derivatives and selection-on-gains.
- Mixture decomposition (Theorem 3 / Corollary 1)
- For a cell comparing cohort g to not-yet-treated units, βTWFE,g,t = λg,t θR,g,t + (1 − λg,t) δL,g,t / μD,g, λg,t = Vg / (Vg + ρg,t μD,g^2), where ρg,t is the conditional share of not-yet-treated among the cell sample.
- Interpretation: OLS slope lies between the within-cohort slope and the level contrast per mean dose; weight shifts toward level contrast as the untreated-share increases or as Var(D) falls relative to μD^2.
- Causal decomposition of the response index (Theorem 4)
- Under smoothness/support conditions, θR = ∫ W(d) q′(d) dd, where q(d) = E[τ(D) | D = d, treated] and W(d) ≥ 0 integrates to 1 (OLS derivative weights).
- q′(d) = ACRT(d | d) + S(d): ACRT is the causal derivative (response to assigned dose), S(d) is a selection derivative capturing how average gains vary across realized-dose groups.
- Thus θR = causal component + selection component; if S(d) = 0 a.e. then θR is a weighted average of causal derivatives. Bounds on |S(d)| yield sensitivity bounds for the causal part.
- Estimation and inference
- Baseline estimators: treated-vs-not-yet-treated mean contrast for τL and within-cohort OLS slope for θR.
- Influence-function-based joint inference across cohort-time comparisons and event-time aggregates is derived (Wald tests, simultaneous confidence sets).
- Covariate-adjusted extension: orthogonal-score / cross-fit DML-style estimators that preserve level margin but (in general) change the definition of the response index if dose is residualized within cohort (the residual-on-residual representation).
- Empirical illustration (hydraulic fracturing, Bartik et al. 2019 data)
- The continuous-dose OLS coefficient mostly reflects the level-contrast component.
- Under county-level i.i.d. inference, level leads reject a joint zero restriction while response-index leads do not.
- Coding choices that affect the control group can change level estimates but leave response estimates unchanged (because the response index uses only treated units).
Data & Methods
- Setup
- Balanced panel with periods 1..T. Each unit has a first-treatment date G (∞ = never-treated) and a scalar dose D > 0 if treated (D = 0 if never-treated). Dose is fixed after adoption.
- For each cohort-time cell c = (g, t) that pairs cohort g with units not yet treated at t, define Yc = Yt − Yg−1 and Dc = D · 1{G = g}. Within that Pc = P(· | Sc = 1) (sample containing cohort g and not-yet-treated units), the cell is a two-period continuous-treatment DiD problem.
- Main assumptions
- Positivity and finite moments (treated and control presence; finite second moments).
- Positive treated-dose variation (Var(D | treated) > 0) for θR to be defined.
- Consistency / no anticipation (potential outcomes setup).
- PT-L (level parallel trends): E[Untreated change | treated] = E[Untreated change | control].
- PT-R (response parallel trends): E[Untreated change | D, treated] = E[Untreated change | treated].
- Additional smoothness/support assumptions for derivative representation: contiguous support of D among treated, differentiability of conditional effect surface τ(u | d), and compatibility condition E[τ(D) | D] = q(D).
- Theoretical results
- Identification theorems for τL and θR under PT-L and PT-R respectively.
- Exact algebraic mixture of TWFE coefficient into response index and level contrast per mean dose (mixing weight λ).
- Derivative-weight representation of θR and decomposition into causal derivatives (ACRT) and selection term S(d), with bounds.
- Estimation & inference
- Baseline: sample mean contrasts for τL and within-cohort OLS slopes for θR; influence function derivations allow constructing joint/simultaneous inference across many cells and event-time aggregates.
- Covariate-adjusted extension: orthogonal-score / DML estimators using cross-fitting that keep τL target unchanged and implement an adjusted θR (residual-on-residual). Estimators avoid kernel/derivative bandwidths and estimate scalar targets directly.
- Empirics & Monte Carlo
- Monte Carlo studies (not detailed in excerpt) and hydraulic fracturing application illustrate behavior: pooled TWFE coefficients largely reflect level contrasts; separation illuminates which margin drives results; sensitivity to coding and control composition documented.
Implications for AI Economics
- When treatment intensity is continuous (e.g., algorithm usage rate, model parameter intensity, % workforce using AI tools, cloud compute per firm), TWFE/pooled continuous-dose OLS estimates can mix two conceptually different objects:
- A between-group level effect (how outcome differs for adopters at their realized intensity vs not-yet-adopters).
- A within-cohort dose-response (how outcome varies with dose among adopters). Ignoring this can lead to misinterpretation of coefficients as a single “marginal effect of dose.”
- Practical recommendations for empirical AI-economics work with staggered continuous adoption:
- Report margins separately: estimate and show both the level contrast (τL / δL) and the response index θR for cohort-time comparisons and aggregated event-time profiles.
- Test/assess parallel-trends assumptions separately: PT-L for the level margin and PT-R for the response margin are distinct and non-nested; it is possible for one to hold while the other fails.
- Decompose pooled OLS using the explicit mixture formula to see which margin drives results and how changing the untreated share would move the coefficient.
- Use the derivative/selection decomposition to diagnose selection-on-gains: if you suspect that units selecting larger AI intensities differ in returns (selection), exploit the representation to bound the causal derivative component or conduct sensitivity analysis on S(d).
- Be careful with covariate adjustment: residualizing dose within cohort (common practice) changes the response index you estimate. Use orthogonal-score / DML procedures to get valid inference and to control flexibly for covariates, but interpret the adjusted response index as a different object unless you specifically preserve the original definition.
- Avoid interpreting a single pooled continuous-dose TWFE coefficient as the marginal causal effect per unit of dose without (i) verifying which margin dominates, (ii) checking PT-L/PT-R, and (iii) accounting for selection on gains.
- Research design implications
- In staggered AI adoption settings, design choices that alter the composition of not-yet-treated units (e.g., sample trimming, coding of missing intensity as zero, calendar cutoff choices) can materially change the pooled OLS slope through the mixing weight even if cohort-level dose-response and treated-only relationships are stable.
- When dose variation within cohorts is small (low Var(D)), pooled estimates will put greater weight on the level contrast per mean dose; conversely, large within-cohort dose dispersion increases the weight on θR.
- If the policy question is about marginal returns to increasing intensity (e.g., marginal benefit of more AI usage), focus on identifying θR and evaluate selection; if the question is about realized average effects of cohorts at their observed intensities (e.g., average effect of a given rollout intensity), focus on τL.
- Overall: for credible causal inference about continuous AI-treatment intensities in staggered DiD designs, researchers should separate and report the level and response margins, explicitly state and assess the corresponding parallel-trends assumptions, use the mixture/decomposition to interpret pooled coefficients, and apply sensitivity analysis or bounds when selection-on-gains is plausible.
Assessment
Claims (9)
| Claim | Direction | Outcome | Confidence & Evidence | Details |
|---|---|---|---|---|
| Under level parallel trends, the level margin—the average treatment effect at the treated cohort's realized doses—is identified by the difference in mean outcome changes between the treated cohort and units that remain untreated at the outcome date. Other | positive | Average treatment effect at realized doses, measured in outcome units |
Reading fidelity
high
Study strength
high
|
not reported
|
| Under response parallel trends, the within-treated-cohort response index identifies the covariance between dose and treatment effects divided by the variance of dose. Other | positive | Within-cohort outcome-change slope with respect to treatment dose |
Reading fidelity
high
Study strength
high
|
not reported
|
| Level parallel trends and response parallel trends are non-nested assumptions: neither implies the other. Governance And Regulation | mixed | Identification assumptions for level and response margins |
Reading fidelity
high
Study strength
high
|
not reported
|
| The continuous-dose OLS coefficient in a cohort-time comparison is a convex combination of the within-treated response index and the treated-control level contrast divided by the treated cohort's mean dose. Organizational Efficiency | mixed | OLS slope of the outcome change on continuous treatment dose |
Reading fidelity
high
Study strength
high
|
not reported
|
| Holding the conditional outcome and dose distributions fixed, increasing the not-yet-treated share shifts weight in the continuous-dose OLS coefficient away from the response index and toward the level contrast per unit of mean treated dose. Organizational Efficiency | mixed | Composition of the continuous-dose OLS coefficient |
Reading fidelity
high
Study strength
high
|
not reported
|
| The response index can combine a causal response component with a selection-on-gains component, so response parallel trends alone does not generally give a purely causal response derivative. Decision Quality | mixed | Dose-response slope and its causal versus selection components |
Reading fidelity
high
Study strength
high
|
not reported
|
| If the selection derivative is zero almost everywhere, the response index equals the causal response component; if the selection derivative is nonnegative, the causal component is no greater than the response index. Decision Quality | positive | Causal response derivative with respect to dose |
Reading fidelity
high
Study strength
high
|
not reported
|
| In the hydraulic-fracturing application, the level leads reject a joint zero restriction, whereas the response-index leads do not. Employment | mixed | Pre-treatment employment contrasts and within-cohort employment-dose response leads |
Reading fidelity
high
Study strength
medium
|
not reported
|
| In the hydraulic-fracturing application, the continuous-dose OLS coefficient places most of its weight on the level-contrast component. Employment | positive | Relative contribution of level and response margins to the continuous-dose OLS coefficient |
Reading fidelity
high
Study strength
medium
|
most of its weight
|