0 cumulative citations
View corpus contextA spectral decomposition lets researchers transport causal estimates between experimental assignment policies in networks by projecting inverse-probability weights onto a prespecified Fourier subspace, but gains require assuming low-degree/low-support interference: misspecification or large policy shifts drive bias or exponential variance.
Citation observations
Cumulative provider counts captured on specific dates; providers are never combined.
In the presence of interference, where the treatment assigned to one unit can affect the outcomes of others, many causal estimands depend on the treatment-assignment policy under which the experiment is conducted. This policy dependence creates a fundamental challenge for off-policy estimation, where the goal is to estimate causal quantities under a hypothetical intervention policy different from the one used to collect data. We study this problem of off-policy estimation of causal effects for heterogeneous Bernoulli policies. By representing exposure-weighted potential outcomes in the biased Fourier basis of the experimental design, we construct, for any prespecified Fourier subspace encoding the assumed interference structure, the unique minimum-$L^2$ weight that transports every function in that subspace. Global and local inverse-probability weights, linear-interference weights, and no-interference weights are special cases. The weight variance is a structured chi-square distance between the experiment and target policies. When the assumed interference structure is misspecified, the introduced bias couples the omitted outcome spectrum with the corresponding policy-shift coefficients, yielding a sharp robustness bound and a bias-variance trade-off. A Fourier-neighborhood-overlap condition gives consistency under structured interference, and we state a Doob-martingale central limit theorem for off-policy estimators. As the variance is not identified, we derive identifiable bounds and associated conservative estimators of the variance. Simulations illustrate these theoretical results for the design and analysis of experiments under network interference and design mismatch.
Summary
Main Finding
The paper develops a principled framework for off-policy causal estimation in networked experiments with interference by representing potential outcomes in the biased Fourier basis induced by a Bernoulli (product) assignment policy. For any analyst-specified Fourier subspace S (encoding which interactions / orders matter), the authors derive the unique minimum-L2 (minimum-variance) weighting function that exactly transports expectations from the experimental (design) policy π to a target policy π′ for every function in that subspace. They characterize (i) the variance of that weight as a structured chi-square distance between π and π′ restricted to S, (ii) the exact bias when S omits true Fourier components, and (iii) consistency, bias–variance trade-offs, and conservative variance bounds under structured interference.
Key Points
-
Problem setup
- Units i = 1..n, full potential-outcome functions yi(z) defined on the Boolean cube {0,1}^n.
- Design policy Z ∼ Pπ is independent Bernoulli with unit probabilities π = (πi). Target policy is π′.
- Many causal estimands (e.g., expected average outcome Eπ′[Yi] or exposure contrasts) depend on the assignment policy → off-policy estimation problem.
-
Biased Fourier representation
- Use π-biased Fourier characters χVπ(z) = ∏_{j∈V} (zj − πj)/√(πj(1−πj)), giving an orthonormal basis under Pπ.
- Any function f on assignments expands as f(z) = ∑_{V} f̂π(V) χVπ(z).
- Policy-shift coefficients: ∆Vππ′ = ∏_{j∈V} (π′j − πj)/√(πj(1−πj)). These equal Eπ′[χVπ(Z)].
-
Minimum-variance spectral transport weight
- For a chosen Fourier support S (a subset of index sets V), define the S-restricted weight wSππ′(z) = ∑_{V∈S} ∆Vππ′ χVπ(z).
- Proposition: If f ∈ HS (has Fourier support contained in S), then Eπ[f(Z) wSππ′(Z)] = Eπ′[f(Z)]. Among all weights that balance every f in HS, wSππ′ uniquely minimizes L2(Pπ) norm (hence variance).
- Variance of the S-weight: Varπ(wSππ′) = d2S(π′,π) = ∑_{V∈S{∅}} (∆Vππ′)^2. This is a structured chi-squared policy distance along the retained Fourier directions.
- The full inverse-probability weight (transport for all functions) is the S = P([n]) case; wS is the orthogonal projection of the full IPW onto HS.
-
Misspecification (bias) characterization
- Exact bias when the true function f is not in HS: BiasS(f; π, π′) = Eπ[f wSππ′] − Eπ′[f] = − ∑_{V∉S} f̂π(V) ∆Vππ′.
- Norm bound: |BiasS| ≤ ‖f − PS f‖2,π · dSc(π′,π). So the bias is the inner product of omitted Fourier coefficients and omitted policy-shift coefficients.
- Local behavior: omitted interactions of order r are scaled by factors ≈ |q−p|^r (for homogeneous shifts), giving O(|q−p|^{r0}) local bias if first omitted order is r0.
-
Bias–variance trade-off and selection of S
- Enlarging S reduces bias (remove omitted terms) but increases variance by the added ∑ (∆V)^2. The choice of S is thus an explicit bias–variance decision.
- Corollary gives an explicit envelope bounding mean-squared error in terms of overlap-adjusted variance (depends on retained d2Si) and squared omitted-spectrum bias.
-
Consistency under structured interference
- Define for each unit an influence set Γi (units whose treatments affect the unit-level summand) and an overlap degree D = max_i |{j ≠ i : Γi ∩ Γj ≠ ∅}|.
- Theorem: If the per-summand second moments are bounded by B and B·(D+1) = o(n), then the average estimator is consistent (variance ≤ B (D+1)/n).
- Effective sample size n_eff = n/(D+1). Gives sufficient scaling conditions for common interference regimes (complete local interference, linear / degree-limited interactions, cluster designs).
-
Variance nonidentifiability and conservative bounds
- Randomization variance of an unbiased estimator is generally not identifiable from a single realized assignment (analogue of Neyman variance nonidentifiability).
- The authors derive identifiable covariance bounds and conservative (valid) variance estimators and provide a Doob-martingale CLT for off-policy estimators under their dependence structure.
-
Practical special cases (Table summary)
- S choices include: full Γ (complete local interference on a subset), degree ≤ d (limits interaction order), linear (only linear terms), SUTVA (only own-unit terms). Corresponding wS simplify to product IPW, low-degree polynomials, linear-in-χ, etc.
Data & Methods
- The work is theoretical / methodological, not empirical-data driven. Main technical elements:
- Model: independent Bernoulli assignment policies (product distributions).
- Analytical tool: π-biased Fourier expansion of Boolean functions; manipulation of Fourier coefficients and expectations across product measures.
- Construction: orthogonal projection of the full inverse-probability weight onto a specified Fourier subspace S yields the minimum-L2 balancing weight wSππ′.
- Analysis:
- Exact algebraic expressions for transported expectations, weight variance (structured chi-square), and misspecification bias.
- Asymptotic results via dependence bounds based on influence sets and a Doob-martingale central limit theorem (proofs in appendices).
- Derived conservative, identifiable bounds for variance/covariance since full randomization variance is not observable from a single assignment.
- Simulations are used (in paper) to illustrate finite-sample behavior, the bias–variance trade-off, the performance of spectral weights, and the effect of design–target mismatch.
Implications for AI Economics
-
Off-policy evaluation in networked settings (platform interventions, algorithmic recommendations, contagion of behaviors) requires structural restrictions:
- Without structure, full IPW transports but has exponential variance in the number of interacting units and is impractical.
- The Fourier-spectral approach gives a flexible, principled way to encode structure (which units and which interaction orders matter) that directly targets the components relevant to policy transport.
-
Practical guidance for experimental design and analysis
- When planning experiments to estimate effects for a family of target policies π′, either:
- Design π to have good overlap with π′ in the relevant Fourier directions, or
- Restrict S to low-degree / small influence sets (e.g., linear terms or neighborhood-limited interactions) to keep variance manageable.
- The paper makes explicit the bias–variance trade-off of choosing S: include more Fourier directions to reduce bias but accept increased variance = d2S(π′,π).
- Use influence-set and overlap diagnostics to assess effective sample size n_eff = n/(D+1). For platform A/B tests where users interact densely, design should reduce overlap (clustering, blocking) or limit target policy shifts.
-
Relevance for policy learning and transfer
- The structured chi-square distance d2S(π′,π) quantifies the difficulty of transporting expectations from π to π′ along the retained interactions; this can guide which policy changes are reliably estimable from existing randomized data.
- The methodology connects to importance-weighting and transfer-learning ideas common in AI: wS is an optimal projection of an importance weight under an assumed function class, balancing variance and approximation error.
-
Cautions & limitations
- Requires independent Bernoulli assignment design — extensions needed for correlated randomization schemes.
- The biased Fourier basis and coefficients depend on the design π; estimating outcome Fourier coefficients (f̂π(V)) in practice requires enough variation and possibly modeling assumptions.
- Misspecification of S can produce biased estimates; the paper provides exact bias formulas and bounds, but in practice one must trade off bias vs feasible variance.
- Randomization variance is not fully identifiable from a single realization; use conservative variance bounds for valid inference.
-
Takeaway for AI economists
- When evaluating or comparing AI-driven interventions on networks (recommendation rule changes, nudges, pricing, algorithmic tweaks), explicitly encode plausible interaction structure (low-degree, neighborhood-based, or linear spillovers) and use the spectral-transport weights to obtain minimum-variance off-policy estimates for that class. Use the provided bias expressions, the structured chi-square distance, and overlap diagnostics to judge whether an off-policy transport is feasible and to design experiments that make desired transports identifiable and low-variance.
Assessment
Claims (10)
| Claim | Direction | Outcome | Confidence & Evidence | Details |
|---|---|---|---|---|
| For any prespecified Fourier subspace containing the assumed interference structure, the proposed spectral-transport weight exactly transports the expectation of every function in that subspace from the design policy to the target policy and uniquely minimizes the L2 norm among all weights with this property. Error Rate | positive | Exactness and L2 efficiency of off-policy expectation transport |
Reading fidelity
high
Study strength
high
|
not reported
|
| The variance of the minimum-L2 spectral-transport weight equals the sum of squared design-to-target policy-shift coefficients over the retained nonconstant Fourier directions. Error Rate | positive | Randomization variance of the off-policy transport weight |
Reading fidelity
high
Study strength
high
|
Varπ(WSππ′) = ΣV∈S\{∅}(ΔVππ′)²
|
| When the retained Fourier supports contain the supports of the exposure-weighted potential-outcome functions, the proposed estimator is unbiased for the target-policy causal contrast; the expected average outcome is a one-term special case. Decision Quality | positive | Target-policy causal contrast and expected average outcome |
Reading fidelity
high
Study strength
high
|
not reported
|
| Without restrictions on the outcome function, off-policy transport requires balancing the full treatment-assignment distribution, corresponding to the global inverse-probability weight; its variance generally grows exponentially with the number of units whose treatments affect the outcome. Error Rate | negative | Variance and scalability of unrestricted off-policy estimation |
Reading fidelity
high
Study strength
high
|
Qni=1(1 + (∆iππ′)²) − 1
|
| Under Fourier-subspace misspecification, the estimator's bias is the negative inner product between the omitted outcome Fourier coefficients and the corresponding policy-shift coefficients, and its absolute value is bounded by the omitted-spectrum L2 norm multiplied by the policy distance along omitted directions. Error Rate | mixed | Bias of off-policy expectation estimation under misspecified interference structure |
Reading fidelity
high
Study strength
high
|
|BiasS(f;π,π′)| ≤ ∥f − PSf∥2,π · dSc(π′,π)
|
| If the first omitted Fourier interaction order is r0 under a homogeneous policy shift, the local misspecification bias is of order O(|q − p|r0) as q approaches p, provided the omitted-spectrum energy and multiplicity are controlled. Error Rate | mixed | Local bias from omitted higher-order interference terms |
Reading fidelity
high
Study strength
high
|
O(|q − p|r0)
|
| For averages of unit-level estimators whose treatment dependencies have spectral overlap degree D(n), the variance is bounded by B(n)(D(n)+1)/n; therefore, the estimator is consistent for its expectation when B(n)(D(n)+1)=o(n). Error Rate | positive | Variance and consistency of averaged off-policy estimators under network overlap |
Reading fidelity
high
Study strength
high
|
Varπ(n)(θ̂(n)) ≤ B(n)(D(n) + 1)/n
|
| Under a fixed homogeneous policy shift and uniformly bounded remaining factors, complete local interference permits consistency when the maximum influence-set size grows only logarithmically in the overlap-adjusted effective sample size, whereas linear interference permits the larger regime in which influence-set size is little-o of that effective sample size. Error Rate | positive | Consistency under structured interference and policy shift |
Reading fidelity
high
Study strength
high
|
s(n) ≤ κ log n_eff(n) for complete local interference; s(n) = o(n_eff(n)) for linear interference
|
| The choice of retained Fourier support creates an explicit bias-variance trade-off: adding a Fourier interaction increases the transport-weight variance by its squared policy-shift coefficient while removing the corresponding misspecification-bias term. Error Rate | mixed | Bias-variance trade-off in off-policy estimation |
Reading fidelity
high
Study strength
high
|
Variance increase: (∆Vππ′)²
|
| The randomization variance of an estimator based on an otherwise unrestricted function is generally not identifiable from a single realized treatment assignment and observed outcome pair. Error Rate | negative | Identifiability of randomization variance |
Reading fidelity
high
Study strength
high
|
not reported
|