The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

Singular-value decompositions of the nonparametric two-way regression kernel produce eigenfunction-based proxies that identify unit and time fixed effects: injectivity holds almost everywhere under mild nondegeneracy and holds uniformly (for a finite truncation) under a local full-rank/immersion condition, enabling identification of the regression function at the realised types.

Nonparametric Identification of Two-Way Unobserved Heterogeneity
Hugo Freeman, Dennis Kristensen · August 27, 2026
arxiv theoretical n/a evidence 7/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Hugo Freeman unresolved corpus identity
  2. Dennis Kristensen unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Hugo Freeman provider ID
  2. Dennis Kristensen provider ID
The paper shows that the SVD of the nonparametric two-way regression kernel yields finite-dimensional eigenfunction proxies for unit and time heterogeneity that are injective (a.e. or uniformly under a local immersion), so the nonparametric fixed-effects regression is identified at realised types up to observational equivalence.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

We study identification of two-way unobserved heterogeneity in the nonparametric panel regression $G_{it}=g(α_i,γ_t)+\varepsilon_{it}$, where identification of the latent types reduces to constructing identified, \emph{injective} proxies for them. To this end we consider the singular value decomposition (SVD) of the bivariate regression function $g(α,γ)$ on a product domain $Ω_α\timesΩ_γ$, whose left singular functions $\{u_r\}$ serve as proxies for the unobserved heterogeneity parameter $α$. The arguments are symmetric for $\{v_r\}$ vis-à-vis $γ$. We work under an \emph{observational-equivalence simplification}: two values of $α$ that induce the same conditional response $g(α,\cdot)$ are identified, so that the response map $α\mapsto g(α,\cdot)$ is injective by construction. We show two things. First, this reduction is \emph{equivalent} to injectivity of the full collection of left singular eigenfunctions, so no further condition is needed over the infinite collection $\{u_r\}_{r\ge1}$. Second, under a single additional \emph{local injectivity} condition, a finite collection of leading eigenfunctions $U_R=(u_1^{\top},\dots,u_R^{\top})^{\top}$ is injective for all sufficiently large $R$. The proof reduces a global univalence question to a local first-order condition plus a topological compactness argument, bypassing the global Jacobian conditions usually required.

Summary

Main Finding

The paper shows that the latent, two-way unobserved heterogeneity in the nonparametric panel regression Git = g(αi, γt) + εit can be identified (up to observational equivalence) by using the singular value decomposition (SVD) of the bivariate regression kernel g. The left/right singular functions (ur, vr) serve as injective, identified proxies for the cross-sectional and time heterogeneity respectively. Two identification regimes are established: (i) an almost-everywhere regime where a finite collection of leading eigenfunctions becomes injective on a set of probability arbitrarily close to 1, and (ii) a uniform regime where under a stronger local full-rank condition a finite number of leading eigenfunctions are globally injective on the whole support.

Key Points

  • Model & goal

    • Observed transformed outcomes Git = G(Yit) with conditional mean g(αi, γt).
    • Aim: identify realised fixed effects (αi, γt) or injective proxies for them and identify g at those realised types under large N, large T.
    • Observational-equivalence simplification: distinguishability is only required up to values of α that induce distinct functions g(α, ·).
  • SVD-based proxy construction

    • Treat g as a Hilbert–Schmidt kernel on Ωα × Ωγ and form the integral operator Tg.
    • SVD: g(α, γ) = Σr σr ur(α) vr(γ) with left singular functions ur ∈ Hα (RdG-valued) and right singular functions vr ∈ L2π(Ωγ).
    • Define eigenfunction maps UR(α) = (u1(α),...,uR(α))⊤ and VR(γ) similarly. These serve as proxies λi = UR(αi) and ft = VR(γt).
  • Two identification regimes (Theorem 2)

    • Almost-everywhere regime (weaker assumption): If the gradient-information matrix M∞(α) := ∫ J(α,γ)⊤J(α,γ) dπγ is positive definite for almost every α, then for any δ>0 there exists finite R(δ) and large-probability compact sets on which UR and VR are injective for all R ≥ R(δ). Thus proxies can be made injective on sets of probability arbitrarily close to one, at the cost of increasing R.
    • Uniform regime (stronger assumption): If a finite “design” of time points yields a stacked Jacobian with full column rank uniformly in α (a local immersion that holds uniformly), then there exists finite R such that UR and VR are globally injective on the whole supports for every R ≥ R.
  • Key technical ingredients

    • Use gradient expansion of J(α,·) = Dαg(α,·) in the vr basis and define MR(α) = Σ_{r≤R} σr^2 Dur(α)⊤Dur(α), with M∞ its limit.
    • Show M∞(α) equals ∫ J⊤J and relate positive-definiteness of M∞ to local injectivity of the infinite eigenfunction map U∞.
    • Reduce global injectivity of finite UR to a local first-order (immersion) condition plus topological compactness; avoids global Jacobian (univalence) conditions.
  • Additional points

    • Theorem 1: If identified finite-dimensional proxies (λi, ft) are injective maps of (αi, γt), then g at realised types is identified and inherits regularity under smoothness of U,V,g.
    • The paper treats both α- and γ-sides symmetrically (γ-analogues required).
    • Additive structure in the kernel across coordinates of the types yields componentwise decomposition that can remove the curse of dimensionality (Section 5.8).

Data & Methods

  • Mathematical / theoretical approach (no new empirical dataset)
    • Functional-analytic framework: Hilbert–Schmidt integral operators on L2 spaces, SVD of compact operators, orthonormal systems of singular functions.
    • Sobolev regularity: eigenfunction smoothness established under Hp assumptions (Appendix A); use of weighted/unweighted Sobolev equivalence given densities bounded away from 0 and ∞.
    • Differential calculations: Jacobian expansions, Bessel and dominated convergence arguments show convergence of differentiated SVD series and closed form of M∞.
    • Topological/compactness arguments: reduce global injectivity to local immersion plus compactness to get finite truncation injectivity.
  • Main assumptions
    • Support assumptions: compact (and for α, convex and connected) supports; densities bounded away from 0 and ∞.
    • Smoothness: g continuously differentiable in α with joint continuity of Dαg on Ωα×Ωγ (so ur are C1).
    • Observational equivalence: the response map α ↦ g(α, ·) is injective after removing observationally equivalent points.
    • Assumption 4 (weaker): det M∞(α) = 0 only on a πα-null set (M∞ ≻ 0 a.e.).
    • Assumption 5 (stronger): existence of a finite list of evaluation points in Ωγ so that a stacked finite Jacobian has full column rank uniformly in α (gives uniform positive-definiteness of M∞).
  • Results are identification proofs; estimation aspects (finite-sample error, rates) are discussed qualitatively (e.g., how truncation order R must grow) but detailed finite-sample theory/algorithms are not developed in this paper.

Implications for AI Economics

  • What it enables
    • Rigorous identification of latent heterogeneous agent/time types in rich nonparametric panel models when heterogeneity affects conditional responses (i.e., when different types generate different conditional distributions of transformed outcomes).
    • A constructive route to build low-dimensional, injective proxies for latent types by spectral decomposition of the estimated regression kernel. Those proxies can be used as sufficient controls in downstream structural or causal analysis.
  • Practical consequences for empirical work in AI economics
    • When studying heterogeneous effects of algorithmic policies, platform changes, or firm-level AI adoption over time, researchers can:
      • Estimate the conditional mean g nonparametrically (or its moments G(Yit)),
      • Compute an empirical SVD of the estimated kernel,
      • Use leading singular-function evaluations as proxies for cross-sectional/time types.
    • The paper gives conditions under which a finite number of leading components suffice to identify types (so using low-rank approximations is justified), and it clarifies the trade-off: under weak assumptions one can achieve injectivity on almost all types by letting the number of components R grow (possibly slowly), while under a stronger local full-rank condition a fixed finite R is enough globally.
  • Guidance on choosing truncation R and interpretation
    • Under the almost-everywhere regime, R must grow with sample sizes to drive down the measure and diameter of the set of conflated types; in examples with bounded type density this can be logarithmic in NT, but no universal rate holds.
    • Under the uniform regime, a finite R* suffices; verifying the finite-design full-rank condition in applications amounts to checking that variation in g across some finite set of γ points spans the α-derivatives.
  • Limitations and cautions for applied researchers
    • Identification is up to observational equivalence: types that produce exactly identical conditional responses cannot be distinguished by any procedure based on g; researchers should check whether types of interest can plausibly be distinguished via data transformations G(·).
    • The theory relies on smoothness, compact-support, and density bounds; violations may break eigenfunction regularity and injectivity conclusions.
    • The paper proves identification; it does not provide a full finite-sample estimation and inference recipe that quantifies sampling error from nonparametric estimation of g and of the operator SVD. Implementing the approach requires careful nonparametric estimation and numerical spectral analysis, with attention to discretisation and regularisation.
  • Directions for empirical AI-economics research
    • Use eigenfunction proxies as controls in causal or structural estimators to account for two-way unobserved heterogeneity (e.g., heterogeneous response to algorithmic recommendations across user types and time regimes).
    • Exploit additive or separable kernel structure (when present) to mitigate curse-of-dimensionality concerns, per the componentwise decomposition in the paper.
    • Combine with asymptotic results for nonparametric estimators of g to develop finite-sample procedures (model selection for R, regularisation) that inherit the identification guarantees here.

Summary recommendation: the paper provides a clean, constructive identification strategy for two-way latent heterogeneity via kernel SVD and gives clear conditions (both weak and strong) under which a finite spectral truncation yields valid, interpretable proxies. Applied researchers in AI economics can leverage these proxies to control for rich unobserved heterogeneity, but must take care when estimating g nonparametrically and when verifying the assumptions required for the desired injectivity regime.

Assessment

Paper Typetheoretical Evidence Strengthn/a — This is a theoretical identification paper with no empirical estimation or causal inference on data; evidence consists of formal theorems, lemmas, and proofs rather than empirical results. Methods Rigorhigh — The paper states clear assumptions, derives lemmas linking the SVD expansion and derivatives, defines gradient information matrices, treats both almost-everywhere and uniform (global) regimes, and provides topological and Sobolev-regularity arguments to reduce global injectivity to a local immersion plus compactness—showing technical care and awareness of edge cases (degeneracy set, eigenvalue/truncation issues). SampleTheoretical model: panel observations Y_it (or moments G(Y_it)) for i=1..N, t=1..T under large-N, large-T asymptotics; unobserved finite-dimensional unit effects α_i ∈ Ω_α ⊂ R^{d_α} and time effects γ_t ∈ Ω_γ ⊂ R^{d_γ}; nonparametric conditional mean g(α,γ) = E[G(Y)|α,γ] in L2 over product domain; assumptions on compact supports, bounded type densities, Sobolev smoothness of g, and existence of a known transformation G applied to Y. Themesproductivity org_design IdentificationConstruct left/right singular-function maps UR(α) and VR(γ) from the SVD of the nonparametric two-way regression kernel g(α,γ); treat these finite-dimensional eigenfunction truncations as proxies for the unobserved unit and time types and prove injectivity of the proxy maps (pointwise a.e. under a mild nondegeneracy and Lipschitz conditions, and uniformly for all sufficiently large truncation order under a stronger local full-rank/immersion condition), which yields identification of g(α,γ) at the realised types (up to observational equivalence). GeneralizabilityRequires finite-dimensional latent types (d_α,d_γ fixed); extension to infinite-dimensional heterogeneity is not addressed, Relies on compactness, bounded densities and Sobolev smoothness of g (so fails if domains are unbounded or g lacks required smoothness), Identification is up to observational equivalence: distinct α that induce identical response functions cannot be separated, Uniform injectivity requires a stronger local full-rank/immersion design (Assumption 5); without it only almost-everywhere identification is guaranteed and convergence/truncation rates depend on unknown collision geometry near degeneracy set, Construction assumes availability/knowledge of a transformation G of Y so that conditional mean g separates types; applicability depends on choosing such G in practice

Claims (8)

ClaimDirectionOutcomeConfidence & EvidenceDetails
Identification of the latent heterogeneity types reduces to constructing identified proxies for the types that are injective. Other positive Identification of latent individual and time heterogeneity
Reading fidelity high
Study strength high
not reported
0.2
Under the observational-equivalence reduction, injectivity of the response map α ↦ g(α, ·) is equivalent to injectivity of the full collection of left singular eigenfunctions, so no additional condition is required for the infinite collection of eigenfunction proxies. Other positive Injectivity of the infinite-dimensional eigenfunction proxy
Reading fidelity high
Study strength high
not reported
0.2
If a local injectivity condition holds, a finite collection of sufficiently many leading left singular eigenfunctions is injective. Other positive Global injectivity of finite-dimensional eigenfunction proxies
Reading fidelity high
Study strength high
not reported
0.2
Given identified and injective proxies λi = U(αi) and ft = V(γt), the normalized regression function g0(λi, ft) identifies the original regression value g(αi, γt) at the realized types. Other positive Identification of the regression function at realized latent types
Reading fidelity high
Study strength high
not reported
0.2
Under the almost-everywhere assumptions, for every δ > 0 there is a finite truncation order R(δ) such that the leading eigenfunction proxies are injective on compact subsets containing at least 1 − δ of the support mass. Other positive Almost-everywhere injectivity of finite eigenfunction proxies
Reading fidelity high
Study strength high
mass at least 1 − δ
0.2
Under the stronger uniform local-rank assumptions, there is a finite truncation order R* such that every truncation of order R ≥ R* is injective over the entire supports of both individual and time heterogeneity. Other positive Uniform global injectivity of finite eigenfunction proxies
Reading fidelity high
Study strength high
not reported
0.2
The stronger finite-design full-rank condition implies that the infinite gradient information matrix M∞(α) is positive definite at every α and has a uniformly positive minimum eigenvalue over the compact support. Other positive Local rank and gradient information for identifying latent types
Reading fidelity high
Study strength high
c0 > 0
0.2
With a fixed finite truncation order, proxy collisions can affect entire rows and columns of the panel, and the fraction of affected cells is of order δα(R) + δγ(R). Error Rate negative Fraction of panel cells affected by proxy misclassification
Reading fidelity high
Study strength medium
of order δ(R) := δα(R) + δγ(R)
0.12

Notes