0 cumulative citations
View corpus contextScientific teams have grown and papers have become more topically concentrated since 2010, yet cited inputs have not narrowed and individual researchers show only modest broadening; deeper prior work in a focal subfield predicts higher citation impact.
Citation observations
Cumulative provider counts captured on specific dates; providers are never combined.
Generative artificial intelligence raises a central question for scientific training and organization. Is research shifting from deep specialization toward broad individual knowledge? We examine this proposition across papers, cited knowledge, contributor histories, and teams using 47,959 articles from six fields over 2010-2025, 51,736 resolved cited works, and chronologically reconstructed prior publication histories for 1,754 randomly selected index contributors. From 2010 to 2022, team size increased by an estimated 37.3% (95% confidence interval [34.4%, 40.3%]), while paper topic breadth declined by 0.0144 on a 0-1 hierarchical distance scale. Cited knowledge was stable to modestly broader, revealing a divergence between focused outputs and the reach of knowledge inputs. Established contributors' prior breadth increased by 0.0190 [-0.0078, 0.0459] by 2019-2022, within a +/-0.05 equivalence bound assessed in sensitivity analysis. In mature citation windows, one standard deviation of focal depth was associated with 8.2% higher 1 + FWCI [1.9%, 14.9%]; average breadth and interaction associations were smaller under the specified equivalence bounds. Post-2022 deviations from earlier trends were not systematic, and recent changes did not vary clearly with baseline AI intensity across 83 subfields. The findings support a differentiated structure of scientific expertise in which focused individual accumulation coexists with expanding collaboration and sustained access to diverse knowledge inputs.
Summary
Main Finding
Scientific production from 2010–2022 shows a differentiated reorganization: teams grew substantially and papers became slightly more topically concentrated, while the knowledge inputs cited remained stable to modestly broader. Individual established researchers’ repertoires changed only modestly (within a small equivalence bound), and deeper prior work in a focal subfield (focal depth) is associated with higher mature citation impact. Overall, focused individual accumulation coexists with larger teams and continued access to diverse knowledge inputs — i.e., a separation of depth at the individual level from breadth in collaborative and input structures.
Key Points
- Samples and fields: analysis covers 47,959 OpenAlex articles (six fields: arts & humanities, computer science, engineering, medicine, physics, social sciences) over 2010–2025; linked references and contributor histories were constructed for nested subsamples.
- Team size rose markedly: modeled log(1 + team size) annual coefficient ≈ 0.01738 (95% CI [0.01602, 0.01875]). Modeled increase ≈ 37.3% (95% CI [34.4%, 40.3%]) over 2010–2022.
- Paper topic breadth contracted slightly: annual decline ≈ 0.00120 (95% CI [–0.00166, –0.00074]) per year; cumulative change ≈ –0.0144 on a 0–1 hierarchical distance scale (2010–2022).
- Cited knowledge breadth was stable to modestly increasing: per-year increase ≈ 0.00228 (HC3 95% CI [0.00002, 0.00454], p = 0.048; clustering by field-year attenuates significance). Coarser measures (share of out-of-field refs, effective number of cited fields) were approximately stable.
- Individual prior breadth: established contributors’ five-year prior breadth rose modestly by 0.0190 (95% CI [–0.0078, 0.0459]); not statistically different from zero but within an equivalence bound of ±0.05 (TOST p = 0.012; 90% CI [–0.0035, 0.0416]).
- Focal depth predicts impact: one standard deviation increase in focal depth → ~8.2% higher 1 + FWCI (95% CI [1.9%, 14.9%], p = 0.010) after adjusting for covariates. Breadth (at mean depth) and the breadth×depth interaction had small, statistically indistinct associations (breadth: 1.8% [–4.7%, 8.8%]; interaction: 0.7% [–4.9%, 6.6%]).
- Post-2022: no systematic deviations from prior trends and no clear heterogeneity in post-2022 changes by baseline AI intensity across 83 subfields.
- Interpretation: evidence supports a “differentiated” organization where concentrated, specialist outputs (and the citation returns to focal depth) coexist with larger collaborative teams and continued access to diverse knowledge inputs.
Data & Methods
- Primary bibliographic source: OpenAlex.
- Article sample: 47,959 randomly sampled articles (500 per field per year across six fields and 16 years), excluding retractions and paratext.
- Reference sample: 3,360 sampled papers (expanded to 35 papers per field-year); resolved 51,736 cited works (≈93.7% resolution); up to 20 references per paper sampled.
- Contributor histories: chronologically reconstructed prior publications for 1,754 index contributors; after eligibility and identity filters, 1,316 eligible contributor events and 959 complete cases for citation-return analyses.
- Key measures:
- Paper topic breadth: hierarchical topic distance (0–1 scale), primary metric often from papers assigned exactly three topics for comparability.
- Cited knowledge breadth: distance among topics of resolved references.
- Individual prior breadth: distribution of recent topics over a 5-year lookback (alternative history windows tested).
- Focal depth: accumulated prior work in the subfield of the focal article (log scale).
- Outcome: mature Field-Weighted Citation Impact (FWCI) with OpenAlex’s 4-year window; models used 1 + FWCI transformed and modeled on log scale.
- Team size: log(1 + team size) and alternative log(team size).
- Estimation details:
- Regression models adjust for field and year fixed effects, career age, recent publication volume, team size, open access status, etc.
- Robustness: HC3 standard errors, clustering within field-year cells for some estimates, holdouts (pilot vs expanded), alternative aggregations and weighting, ORCID subsample checks.
- Multiple-testing control: six pre-specified primary tests, Holm adjustment for familywise error.
- Equivalence tests: two one-sided tests (TOST) with specified bounds (e.g., ±0.05 on breadth); sensitivity bounds used in impact checks (e.g., ±log(1.10)).
- Additional checks: predictive counterfactual forecasting and continuous difference-in-differences for post-2022 / AI-intensity analyses across 83 subfields (pre-specified exposure window).
Implications for AI Economics
- AI tools have not (through 2022/early 2023 evidence) induced a large, measurable shift of individual scientific talent from depth toward breadth. Instead, organizational and input-side changes have been more salient (larger teams, stable or modestly broader cited inputs).
- Returns to specialization remain important: focal subfield depth predicts higher mature citation impact. This implies that specialist human capital retains value even as AI lowers information access costs.
- Complementarity of AI and specialization: generative AI may lower search, synthesis, and coordination costs, enabling teams and individuals to incorporate more diverse inputs without eroding the productivity premium of deep subfield expertise. In labor-market terms, demand for specialists with deep tacit knowledge likely persists; AI augments them rather than replaces the value of depth.
- Skill-composition and training:
- Short-to-medium term: investments in deeper disciplinary training remain warranted, alongside training in cross-domain search, prompt engineering, and tools that facilitate integrating external inputs.
- Institutions (universities, funders, firms) should balance incentives: reward depth that drives impact (per focal depth returns) while supporting collaborative and integrative capacities (teamwork, data/tool access).
- Organizational design and hiring:
- Teams are growing; organizations may prioritize coordination, project management, and integrative roles that combine limited breadth with mechanisms for accessing diverse inputs (e.g., AI-enabled literature synthesis).
- Recruitment signals: evidence supports continued premium for demonstrated focal depth in impact-sensitive roles; evaluation metrics should account for team size and collaborative contributions.
- Funding and evaluation policy:
- Bibliometric evaluation should distinguish between input diversity (references, sources) and output topical breadth — both matter but capture different phenomena. Funding programs that seek interdisciplinarity should explicitly define whether they target breadth of inputs (source diversity) or breadth of outputs (topic span).
- Given the persistence of depth’s returns, funding mechanisms that reward concentrated, risky specialization remain justified alongside interdisciplinary initiatives.
- Market and productivity implications:
- AI adoption may raise aggregate productivity by easing access to knowledge inputs and coordination (consistent with larger teams and stable/broader cited inputs), but gains are likely complementary to — not substitutes for — accumulated specialist human capital.
- Wage and career dynamics: unless AI fundamentally reduces the marginal value of tacit, practice-based expertise, specialists should continue to command favorable returns; generalists who combine coordination, tool literacy, and some domain competence may see rising demand.
- Cautions for policymakers and economists:
- Observational limits: the paper’s timeframe mostly precedes very recent waves of generative AI deployment; continued monitoring is needed to detect later shifts.
- Field heterogeneity: effects vary by field; sectoral and subfield-specific AI exposure matters for local labor-market impacts.
- Measurement caveats: citation impact is an imperfect proxy for social/economic value; broader metrics (translation to products, patents, applied outcomes) may show different patterns.
Limitations noted by the authors (relevant to inference): - Observational design (causal claims limited). - Coverage and measurement choices (OpenAlex coverage, topic assignments, reference resolution). - Time window: many generative-AI developments post-date the primary pre-2023 capability comparisons; effects may evolve as tools and practices diffuse.
Bottom line: up through the study window, scientific systems are showing larger teams and somewhat more concentrated outputs while still drawing on diverse inputs; individual-level breadth hasn’t demonstrably replaced depth, and depth continues to matter for citation impact. For AI economics, this suggests AI is reshaping coordination and information access more than it has (so far) reallocated the returns from specialization to generalism.
Assessment
Claims (10)
| Claim | Direction | Outcome | Confidence & Evidence | Details |
|---|---|---|---|---|
| From 2010 to 2022, paper topic breadth declined by 0.0144 on a 0–1 hierarchical distance scale. Research Productivity | negative | Paper topic breadth |
Reading fidelity
high
Study strength
high
|
n=47959
-0.0144 on a 0–1 hierarchical distance scale
|
| Team size increased substantially from 2010 to 2022, with a modelled increase of 37.3%. Team Performance | positive | Scientific research team size |
Reading fidelity
high
Study strength
high
|
n=38285
37.3% increase [34.4%, 40.3%]
|
| Cited knowledge breadth increased slightly over time, while coarser measures of cited-field diversity were approximately stable. Research Productivity | mixed | Breadth and diversity of cited knowledge inputs |
Reading fidelity
high
Study strength
medium
|
n=2619
0.00228 per year [0.00002, 0.00454] in the unclustered estimate
|
| Established contributors' five-year prior topic breadth was only modestly higher in 2019–2022 than in 2010–2014, and the change was within the prespecified ±0.05 equivalence bound. Skill Acquisition | null_result | Established contributors' prior topic breadth |
Reading fidelity
high
Study strength
medium
|
n=1292
0.0190 on the 0–1 breadth scale [–0.0078, 0.0459]
|
| Focal depth was not clearly changed between 2019–2022 and 2010–2014. Skill Acquisition | null_result | Accumulated individual depth in the focal subfield |
Reading fidelity
high
Study strength
medium
|
n=1292
-0.0274 [–0.1451, 0.0903]
|
| Among contributor events with mature citation windows, one standard deviation greater focal depth was associated with 8.2% higher 1 + FWCI. Research Productivity | positive | Field-weighted citation impact of the focal article |
Reading fidelity
high
Study strength
high
|
n=959
8.2% higher 1 + FWCI per one standard deviation increase in focal depth [1.9%, 14.9%]
|
| Prior breadth had no statistically distinguishable association with mature citation impact at mean focal depth. Research Productivity | null_result | Field-weighted citation impact of the focal article |
Reading fidelity
high
Study strength
medium
|
n=959
1.8% [–4.7%, 8.8%]
|
| The interaction between prior breadth and focal depth was not statistically distinguishable from zero in its association with mature citation impact. Research Productivity | null_result | Interaction of prior breadth and focal depth on field-weighted citation impact |
Reading fidelity
high
Study strength
medium
|
n=959
0.7% [–4.9%, 6.6%]
|
| Post-2022 deviations from trends forecast using earlier years were not statistically distinguishable from zero. Other | null_result | Post-2022 deviation from forecasted scientific capability trends |
Reading fidelity
high
Study strength
medium
|
n=47959
-0.0108 [-0.026, 0.004]
|
| Recent changes did not vary clearly with baseline AI intensity across 83 subfields. Automation Exposure | null_result | Differential post-2022 change by baseline AI intensity |
Reading fidelity
high
Study strength
low
|
n=83
|