The Commonplace
Home Three-study pilot Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

Scientific teams have grown and papers have become more topically concentrated since 2010, yet cited inputs have not narrowed and individual researchers show only modest broadening; deeper prior work in a focal subfield predicts higher citation impact.

Has Scientific Talent Shifted from Depth to Breadth?Evidence across Papers, Knowledge Inputs, Careers, and Teams
Xiaoshn Nee, Haobo Zhong, Xiaomin Ni · September 13, 2026
arxiv descriptive medium evidence 7/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Xiaoshn Nee unresolved corpus identity
  2. Haobo Zhong unresolved corpus identity
  3. Xiaomin Ni unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Xiaoshn Nee provider ID
  2. Hao-Bo Zhong provider ID
  3. Xiao-Ming Ni provider ID
From 2010–2022, scientific teams grew substantially while paper topic breadth modestly contracted, cited knowledge stayed stable-to-broader, individual contributor breadth rose only slightly, and accumulated focal depth—rather than breadth—was positively associated with mature citation impact.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

Generative artificial intelligence raises a central question for scientific training and organization. Is research shifting from deep specialization toward broad individual knowledge? We examine this proposition across papers, cited knowledge, contributor histories, and teams using 47,959 articles from six fields over 2010-2025, 51,736 resolved cited works, and chronologically reconstructed prior publication histories for 1,754 randomly selected index contributors. From 2010 to 2022, team size increased by an estimated 37.3% (95% confidence interval [34.4%, 40.3%]), while paper topic breadth declined by 0.0144 on a 0-1 hierarchical distance scale. Cited knowledge was stable to modestly broader, revealing a divergence between focused outputs and the reach of knowledge inputs. Established contributors' prior breadth increased by 0.0190 [-0.0078, 0.0459] by 2019-2022, within a +/-0.05 equivalence bound assessed in sensitivity analysis. In mature citation windows, one standard deviation of focal depth was associated with 8.2% higher 1 + FWCI [1.9%, 14.9%]; average breadth and interaction associations were smaller under the specified equivalence bounds. Post-2022 deviations from earlier trends were not systematic, and recent changes did not vary clearly with baseline AI intensity across 83 subfields. The findings support a differentiated structure of scientific expertise in which focused individual accumulation coexists with expanding collaboration and sustained access to diverse knowledge inputs.

Summary

Main Finding

Scientific production from 2010–2022 shows a differentiated reorganization: teams grew substantially and papers became slightly more topically concentrated, while the knowledge inputs cited remained stable to modestly broader. Individual established researchers’ repertoires changed only modestly (within a small equivalence bound), and deeper prior work in a focal subfield (focal depth) is associated with higher mature citation impact. Overall, focused individual accumulation coexists with larger teams and continued access to diverse knowledge inputs — i.e., a separation of depth at the individual level from breadth in collaborative and input structures.

Key Points

  • Samples and fields: analysis covers 47,959 OpenAlex articles (six fields: arts & humanities, computer science, engineering, medicine, physics, social sciences) over 2010–2025; linked references and contributor histories were constructed for nested subsamples.
  • Team size rose markedly: modeled log(1 + team size) annual coefficient ≈ 0.01738 (95% CI [0.01602, 0.01875]). Modeled increase ≈ 37.3% (95% CI [34.4%, 40.3%]) over 2010–2022.
  • Paper topic breadth contracted slightly: annual decline ≈ 0.00120 (95% CI [–0.00166, –0.00074]) per year; cumulative change ≈ –0.0144 on a 0–1 hierarchical distance scale (2010–2022).
  • Cited knowledge breadth was stable to modestly increasing: per-year increase ≈ 0.00228 (HC3 95% CI [0.00002, 0.00454], p = 0.048; clustering by field-year attenuates significance). Coarser measures (share of out-of-field refs, effective number of cited fields) were approximately stable.
  • Individual prior breadth: established contributors’ five-year prior breadth rose modestly by 0.0190 (95% CI [–0.0078, 0.0459]); not statistically different from zero but within an equivalence bound of ±0.05 (TOST p = 0.012; 90% CI [–0.0035, 0.0416]).
  • Focal depth predicts impact: one standard deviation increase in focal depth → ~8.2% higher 1 + FWCI (95% CI [1.9%, 14.9%], p = 0.010) after adjusting for covariates. Breadth (at mean depth) and the breadth×depth interaction had small, statistically indistinct associations (breadth: 1.8% [–4.7%, 8.8%]; interaction: 0.7% [–4.9%, 6.6%]).
  • Post-2022: no systematic deviations from prior trends and no clear heterogeneity in post-2022 changes by baseline AI intensity across 83 subfields.
  • Interpretation: evidence supports a “differentiated” organization where concentrated, specialist outputs (and the citation returns to focal depth) coexist with larger collaborative teams and continued access to diverse knowledge inputs.

Data & Methods

  • Primary bibliographic source: OpenAlex.
  • Article sample: 47,959 randomly sampled articles (500 per field per year across six fields and 16 years), excluding retractions and paratext.
  • Reference sample: 3,360 sampled papers (expanded to 35 papers per field-year); resolved 51,736 cited works (≈93.7% resolution); up to 20 references per paper sampled.
  • Contributor histories: chronologically reconstructed prior publications for 1,754 index contributors; after eligibility and identity filters, 1,316 eligible contributor events and 959 complete cases for citation-return analyses.
  • Key measures:
    • Paper topic breadth: hierarchical topic distance (0–1 scale), primary metric often from papers assigned exactly three topics for comparability.
    • Cited knowledge breadth: distance among topics of resolved references.
    • Individual prior breadth: distribution of recent topics over a 5-year lookback (alternative history windows tested).
    • Focal depth: accumulated prior work in the subfield of the focal article (log scale).
    • Outcome: mature Field-Weighted Citation Impact (FWCI) with OpenAlex’s 4-year window; models used 1 + FWCI transformed and modeled on log scale.
    • Team size: log(1 + team size) and alternative log(team size).
  • Estimation details:
    • Regression models adjust for field and year fixed effects, career age, recent publication volume, team size, open access status, etc.
    • Robustness: HC3 standard errors, clustering within field-year cells for some estimates, holdouts (pilot vs expanded), alternative aggregations and weighting, ORCID subsample checks.
    • Multiple-testing control: six pre-specified primary tests, Holm adjustment for familywise error.
    • Equivalence tests: two one-sided tests (TOST) with specified bounds (e.g., ±0.05 on breadth); sensitivity bounds used in impact checks (e.g., ±log(1.10)).
    • Additional checks: predictive counterfactual forecasting and continuous difference-in-differences for post-2022 / AI-intensity analyses across 83 subfields (pre-specified exposure window).

Implications for AI Economics

  • AI tools have not (through 2022/early 2023 evidence) induced a large, measurable shift of individual scientific talent from depth toward breadth. Instead, organizational and input-side changes have been more salient (larger teams, stable or modestly broader cited inputs).
  • Returns to specialization remain important: focal subfield depth predicts higher mature citation impact. This implies that specialist human capital retains value even as AI lowers information access costs.
  • Complementarity of AI and specialization: generative AI may lower search, synthesis, and coordination costs, enabling teams and individuals to incorporate more diverse inputs without eroding the productivity premium of deep subfield expertise. In labor-market terms, demand for specialists with deep tacit knowledge likely persists; AI augments them rather than replaces the value of depth.
  • Skill-composition and training:
    • Short-to-medium term: investments in deeper disciplinary training remain warranted, alongside training in cross-domain search, prompt engineering, and tools that facilitate integrating external inputs.
    • Institutions (universities, funders, firms) should balance incentives: reward depth that drives impact (per focal depth returns) while supporting collaborative and integrative capacities (teamwork, data/tool access).
  • Organizational design and hiring:
    • Teams are growing; organizations may prioritize coordination, project management, and integrative roles that combine limited breadth with mechanisms for accessing diverse inputs (e.g., AI-enabled literature synthesis).
    • Recruitment signals: evidence supports continued premium for demonstrated focal depth in impact-sensitive roles; evaluation metrics should account for team size and collaborative contributions.
  • Funding and evaluation policy:
    • Bibliometric evaluation should distinguish between input diversity (references, sources) and output topical breadth — both matter but capture different phenomena. Funding programs that seek interdisciplinarity should explicitly define whether they target breadth of inputs (source diversity) or breadth of outputs (topic span).
    • Given the persistence of depth’s returns, funding mechanisms that reward concentrated, risky specialization remain justified alongside interdisciplinary initiatives.
  • Market and productivity implications:
    • AI adoption may raise aggregate productivity by easing access to knowledge inputs and coordination (consistent with larger teams and stable/broader cited inputs), but gains are likely complementary to — not substitutes for — accumulated specialist human capital.
    • Wage and career dynamics: unless AI fundamentally reduces the marginal value of tacit, practice-based expertise, specialists should continue to command favorable returns; generalists who combine coordination, tool literacy, and some domain competence may see rising demand.
  • Cautions for policymakers and economists:
    • Observational limits: the paper’s timeframe mostly precedes very recent waves of generative AI deployment; continued monitoring is needed to detect later shifts.
    • Field heterogeneity: effects vary by field; sectoral and subfield-specific AI exposure matters for local labor-market impacts.
    • Measurement caveats: citation impact is an imperfect proxy for social/economic value; broader metrics (translation to products, patents, applied outcomes) may show different patterns.

Limitations noted by the authors (relevant to inference): - Observational design (causal claims limited). - Coverage and measurement choices (OpenAlex coverage, topic assignments, reference resolution). - Time window: many generative-AI developments post-date the primary pre-2023 capability comparisons; effects may evolve as tools and practices diffuse.

Bottom line: up through the study window, scientific systems are showing larger teams and somewhat more concentrated outputs while still drawing on diverse inputs; individual-level breadth hasn’t demonstrably replaced depth, and depth continues to matter for citation impact. For AI economics, this suggests AI is reshaping coordination and information access more than it has (so far) reallocated the returns from specialization to generalism.

Assessment

Paper Typedescriptive Evidence Strengthmedium — Large, pre-specified, multi-level bibliometric samples with robust SEs, familywise corrections, equivalence tests, and sensitivity checks provide consistent descriptive evidence on trends; however the analysis is observational without a clear causal identification strategy linking generative AI to observed changes, and measures (topic assignments, breadth metrics, OpenAlex coverage) may contain measurement error and selection biases. Methods Rigorhigh — The authors use randomized sampling across fields/years, resolve a high fraction of references, pre-specify a family of tests with Holm correction, use robust HC3 clustering, equivalence testing, reweighting and multiple sensitivity checks (pilot/holdout/ORCID subsamples, alternative lookbacks), and link multiple complementary units of analysis (papers, cited works, contributor histories, citation returns). Remaining concerns are observational design limitations and dependence on algorithmic topic classifications. SampleRandom sample of ~47,959 OpenAlex articles (500 per field per year across six fields and 16 years 2010–2025, after exclusions). Cited-works branch: 3,360 sampled articles (3,215 eligible) with up to 20 references per paper and 51,736 resolved cited works. Contributor branch: 1,754 unique index contributors with 1,316 eligible contributor events after identity/history filters; 959 complete cases for citation-return models (focal articles through 2022, allowing mature FWCI). Themeshuman_ai_collab org_design GeneralizabilityLimited to six broad fields and authors/papers covered in OpenAlex; not necessarily representative of all disciplines (e.g., clinical, industrial R&D, non-English literature)., Relies on algorithmic topic assignments and hierarchical distance measures that may misclassify interdisciplinarity or intellectual distance., Observational trends may reflect contemporaneous changes (publication practices, indexing coverage, team authorship norms) not causal effects of generative AI., Contributor sample selects one author per paper and imposes publication-history filters, which may underrepresent early-career or highly mobile researchers., Citation-based impact (FWCI) is field-adjusted but still subject to citation practice heterogeneity and time-varying citation dynamics.

Claims (10)

ClaimDirectionOutcomeConfidence & EvidenceDetails
From 2010 to 2022, paper topic breadth declined by 0.0144 on a 0–1 hierarchical distance scale. Research Productivity negative Paper topic breadth
Reading fidelity high
Study strength high
n=47959
-0.0144 on a 0–1 hierarchical distance scale
0.3
Team size increased substantially from 2010 to 2022, with a modelled increase of 37.3%. Team Performance positive Scientific research team size
Reading fidelity high
Study strength high
n=38285
37.3% increase [34.4%, 40.3%]
0.3
Cited knowledge breadth increased slightly over time, while coarser measures of cited-field diversity were approximately stable. Research Productivity mixed Breadth and diversity of cited knowledge inputs
Reading fidelity high
Study strength medium
n=2619
0.00228 per year [0.00002, 0.00454] in the unclustered estimate
0.18
Established contributors' five-year prior topic breadth was only modestly higher in 2019–2022 than in 2010–2014, and the change was within the prespecified ±0.05 equivalence bound. Skill Acquisition null_result Established contributors' prior topic breadth
Reading fidelity high
Study strength medium
n=1292
0.0190 on the 0–1 breadth scale [–0.0078, 0.0459]
0.18
Focal depth was not clearly changed between 2019–2022 and 2010–2014. Skill Acquisition null_result Accumulated individual depth in the focal subfield
Reading fidelity high
Study strength medium
n=1292
-0.0274 [–0.1451, 0.0903]
0.18
Among contributor events with mature citation windows, one standard deviation greater focal depth was associated with 8.2% higher 1 + FWCI. Research Productivity positive Field-weighted citation impact of the focal article
Reading fidelity high
Study strength high
n=959
8.2% higher 1 + FWCI per one standard deviation increase in focal depth [1.9%, 14.9%]
0.3
Prior breadth had no statistically distinguishable association with mature citation impact at mean focal depth. Research Productivity null_result Field-weighted citation impact of the focal article
Reading fidelity high
Study strength medium
n=959
1.8% [–4.7%, 8.8%]
0.18
The interaction between prior breadth and focal depth was not statistically distinguishable from zero in its association with mature citation impact. Research Productivity null_result Interaction of prior breadth and focal depth on field-weighted citation impact
Reading fidelity high
Study strength medium
n=959
0.7% [–4.9%, 6.6%]
0.18
Post-2022 deviations from trends forecast using earlier years were not statistically distinguishable from zero. Other null_result Post-2022 deviation from forecasted scientific capability trends
Reading fidelity high
Study strength medium
n=47959
-0.0108 [-0.026, 0.004]
0.18
Recent changes did not vary clearly with baseline AI intensity across 83 subfields. Automation Exposure null_result Differential post-2022 change by baseline AI intensity
Reading fidelity high
Study strength low
n=83
0.09

Notes