The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References About 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

Since 2022 scientists have moved into more distant fields and formed more interdisciplinary collaborations as LLM use rises, with AI-writing adopters widening an existing lead; at the same time research teams show sharper division of labor, shifting toward software and validation roles and away from conceptual and managerial roles.

Scientific exploration, collaboration and labor division in the large language model era
Xiang Zheng, Xi Hong, Jialin Liu, Chaoqun Ni · July 23, 2026
arxiv correlational medium evidence 8/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Xiang Zheng unresolved corpus identity
  2. Xi Hong unresolved corpus identity
  3. Jialin Liu unresolved corpus identity
  4. Chaoqun Ni unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Xiang Zheng provider ID
  2. Xi Hong provider ID
  3. Jialin Liu provider ID
  4. Chaoqun Ni provider ID
Around and after 2022, scientific work became more interdisciplinary and exploratory and teams showed greater role specialization—changes that are concentrated among authors with stronger text-based signals of AI writing—though these results are associational rather than causal.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

Large language models (LLMs) have rapidly and significantly entered scientific workflows, but it remains unclear how their diffusion is associated with changes in scientists' strategies in research directions and team building. We link PubMed Central full text with OpenAlex publication and collaboration histories for 775,323 scientists and analyze CRediT contribution statements from 137,120 multi-author papers. After 2022, scientists increasingly published across more intellectually distant fields and entered fields in which they had not previously worked. These increases in interdisciplinarity and exploration were especially pronounced among established scientists and scientists from non-English-speaking low- and middle-income countries. Authors with stronger AI-writing signals were already more interdisciplinary and exploratory before the widespread adoption of LLMs, and the gap widened further after 2022 compared with authors with weaker AI-writing signals. Scientists' collaboration networks also became more interdisciplinary after 2022. Yet, among authors with stronger AI-writing signals, research interdisciplinarity was less closely tied to the disciplinary diversity of their collaborators. The division of labor within research teams also became more differentiated. Contributors on papers published after 2022 reported narrower role sets on average, coauthors shared fewer roles in common, and their role profiles became less rigid and more fluid. Software and validation roles increased, while conceptual and management roles decreased. These patterns suggest that team members are taking on more distinct responsibilities and may rely less on one another to perform research tasks. Overall, this study indicates that the LLM era coincides with a broader reorganization of scientific exploration, collaboration, and the division of labor.

Summary

Main Finding

The public diffusion of large language models (LLMs) after late 2022 is associated with a broad reorganization of scientific work: scientists (especially established researchers and those in non‑English, low/middle‑income countries) published more interdisciplinarily and explored new fields more often; collaboration networks became more disciplinary‑diverse; and within teams, contribution roles became more specialized and task‑differentiated. Authors and papers with stronger LLM‑style writing signals show larger increases, although the design is descriptive and not causal.

Key Points

  • Expansion of portfolios (population-level patterns, 2011–2025):
    • Average number of distinct primary fields per author rose from 2.3 (2022) to 2.8 (2025).
    • Field Shannon entropy and Rao–Stirling interdisciplinarity increased (Rao–Stirling 0.176 → 0.199 by 2025; ~13% above pre‑2023 trend).
    • Share of authors entering at least one new field rose from 30.4% (2022) to 35.5% (2025).
    • Average distance between papers and an author’s primary field increased (~13% above trend by 2025).
  • Exploration and topic dispersion:
    • References became more cross‑field (share of references outside citing paper’s field rose; topical Gini declined → publications spread across more topics).
    • More concurrent publications spanning biomedical + engineering/technical fields.
  • Heterogeneity:
    • Advanced (senior) scientists show the clearest post‑2022 pivot increases; junior scientists show smaller changes.
    • Non‑English, low/middle‑income country authors exhibit the largest increases in exploration/interdisciplinarity.
  • AI‑writing intensity:
    • Paper‑level AI‑writing fraction (language resemblance to LLM‑assisted text) rose from ~0.05 (2021/2022 baseline) to ~0.19 (2025).
    • Authors classified as high‑AI‑writing (avg paper fraction >0.15) were already more interdisciplinary before 2023 and their advantage widened after 2022.
    • High‑AI‑writing authors became more productive post‑2022 relative to matched low‑AI peers.
  • Collaboration networks:
    • Average number of distinct collaborator fields per author increased (expected 4.4 → observed 5.2 by 2025).
    • Share of collaborators outside an author’s primary field rose (40.9% → 43.7%).
    • Collaborator‑based Rao–Stirling index increased (~12% above pre‑2023 trend by 2025).
    • High‑AI‑writing authors had higher collaborator interdisciplinarity both before and after 2022; however, for them the link between their research interdisciplinarity and collaborator disciplinary diversity weakened.
  • Division of labor (from CRediT statements; 137,120 multi‑author papers):
    • Contributors reported narrower role sets on average; coauthors shared fewer roles in common.
    • Role profiles became less rigid and more fluid across papers.
    • Increases in roles like software and validation; decreases in conceptual and managing roles.
    • Interpretation: team members take on more distinct, specialized responsibilities and may rely less on coauthors for some tasks.

Data & Methods

  • Data sources:
    • PubMed Central (full text) linked to OpenAlex author publication and collaboration histories.
    • CRediT contribution statements drawn from PLOS and PMC (137,120 multi‑author papers).
    • Sample: analyses across ~775,323 scientists; time window 2011–2025, with focus on changes after 2022/2023.
  • Key measures:
    • Portfolio breadth: count of distinct primary fields, Shannon entropy, Herfindahl‑Hirschman Index.
    • Interdisciplinarity: Rao–Stirling index (field distances incorporated), share of out‑of‑field references.
    • Exploration/pivot: document‑level pivot size based on divergence from an author’s recent reference profile; count of new fields entered.
    • Collaboration composition: number of distinct collaborator fields, share of out‑of‑field collaborators, collaborator Rao–Stirling.
    • AI‑writing fraction: per‑paper score estimating resemblance to LLM‑assisted writing (language features); author AI‑writing rate = mean of post‑2022 paper fractions. Authors categorized as high (>0.15) or low (<0.05).
    • Team labor: CRediT roles frequency, role overlap among coauthors, role set size, role rigidness/fluidity metrics.
  • Identification/contrast strategy:
    • Descriptive temporal comparisons (pre‑2023 trend extrapolation vs observed post‑2022).
    • Matched comparisons (coarsened exact matching on career stage, primary field, affiliation country group, productivity bins) for high vs low AI‑writing authors.
    • Robustness: alternative disciplinary classifications, exclusion of low‑productivity authors, author fixed effects and publication‑count controls, bootstrap confidence intervals.
  • Limitations called out by authors:
    • Study is descriptive—not a causal estimate of LLM use.
    • AI‑writing fraction is a proxy (language resemblance) and may reflect correlated unobserved traits (early adopters, resourcing, prior interdisciplinarity).
    • Data centered on biomedical and adjacent fields (PMC, PLOS) → generalizability to all sciences should be cautious.

Implications for AI Economics

  • Reduced coordination and knowledge‑acquisition costs: LLMs (or tools associated with LLM‑like writing signals) are plausibly lowering language, literature‑search, and translational costs that impede cross‑field exploration, which can shift the returns to exploration vs exploitation for researchers and institutions.
  • Skill and task reallocation:
    • Increased demand for technical implementation and validation skills (software, validation roles) and reduced relative demand for some conceptual/managerial tasks suggests reweighting of task complementarities within teams — relevant for hiring, training, and compensation.
    • Automation/substitution possibilities for writing/communication tasks may change bargaining power and wages for roles centered on drafting and editing; conversely, tasks requiring domain insight, experimental design, or high‑stakes judgment may retain or increase premium.
  • Distributional effects and international leveling:
    • Larger portfolio and exploration gains among non‑English, lower‑income countries suggest LLMs may lower language and access barriers, potentially narrowing some international research gaps. This has implications for global human capital accumulation and location choices for research talent.
  • Returns to interdisciplinarity and team composition:
    • If interdisciplinarity becomes less dependent on collaborator diversity for high‑AI users, returns to assembling diverse teams may shift (e.g., more solo pivots or looser team complementarities). Funders and institutions may need to rethink incentives for team formation and credit allocation.
  • Research productivity and incentives:
    • Associations between higher AI‑writing intensity and increased productivity/interdisciplinarity could alter evaluation metrics (publications, topical breadth). Policymakers and funders should account for tool‑mediated changes in productivity when designing performance evaluations and tenure/promotion rules.
  • Measurement and crediting:
    • More differentiated contribution roles (CRediT changes) call for updated measurement of research inputs and more granular attribution in funding, hiring, and reward systems to reflect shifting labor division.
  • Caution for policy and markets:
    • Evidence is correlational: policy responses (training, regulation, intellectual property, authorship rules) should be robust to uncertainty about causality.
    • Potential unintended consequences: if LLMs lower costs of exploring many fields, this could dilute depth in some areas or change the nature of expertise accumulation; labor market effects (displacement vs upskilling) will depend on how tasks evolve and on institutional incentives.

Overall, the paper provides large‑scale descriptive evidence that the LLM era coincides with broader changes in what scientists study, whom they involve, and how teams divide labor. For economists, these patterns point to evolving task structures, changing returns to specialization vs exploration, and distributional shifts across career stages and countries — all of which merit targeted causal research and policy consideration.

Assessment

Paper Typecorrelational Evidence Strengthmedium — Very large, rich observational datasets (hundreds of thousands of authors; >137k papers with CRediT statements) and careful measurement of interdisciplinarity and role composition provide credible descriptive evidence of correlated change around 2022, but causal inference is limited by reliance on a proxy for LLM use, potential unobserved confounders, contemporaneous shocks, and possible selection into the sample (CRediT availability, PMC coverage). Methods Rigormedium — The study links multiple high-quality bibliometric sources, uses granular CRediT contribution data and network measures, and implements comparative pre/post and between-group analyses; however, rigor is curtailed by reliance on an imperfect text-based AI-use indicator, likely observational confounding, and no clear exogenous identification strategy to rule out alternative explanations. SampleLinked PubMed Central full text and OpenAlex publication/collaboration histories covering 775,323 scientists and 137,120 multi-author papers with CRediT contribution statements; analysis compares patterns before and after 2022 and stratifies authors by a text-derived AI-writing signal; includes authors across many fields and countries with identifiable affiliation/language metadata (including non-English-speaking low- and middle-income countries). Themeshuman_ai_collab innovation IdentificationAssociational pre/post analysis around 2022 combined with cross-sectional stratification by a text-based "AI-writing signal" proxy; compares changes in interdisciplinarity, exploration, collaboration networks, and contribution roles before and after widespread LLM adoption and between authors with stronger vs weaker AI-writing signals, with adjustment for observable covariates (no random assignment or exogenous instrument). GeneralizabilityBiomedicine-heavy sample due to PubMed Central coverage — findings may not generalize to social sciences, physical sciences, or humanities., Only includes papers with CRediT statements and PMC full text, creating selection bias toward journals/papers that report contributions and toward open-access content., AI-writing signal is a proxy derived from text and may misclassify actual LLM use (false positives/negatives), limiting inference about individual-level adoption., Temporal confounding: other post-2022 factors (pandemic-related shifts, funding trends, policy changes) could drive observed changes., Likely underrepresents industry research and non-published forms of scientific work (preprints, internal reports).

Claims (10)

ClaimDirectionOutcomeConfidence & EvidenceDetails
After 2022, scientists increasingly published across more intellectually distant fields and entered fields in which they had not previously worked. Research Productivity positive interdisciplinarity and field-entry (novel fields entered)
Reading fidelity high
Study strength medium
n=775323
0.3
These increases in interdisciplinarity and exploration were especially pronounced among established scientists and scientists from non-English-speaking low- and middle-income countries. Research Productivity positive interdisciplinarity and exploration (by subgroup)
Reading fidelity high
Study strength medium
n=775323
0.3
Authors with stronger AI-writing signals were already more interdisciplinary and exploratory before the widespread adoption of LLMs. Research Productivity positive interdisciplinarity and exploration (pre-2022 differences by AI-writing signal)
Reading fidelity high
Study strength medium
not reported
0.3
The gap in interdisciplinarity and exploration between authors with stronger versus weaker AI-writing signals widened further after 2022. Research Productivity positive change in interdisciplinarity and exploration gap (strong vs weak AI-writing signal)
Reading fidelity high
Study strength medium
not reported
0.3
Scientists' collaboration networks became more interdisciplinary after 2022. Team Performance positive collaboration network interdisciplinarity (disciplinary diversity of collaborators)
Reading fidelity high
Study strength medium
n=775323
0.3
Among authors with stronger AI-writing signals, research interdisciplinarity was less closely tied to the disciplinary diversity of their collaborators. Team Performance negative strength of association between individual interdisciplinarity and collaborators' disciplinary diversity
Reading fidelity high
Study strength medium
not reported
0.3
The division of labor within research teams became more differentiated after 2022: contributors reported narrower role sets on average, coauthors shared fewer roles in common, and their role profiles became less rigid and more fluid. Task Allocation mixed team division of labor indicators (roles per contributor, role overlap, role rigidity/fluidity)
Reading fidelity high
Study strength medium
n=137120
0.3
Software and validation roles increased after 2022, while conceptual and management roles decreased. Task Allocation mixed frequency of CRediT roles (software, validation, conceptual, management)
Reading fidelity high
Study strength medium
n=137120
0.3
These patterns suggest that team members are taking on more distinct responsibilities and may rely less on one another to perform research tasks. Organizational Efficiency negative inference about interdependence among team members (reliance on one another)
Reading fidelity high
Study strength speculative
not reported
0.05
Overall, the LLM era coincides with a broader reorganization of scientific exploration, collaboration, and the division of labor. Research Productivity mixed aggregate changes in exploration, collaboration, and division of labor
Reading fidelity high
Study strength medium
n=775323
0.3

Notes