The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

Lab groups treat an AI teammate like a human one: an online public-goods experiment finds cooperation determined by group reciprocity and behavioural inertia, not by whether a programmed partner is labeled 'AI' or 'human', suggesting cooperative norms transfer to artificial agents.

Normative Equivalence in Human-AI Cooperation: Behaviour, Not Identity, Drives Cooperation in Mixed-Agent Groups
Nico Mutzner, Taha Yasseri, Heiko Rauhut · January 28, 2026
arxiv rct medium evidence 7/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Nico Mutzner unresolved corpus identity
  2. Taha Yasseri unresolved corpus identity
  3. Heiko Rauhut unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Nico Mutzner provider ID
  2. T. Yasseri provider ID
  3. Heiko Rauhut provider ID
In online four-player public goods experiments, cooperation was driven by reciprocal group dynamics and inertia rather than whether a scripted partner was labeled 'AI' or 'human', producing similar cooperation levels and norm persistence across labels.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

The introduction of artificial intelligence (AI) agents into human group settings raises essential questions about how these novel participants influence cooperative social norms. While previous studies on human-AI cooperation have primarily focused on dyadic interactions, little is known about how integrating AI agents affects the emergence and maintenance of cooperative norms in small groups. This study addresses this gap through an online experiment using a repeated four-player Public Goods Game (PGG). Each group consisted of three human participants and one bot, which was framed either as human or AI and followed one of three predefined decision strategies: unconditional cooperation, conditional cooperation, or free-riding. In our sample of 236 participants, we found that reciprocal group dynamics and behavioural inertia primarily drove cooperation. These normative mechanisms operated identically across conditions, resulting in cooperation levels that did not differ significantly between human and AI labels. Furthermore, we found no evidence of differences in norm persistence in a follow-up Prisoner's Dilemma, or in participants' normative perceptions. Participants' behaviour followed the same normative logic across human and AI conditions, indicating that cooperation depended on group behaviour rather than partner identity. This supports a pattern of normative equivalence, in which the mechanisms that sustain cooperation function similarly in mixed human-AI and all human groups. These findings suggest that cooperative norms are flexible enough to extend to artificial agents, blurring the boundary between humans and AI in collective decision-making.

Summary

Main Finding

When a programmed agent joined three humans in a repeated four-player Public Goods Game, cooperation was driven by observed group behaviour (reciprocity, conditional responses, and behavioural inertia), not by whether the agent was labelled “human” or “AI.” Contributions, norm perceptions, and carry-over cooperation in a follow-up Prisoner’s Dilemma did not differ meaningfully by agent label. The authors term this result “normative equivalence”: the social-norm mechanisms that sustain cooperation operate similarly in mixed human–AI and all-human groups under the study conditions.

Key Points

  • Experimental design: 2 × 3 between-subjects (agent label: human vs. AI) × (bot strategy: unconditional cooperator, conditional cooperator, free-rider). Groups: 3 humans + 1 bot.
  • Sample: final N = 236 participants (near-equal across human vs. AI label). Preregistered; some deviations in power/sample justification reported transparently.
  • Behavioural outcome: Ten rounds of a linear Public Goods Game (endowment 100 tokens, contributions multiplied by 1.5 and split equally). Aggregate contributions began ~40–50 tokens and declined to ~30–40; trajectories largely overlapped across labels and strategies.
  • No meaningful label effect: Differences between human- and AI-labelled conditions were small (90% CI rules out label effects > ~4 tokens) and not statistically substantive.
  • Bot-strategy effects: Small deviations by bot strategy (unconditional cooperator slightly higher contributions; conditional and free-rider slightly lower), but trajectories overlapped and effects were negligible in substantive terms.
  • Norm persistence: One-shot Prisoner’s Dilemma (partner behaviour simulated) showed no systematic differences by label or strategy in carry-over cooperation.
  • Norm perceptions: Krupka–Weber appropriateness ratings, empirical and injunctive expectation measures aligned across conditions; participants’ normative beliefs matched their behaviour regardless of label.
  • Mechanisms observed: responsiveness to group contributions (conditional cooperation), within-player inertia (stable individual tendencies), and gradual decline over rounds—these governed behaviour equally in human and AI label conditions.
  • Robustness: Excluding participants who doubted the manipulation (~18.6%) did not change substantive results.

Data & Methods

  • Platform and recruitment: Online experiment run in oTree; participants recruited on Prolific; payment: fixed fee plus token-based bonus (average effective hourly payment ≈ £8.95).
  • Treatment manipulation: Fourth player was a computer-controlled agent; label shown to participants as either “human” or “AI.” Bot strategies (not disclosed to participants): 1) Unconditional cooperator (always contributes 100), 2) Conditional cooperator (matches group average contribution from previous round), 3) Free-rider (always contributes 0).
  • Observability: Participants saw only aggregate group contribution each round (no individual-level contributions), mimicking opacity in large collective action.
  • Outcome measures:
    • PGG contributions (0–100 tokens) across 10 rounds; analyzed with mixed-effects regressions and mean-contribution linear models.
    • One-shot Prisoner’s Dilemma (binary cooperate/defect); analyzed with logistic regression.
    • Norm elicitation: Krupka–Weber-style appropriateness ratings (0,25,50,75,100), plus empirical and injunctive expectation measures.
    • Post-task survey indices: trust, fairness, cohesion, normative pressure; AI-specific items where applicable.
  • Pre-registration: Registered (AsPredicted #234846). Deviations: adjusted sample/power approach after pilot effect sizes close to zero; full documentation provided.
  • Sample size & precision: Final N provided ≈ 40 observations per cell; authors report sensitivity allowing exclusion of large label effects.

Implications for AI Economics

  • Modeling and policy: In low-social-presence group contexts where agents’ behaviour is visible only at an aggregate level, treating AI agents as functionally equivalent to humans for norm dynamics may be appropriate. Economic models of collective action and public-good provision that include AI participants should emphasize behavioural rules over categorical identity labels.
  • AI design for cooperation: Designers should prioritize norm-consistent behavioural policies (e.g., reciprocation, conditional cooperation) to foster integration into human groups rather than relying solely on anthropomorphic framing or human-like labels.
  • Institutional deployment: Organizations introducing autonomous agents into team or collective settings can potentially rely on behavioural signals (consistent, reciprocal actions) to maintain cooperation; transparency about AI behaviour may be more consequential than human-like presentation.
  • Cautions and boundary conditions:
    • External validity: The experiment used non-adaptive, scripted bots, minimal social presence, no communication, and an online Prolific sample. Effects may differ with adaptive/learning agents, richer social cues, repeated cross-group interactions, higher stakes, or field settings.
    • Agency, accountability, and trust: The lack of label effects in this setting does not remove normative, ethical, or legal reasons to treat AI and humans differently (e.g., accountability, liability, algorithmic bias).
  • Directions for future economic research: test when normative equivalence breaks—introduce adaptive/strategic AI, communication channels, visibility of individual actions, higher stakes, reputation systems, or mixed incentives; examine long-term dynamics and market/organizational outcomes when AIs are integrated into real-world collective-action problems (climate agreements, resource allocation, coordinated markets).

Assessment

Paper Typerct Evidence Strengthmedium — Internal validity is strong because treatment (agent label and strategy) appears randomized, allowing credible causal claims about the immediate effect of labeling/strategy on cooperation; however, the study reports a null effect, sample size is moderate (limiting power to detect small effects or to establish equivalence), and external validity is limited because the setting is an online, short-duration laboratory game with simple scripted agents rather than deployed, adaptive AI systems. Methods Rigormedium — The experimental design (random assignment of label and strategy, repeated PGG and follow-up PD) is appropriate and provides good control over confounds, but the description lacks information on pre-registration, power/equivalence testing, participant recruitment details, checks on label credibility, and robustness checks, which would be needed to raise the rigor rating to high. SampleOnline experiment with N=236 human participants recruited online (platform not specified), organized into repeated four-player public goods games composed of 3 human players plus 1 programmed agent (bot) which was randomly assigned a 'human' vs 'AI' label and one of three fixed decision strategies; participants also played a follow-up Prisoner's Dilemma to assess norm persistence. Themeshuman_ai_collab org_design IdentificationRandomized between-subjects experiment: groups of three human participants were paired with a single programmed agent whose label (framed as 'human' or 'AI') and decision strategy (unconditional cooperator, conditional cooperator, or free-rider) were randomly assigned; causal effects of agent identity/strategy are identified by comparing cooperation outcomes across these randomized treatment arms, with additional tests using a follow-up Prisoner's Dilemma to probe norm persistence. GeneralizabilityConvenience online sample (platform, demographics not specified) may not represent broader populations or workplace teams, Small, stylized laboratory games (repeated PGG and one-shot PD) abstract from real-world organizational complexity and stakes, Programmed bots with simple, fixed strategies do not capture adaptive, learning, or anthropomorphic behaviours of deployed AI systems, Short-term interactions limit inference about long-run norm formation and institutional adoption, Framing used for 'AI' vs 'human' labels may not map onto real-world trust, transparency, or accountability cues

Claims (10)

ClaimDirectionOutcomeConfidence & EvidenceDetails
We ran an online experiment using a repeated four-player Public Goods Game (PGG) in which each group consisted of three human participants and one bot. Other positive experimental_design_implementation
Reading fidelity high
Study strength high
not reported
1.0
The bot was framed either as human or AI and followed one of three predefined decision strategies: unconditional cooperation, conditional cooperation, or free-riding. Other positive treatment_manipulation
Reading fidelity high
Study strength high
not reported
1.0
The sample included 236 participants. Other positive sample_size
Reading fidelity high
Study strength high
n=236
1.0
Reciprocal group dynamics and behavioural inertia primarily drove cooperation. Team Performance positive cooperation levels / drivers of cooperation
Reading fidelity high
Study strength medium
n=236
0.6
These normative mechanisms operated identically across conditions, resulting in cooperation levels that did not differ significantly between human and AI labels. Team Performance null_result cooperation levels
Reading fidelity high
Study strength medium
n=236
0.6
We found no evidence of differences in norm persistence in a follow-up Prisoner's Dilemma. Team Performance null_result norm persistence (in follow-up Prisoner's Dilemma)
Reading fidelity high
Study strength medium
n=236
0.6
We found no evidence of differences in participants' normative perceptions between human- and AI-labelled conditions. Other null_result participants' normative perceptions
Reading fidelity high
Study strength medium
n=236
0.6
Participants' behaviour followed the same normative logic across human and AI conditions, indicating that cooperation depended on group behaviour rather than partner identity. Team Performance null_result behavioral conditionality / drivers of cooperation
Reading fidelity high
Study strength medium
n=236
0.6
This supports a pattern of normative equivalence, in which the mechanisms that sustain cooperation function similarly in mixed human-AI and all human groups. Team Performance null_result mechanisms sustaining cooperation / normative equivalence
Reading fidelity high
Study strength medium
n=236
0.6
These findings suggest that cooperative norms are flexible enough to extend to artificial agents, blurring the boundary between humans and AI in collective decision-making. Team Performance positive extension of cooperative norms to artificial agents
Reading fidelity high
Study strength speculative
n=236
0.1

Notes