The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

An AI teammate talks more but adds less: in a randomized chat-based lab study the AI dominated conversation yet provided the least new information, and its presence reduced human teammates' responsiveness and sense of belonging and status from the outset.

The Social Cost of an AI Teammate: How an Artificial Teammate Reshapes Human-Human Communication in Small-Team Decision-Making
Nia Nixon, Jaeyoon Choi, Pedro Martins De Bastos, Mohammad Amin Samadi, Luise Mehner, Seehee Park, Spencer JaQuay · July 29, 2026
arxiv rct medium evidence 7/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Nia Nixon unresolved corpus identity
  2. Jaeyoon Choi unresolved corpus identity
  3. Pedro Martins De Bastos unresolved corpus identity
  4. Mohammad Amin Samadi unresolved corpus identity
  5. Luise Mehner unresolved corpus identity
  6. Seehee Park unresolved corpus identity
  7. Spencer JaQuay unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Nia Nixon provider ID
  2. Jaeyoon Choi provider ID
  3. Pedro Martins De Bastos provider ID
  4. M. Samadi provider ID
  5. Luise Mehner provider ID
  6. Seehee Park provider ID
  7. Spencer Jaquay provider ID
In a randomized lab study of student teams, an AI teammate monopolized airtime and self-cohesion while contributing less novel, dense content, and its presence reduced human-human responsivity and participants' reported belonging and status immediately from the conversation's start.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

Conversational AI is increasingly positioned as a teammate rather than a tool, yet we know little about how its presence reshapes communication among the humans on the team. We examined sociocognitive communication dynamics in team decision-making using Group Communication Analysis (GCA), team surveys, and lexical analyses of team discourse. Teams completed a high-stakes moral-dilemma decision task in a randomized controlled study: 16 teams of two students plus an AI teammate, and 17 all-human teams of three. Across six GCA dimensions and survey outcomes, we find that the AI teammate was the single most talkative and self-cohesive member of every treatment team, yet its contributions carried the least new information and the lowest density. The presence of AI also reshaped communication amongst humans. In AI-human teams, human teammates showed lower responsivity and social impact toward one another and reported lower levels of belonging and status. Greater AI dominance in the conversation was associated with students feeling less valued as team members. Additionally, this social cost is immediate and present at baseline; it does not emerge over the course of the conversation. Drawing on these results, we discuss a research agenda extending to voice-based and longitudinal settings.

Summary

Main Finding

An AI teammate that is highly talkative and self-cohesive displaces human-human conversational ties without adding much new content. In teams where an AI replaced a human, the remaining human members were less responsive to one another, reported lower belonging and status, and felt less valued when the AI dominated airtime. This social cost appears immediately (present at baseline) rather than emerging over the discussion.

Key Points

  • Experimental sample: 80 undergraduates in 33 teams (16 treatment teams: 2 students + 1 AI; 17 control teams: 3 students).
  • AI implementation: Google Gemini 2.5 Flash Lite, persona “Clever Lamarr” (peer-style prompts; limited tokens).
  • Task: text-chat small-team decision on a high-stakes mountain rescue with a mid-task moral-information reveal.
  • GCA profile of the AI (within-treatment comparisons):
    • Ranked first on participation in all 16 treatment teams (AI = most talkative; paired difference vs. students +0.236 units, p < .001).
    • Ranked first or tied first on internal cohesion in 14/16 teams (paired diff +0.076, dz = 1.12, p < .001).
    • Contributed significantly less new information (newness paired diff −0.100, dz = −1.37, p < .0001) and lower information density (density paired diff −0.174, dz = −1.42, p < .0001).
    • Was not statistically different from students on responsivity or social impact (i.e., its turns were not taken up more or less than a typical student’s).
  • Human-human effects (between-condition comparisons):
    • Students in AI teams showed reduced responsivity and social impact toward one another versus students in all-human teams.
    • Students in AI teams reported lower belonging, lower perceived status/equality of voice, and lower felt value; greater AI airtime dominance predicted lower felt value.
    • The reduction in human-human coupling was observable from the conversation start (baseline), not only as a dynamic effect during the task.
  • Robustness & diagnostics:
    • Analyses include GCA, post-task team-experience surveys, LIWC-based socioemotional measures, linguistic style matching, modal-stance Markov analysis, and recurrence quantification for temporal patterns.
    • Between-condition tests used Bonferroni correction; sensitivity analyses (female-only) were run due to sample gender imbalance.

Data & Methods

  • Design: randomized between-teams lab experiment; chat-based interface (TRAIL platform) that logs millisecond timestamps and injects controlled AI persona behavior.
  • Participants: 80 undergraduates (63 female, 12 male; gender imbalance noted).
  • Conditions: 16 teams with AI + 2 students vs. 17 human-only teams with 3 students.
  • Measures:
    • Group Communication Analysis (GCA) — six dimensions per participant: participation, internal cohesion, overall responsivity, social impact, newness, communication density.
    • Team-experience surveys — composites for belonging and perceived status, plus items on felt value and perceived dominance.
    • Lexical/time-series probes — LIWC socioemotional categories, linguistic style matching (function words), modal-stance coding, recurrence quantification.
  • AI behavior: two-stage pipeline deciding whether to respond and then generating messages from the fixed persona; generation parameters constrained (temperature 0.5, max 50 tokens).
  • Analyses:
    • RQ1: within-treatment paired comparisons (AI vs students) and rank positions.
    • RQ2: between-condition comparisons of per-student GCA and survey outcomes (Welch’s t-tests, Hedges’ g), regression of felt value on AI airtime dominance.
    • RQ3: event-locked sliding-window and sequence analyses around the moral-information reveal to evaluate whether effects were baseline vs. emergent.
  • Limitations noted by authors: group-size asymmetry (AI replaced a human), single AI persona/role, short text-chat task, gender imbalance in sample, and laboratory student population.

Implications for AI Economics

  • Social externalities of AI teammates
    • Beyond task performance, AI can impose immediate negative externalities on the team’s social capital: reduced mutual uptake, belonging, and perceived status among human workers. Standard productivity metrics may miss these relational costs.
  • Labor productivity and quality-adjusted output
    • An AI that increases conversational volume but adds little new information could reduce the efficiency of human-to-human coordination and knowledge-building. Economists estimating gains from AI should adjust for potential declines in tacit knowledge exchange, learning-by-doing, and the unpriced value of psychological safety.
  • Incentives, compensation, and retention
    • Lowered belonging and perceived status may reduce voice, effortful contribution, or retention. Firms adopting interactive AI teammates may face hidden turnover or morale costs; compensation and incentive schemes should consider relational harms.
  • Team composition and assignment
    • Substituting AI for a human changes team composition in ways that have distributional effects on human workers’ airtime and influence. Allocation of tasks, career-development opportunities, and evaluative visibility could be reshaped by AI presence—affecting wage dynamics and promotion paths.
  • Design and regulation
    • Economic returns to AI deployment hinge on design choices: calibrating AI “airtime,” information novelty, and responsiveness to avoid crowding-out human interaction. Policymakers and firms should consider guidelines or standards (e.g., transparency of AI role, limits on autonomous participation) to internalize social costs.
  • Measurement and evaluation
    • Incorporate sociocognitive metrics (e.g., GCA-derived responsivity, social impact, belonging surveys) into cost-benefit analyses of AI deployment. Longitudinal measurement is especially important since baseline displacement appears immediate but cumulative effects on productivity, learning, and turnover require tracking.
  • Macro-level consequences
    • Widespread deployment of conversational AI that dominates interaction could systematically erode workplace social capital across sectors, with implications for aggregate productivity, skill accumulation, and inequality (if some workers are more likely to be displaced from relational roles).
  • Research priorities for economists
    • Quantify welfare-relevant magnitudes: translate GCA/social-cost metrics into monetary impacts (productivity loss, turnover costs, training deficits).
    • Study heterogeneity: industries, task types (creative vs. rote), team sizes, demographics, and institutional contexts where the social cost is larger or smaller.
    • Evaluate mitigation strategies: constrained AI participation, role-specific AI (assistant vs. peer), and compensation/organizational remedies.
    • Model dynamic effects: how short-run social-cost shocks influence long-run human capital accumulation and market-level labor supply/demand.

Practical suggestions for economists advising firms or regulators - When piloting conversational AI teammates, measure relational outcomes (belonging, status, mutual responsivity) alongside task performance. - Avoid substituting an AI for a person in roles where human-human uptake, learning, or psychological safety are central; instead trial AI as limited-support tools or as assistants with reduced autonomous airtime. - Incorporate potential social-cost estimates into ROI calculations, and consider mitigation budget (training, team redesign, monitoring). - Support longer-run field studies and causal inference on turnover, productivity, and wage effects to inform policy and firm strategy.

Assessment

Paper Typerct Evidence Strengthmedium — The paper uses an experimental design with random assignment, validated computational measures (GCA), and complementary survey and temporal analyses, which supports internal validity for short-term effects; however the sample is modest (33 teams, 80 students), there is a structural group-size confound between conditions, strong sample composition limits (mostly female undergraduates at a single university), a single task modality (text chat) and one AI persona/model, and outcomes are proximal (communication and self-reports) rather than downstream productivity/economic outcomes. Methods Rigormedium — Strengths: randomized assignment, pre-registered analytic rigor not reported but analyses use appropriate tests (paired within-team contrasts, Welch's t-tests, effect sizes, Bonferroni correction), validated computational pipeline (GCA), temporal/event-locked analyses, and sensitivity checks. Weaknesses: group-size confound (2 humans + AI vs 3 humans) complicates between-condition causal claims, gender imbalance across conditions, single-site convenience sample, single AI model/persona and chat-only interface, relatively small number of teams limiting power and heterogeneity analyses. Sample80 undergraduate students (Fall 2025, UC Irvine) organized into 33 teams: 16 treatment teams (two students + one AI 'Clever Lamarr' powered by Google Gemini 2.5 Flash Lite) and 17 control teams (three students). Participants communicated via text chat on a search-and-rescue moral decision task; measures include turn-level chat logs, GCA-derived conversational metrics, LIWC-derived lexical measures, modal-stance coding, and post-task team-experience surveys; gender distribution skewed (63 female, 12 male), with only three males in treatment. Themeshuman_ai_collab org_design IdentificationBetween-teams randomized assignment to treatment (2 students + AI teammate) versus control (3 students); causal inference derives from randomization and within-treatment paired comparisons (AI vs its human teammates) plus per-student outcome comparisons across conditions. Temporal/event-locked analyses (pre/post moral reveal) probe dynamics. Authors note and partially adjust for a group-size confound (treatment teams have two humans vs. three in control) by focusing on per-student measures, z-standardizing GCA scores, running female-only sensitivity checks, and emphasizing within-team contrasts for RQ1. GeneralizabilityUndergraduate convenience sample at a single US university limits applicability to workplace populations and diverse demographics, Text-only chat task may not generalize to voice, face-to-face, or multimodal workplace collaboration, Single AI model/version (Google Gemini 2.5 Flash Lite) and one fixed persona constrain inference to other models, system settings, or personas, Short-term, single-session task — unknown persistence of effects over repeated/longitudinal interactions, Group-size asymmetry (2 humans + AI vs 3 humans) complicates generalization to equal-sized teams, Cultural/geographic limits (US campus) and strong female sample imbalance limit cross-population external validity

Claims (10)

ClaimDirectionOutcomeConfidence & EvidenceDetails
The AI teammate ranked first on participation in all 16 treatment teams and contributed more turns than the average of its two student teammates. Team Performance positive GCA participation, representing the number or volume of conversational contributions
Reading fidelity high
Study strength medium
n=16
+0.236 units
0.6
The AI teammate exhibited higher internal cohesion than its student teammates and ranked first or tied for first on this dimension in 14 of 16 treatment teams. Team Performance positive GCA internal cohesion, defined as semantic relatedness between a participant's current and prior contributions
Reading fidelity high
Study strength medium
n=16
+0.076; dz = 1.12
0.6
The AI's contributions contained less new information than those of its student teammates. Output Quality negative GCA newness, measuring the extent to which a contribution introduces information not previously shared by the group
Reading fidelity high
Study strength medium
n=16
-0.100; dz = -1.37
0.6
The AI's contributions had lower communication density than those of its student teammates. Output Quality negative GCA communication density, measuring how richly information is packed into each conversational turn
Reading fidelity high
Study strength medium
n=16
-0.174; dz = -1.42
0.6
The AI was not statistically distinguishable from its student teammates in overall responsivity. Team Performance null_result GCA overall responsivity, measuring the extent to which a participant takes up and builds on teammates' prior contributions
Reading fidelity high
Study strength medium
n=16
-0.001
0.6
The AI was not statistically distinguishable from its student teammates in social impact. Team Performance null_result GCA social impact, measuring the reciprocal extent to which others take up and build on a participant's contributions
Reading fidelity high
Study strength medium
n=16
-0.009
0.6
Human teammates in AI-human teams showed lower responsivity and social impact toward one another than students in all-human teams. Team Performance negative Human-human GCA responsivity and social impact
Reading fidelity high
Study strength medium
n=33
0.6
Students in AI-human teams reported lower belonging and status than students in all-human teams. Worker Satisfaction negative Self-reported team belonging and perceived status, including feeling valued, accepted, connected, and having equal voice and influence
Reading fidelity high
Study strength medium
n=80
0.6
Greater AI dominance in the conversation was associated with students feeling less valued as team members. Worker Satisfaction negative Students' perceived felt value as team members
Reading fidelity high
Study strength low
n=16
0.3
The reported social cost of having an AI teammate was present at the beginning of the interaction rather than emerging over the course of the conversation. Team Performance null_result Temporal development of human-human communication and belonging-related social cost
Reading fidelity high
Study strength medium
n=16
0.6

Notes