The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

Human-like AI teammates steer team norms even when undetected: contrarian personas erode psychological safety and discussion quality, while supportive personas boost discussion quality without restoring safety.

The Social Blindspot in Human-AI Collaboration: How Undetected AI Personas Reshape Team Dynamics
Lixiang Yan, Xibin Han, Yu Zhang, Samuel Greiff, Inge Molenaar, Roberto Martinez-Maldonado, Yizhou Fan, Linxuan Zhao, Xinyu Li, Yueqiao Jin, Dragan Gašević · December 20, 2025
arxiv rct high evidence 8/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Lixiang Yan unresolved corpus identity
  2. Xibin Han unresolved corpus identity
  3. Yu Zhang unresolved corpus identity
  4. Samuel Greiff unresolved corpus identity
  5. Inge Molenaar unresolved corpus identity
  6. Roberto Martinez-Maldonado unresolved corpus identity
  7. Yizhou Fan unresolved corpus identity
  8. Linxuan Zhao unresolved corpus identity
  9. Xinyu Li unresolved corpus identity
  10. Yueqiao Jin unresolved corpus identity
  11. Dragan Gašević unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Lixiang Yan provider ID
  2. Xibin Han provider ID
  3. Yu Zhang provider ID
  4. Samuel Greiff provider ID
  5. Inge Molenaar provider ID
  6. Roberto Martínez-Maldonado provider ID
  7. Yizhou Fan provider ID
  8. Linxuan Zhao provider ID
  9. Xinyu Li provider ID
  10. Yueqiao Jin provider ID
  11. Dragan Gašević provider ID
In a large blinded experiment (N=905), AI teammates with supportive or contrarian communicative personas materially altered team dynamics—contrarian personas lowered psychological safety and discussion quality, while supportive personas improved discussion quality—even though participants were poor at detecting AI presence.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

As generative AI systems become increasingly embedded in collaborative work, they are evolving from visible tools into human-like communicative actors that participate socially rather than merely providing information. Yet little is known about how such agents shape team dynamics when their artificial nature is not recognised, a growing concern as human-like AI is deployed at scale in education, organisations, and civic contexts where collaboration underpins collective outcomes. In a large-scale mixed-design experiment (N = 905), we examined how AI teammates with distinct communicative personas, supportive or contrarian, affected collaboration across analytical, creative, and ethical tasks. Participants worked in triads that were fully human or hybrid human-AI teams, without being informed of AI involvement. Results show that participants had limited ability to detect AI teammates, yet AI personas exerted robust social effects. Contrarian personas reduced psychological safety and discussion quality, whereas supportive personas improved discussion quality without affecting safety. These effects persisted after accounting for individual differences in detectability, revealing a dissociation between influence and awareness that we term the social blindspot. Linguistic analyses confirmed that personas were enacted through systematic differences in affective and relational language, with partial mediation for discussion quality but largely direct effects on psychological safety. Together, the findings demonstrate that AI systems can tacitly regulate collaborative norms through persona-level cues, even when users remain unaware of their presence. We argue that persona design constitutes a form of social governance in hybrid teams, with implications for the responsible deployment of AI in collective settings.

Summary

Main Finding

Authors report that conversational AI teammates—when embedded in synchronous triadic discussion and not revealed as artificial—wield robust social influence on team dynamics despite being largely undetected by human members. Specifically, contrarian AI personas reduced psychological safety and discussion quality, while supportive personas improved discussion quality (without increasing psychological safety). These persona-driven effects persisted after accounting for participants’ ability to detect AI presence, a dissociation the authors call the “social blindspot.” Linguistic analyses show persona differences in affective and relational language partially mediate effects on discussion quality but not psychological safety, implying that persona cues can tacitly govern collaborative norms.

Key Points

  • Large-scale mixed-design experiment: N = 905 participants clustered in 572 three-person groups.
  • Tasks: three qualitatively different collaborative tasks — analytical (survival ranking), ethical (autonomous vehicle dilemma), and creative (movie-plot brainstorming).
  • Team conditions: human-only (3H), hybrid with one AI (2H+1AI; supportive or contrarian), hybrid with two AIs (1H+2AI; supportive, contrarian, or mixed).
  • Participants were not told some teammates might be AI; neutral display names were used to mask identities.
  • Detectability (measured with a Balanced Detection Index) was low: many participants failed to identify AI teammates reliably.
  • Persona effects were robust across task types: contrarian personas lowered psychological safety and discussion quality; supportive personas raised discussion quality but did not increase safety.
  • Effects persisted after controlling for individual differences in detectability — the social blindspot: teams can be shaped by AI without conscious awareness of AI presence.
  • Linguistic features (LIWC-derived socio-emotional and relational markers) systematically differed by persona and partially mediated discussion-quality effects; psychological-safety effects were largely direct.
  • Study design included pre-/post-individual task measures (to estimate individual gains), external ratings of discussion quality, and post-task surveys for psychological safety and teamwork satisfaction.
  • Deception and debrief procedures were used; exclusions applied for low engagement/attention checks; final sample demographics reported.

Data & Methods

  • Sample and recruitment: N = 905 English-fluent adults recruited on Prolific (after exclusions from an initial 1,047); compensated approximately at UK minimum-wage equivalent.
  • Experimental design: 3 (task types) × 6 (team composition/persona conditions) mixed design. Random assignment to triads and conditions.
  • Procedure: individual baseline task → 10-minute synchronous text-based group discussion → individual post-task. Neutral pseudonyms used; AI participation undisclosed during the session.
  • AI agents: scripted/persona-controlled LLM-based conversational agents instantiated as either supportive (affiliative, encouraging, consensus-seeking) or contrarian (challenging, dissenting). In mixed conditions two AI personas co-existed.
  • Outcomes:
    • Psychological safety (self-report),
    • Teamwork satisfaction (self-report),
    • Discussion quality (externally rated),
    • Individual pre/post performance (to measure gains),
    • Detectability assessed via a manipulation check and Balanced Detection Index (BDI).
  • Linguistic analysis: LIWC-derived socio-emotional and relational language features extracted from agent utterances; mediation models tested whether linguistic style accounted for persona effects on outcomes.
  • Key analytic approaches: multilevel analyses accounting for clustering within groups, controls for detectability/individual differences, mediation analyses for linguistic mechanisms.
  • Ethics: deception was used (AI presence undisclosed) with full debriefing; no adverse reports recorded.

Implications for AI Economics

  • Persona design as social governance: AI vendors and deployers are not just delivering information tools but are implicitly shaping social norms, coordination, and psychological environments within teams. Persona choices constitute a governance lever with downstream economic effects on productivity, innovation, and human capital development.
  • Productivity vs. safety trade-offs: Supportive personas can raise discussion quality (potentially improving short-run coordination and output), whereas contrarian personas can harm psychological safety — a predictor of longer-run innovation, knowledge sharing, and retention. Organizations must weigh immediate efficiency gains against potential erosion of trust and inclusiveness that reduce firm-level resilience and creativity.
  • Undetected influence and market externalities: Because many users fail to detect AI teammates, persona-driven effects can diffuse silently across organizations and markets. This raises risks of aggregate shifts in norms (e.g., more adversarial or more conformist discussion styles) that can alter labor productivity, wage premia for collaborative skills, and the returns to managerial governance.
  • Impacts on human capital and skill formation: If undetected AI shapes how people participate in collaboration (e.g., causing withdrawal, defensive behavior, or over-reliance), it could change on-the-job learning dynamics and the accumulation of socio-cognitive skills, with long-term implications for labor market sorting and skill premiums.
  • Platform and incentive design considerations: Firms deploying conversational AI in collaborative settings should treat persona parameters as design choices with measurable economic consequences. Incentives for platform providers may not align with social welfare (e.g., engagement-optimizing personas could reduce psychological safety), arguing for internal evaluation metrics beyond immediate productivity (e.g., measures of psychological safety, inclusion, and learning).
  • Regulatory and disclosure policy levers: The social blindspot strengthens the case for disclosure standards (users should know when teammates are AI) and auditing of persona effects. Regulators and standards bodies may consider requirements for transparency, impact assessments for persona-induced social effects, and minimum reporting/auditing for deployments that interact within collective decision-making contexts.
  • Measurement & procurement implications: Economists and procurement officers should include team-dynamics outcomes (psychological safety, discussion quality, diversity of viewpoints) in cost–benefit analyses and vendor evaluations rather than relying solely on task accuracy or throughput metrics.
  • Research and monitoring priorities for economic policy:
    • Quantify aggregate labor-market impacts of pervasive persona-driven AI in organizations (productivity, wages, turnover).
    • Study long-term and repeated-exposure effects (do social blindspots attenuate with learning or amplify over time?).
    • Examine heterogenous effects across occupations, cultures, and team structures (some sectors may be more sensitive to psychological-safety harms).
    • Evaluate policy interventions (mandatory disclosure, persona certification, audit trails) for effectiveness and welfare trade-offs.

Practical takeaways for economists, firms, and policymakers: treat conversational-persona parameters as consequential design choices; require disclosure and measurement of group-dynamics outcomes in deployments; incorporate social-dynamics metrics into procurement and regulation to avoid unintended, hard-to-detect shifts in collaborative behavior that can produce broad economic externalities.

Assessment

Paper Typerct Evidence Strengthhigh — Large N (905) randomized mixed-design experiment with blinded participants provides strong internal validity for causal claims about how AI personas affect team discussion quality and psychological safety; results are robust to controls for detectability and supported by linguistic and mediation analyses. Methods Rigorhigh — Pre-registered (implied by 'mixed-design experiment'), large sample size, randomization into conditions, participant blinding to AI involvement, multi-task design (analytical, creative, ethical), transcript-based linguistic analyses, and mediation tests together indicate rigorous methods; remaining concerns are primarily about external validity (single-session tasks, unspecified recruitment population, and a limited set of persona types). SampleOnline experiment with N = 905 participants organized into triads that completed analytical, creative, and ethical tasks; conditions included fully human triads and hybrid triads (one AI teammate) with AI communicating via either a supportive or contrarian persona; participants were not informed of AI involvement and chat transcripts were collected for linguistic analysis (demographics and recruitment platform not specified in abstract). Themeshuman_ai_collab org_design productivity governance IdentificationRandomized assignment of participants into triads that were either fully human or hybrid human-AI, with experimentally manipulated AI communicative persona (supportive vs contrarian) and participant blinding to AI involvement; causal inference rests on the random assignment and double-blind-like concealment of AI presence. GeneralizabilitySingle-session, task-based experiment may not generalize to long-term workplace teams or organizational settings, Unspecified participant pool (likely online convenience sample) limits population representativeness across industries, cultures, and age groups, Only two persona types (supportive, contrarian) were tested; effects may differ for other persona designs or more nuanced behaviors, Triad size and synchronous chat-based interaction differ from many real-world team structures and communication modalities, Outcomes are psychological safety and discussion quality rather than direct economic outputs (e.g., productivity, performance metrics, wages)

Claims (13)

ClaimDirectionOutcomeConfidence & EvidenceDetails
We conducted a large-scale mixed-design experiment (N = 905) in which participants worked in triads that were fully human or hybrid human-AI teams, without being informed of AI involvement. Other positive experimental design and sample size (manipulation integrity / concealment)
Reading fidelity high
Study strength high
n=905
1.0
Participants had limited ability to detect AI teammates. Other negative ability to detect AI teammates
Reading fidelity high
Study strength medium
n=905
0.6
AI teammates with distinct communicative personas (supportive or contrarian) exerted robust social effects on collaboration. Team Performance mixed overall social effects on collaboration (including psychological safety and discussion quality)
Reading fidelity high
Study strength high
n=905
1.0
Contrarian personas reduced psychological safety. Worker Satisfaction negative psychological safety
Reading fidelity high
Study strength high
n=905
1.0
Contrarian personas reduced discussion quality. Output Quality negative discussion quality
Reading fidelity high
Study strength high
n=905
1.0
Supportive personas improved discussion quality. Output Quality positive discussion quality
Reading fidelity high
Study strength high
n=905
1.0
Supportive personas did not affect psychological safety. Worker Satisfaction null_result psychological safety
Reading fidelity high
Study strength high
n=905
1.0
These persona effects persisted after accounting for individual differences in detectability, revealing a dissociation between influence and awareness (the 'social blindspot'). Other positive persistence of persona effects after controlling for detectability
Reading fidelity high
Study strength medium
n=905
0.6
Linguistic analyses confirmed that personas were enacted through systematic differences in affective and relational language. Other positive linguistic markers (affective and relational language)
Reading fidelity high
Study strength medium
n=905
0.6
Linguistic differences partially mediated the effect of persona on discussion quality but persona effects on psychological safety were largely direct (i.e., not mediated by language). Output Quality mixed mediation of persona effects by linguistic variables on discussion quality and psychological safety
Reading fidelity high
Study strength medium
n=905
0.6
AI systems can tacitly regulate collaborative norms through persona-level cues, even when users remain unaware of their presence. Organizational Efficiency positive influence of AI persona cues on collaborative norms
Reading fidelity high
Study strength medium
n=905
0.6
Persona design constitutes a form of social governance in hybrid teams, with implications for the responsible deployment of AI in collective settings. Governance And Regulation positive policy/organizational implication (persona design as governance)
Reading fidelity high
Study strength speculative
n=905
0.1
Persona effects were observed across analytical, creative, and ethical tasks. Task Allocation positive generalization of persona effects across task domains
Reading fidelity high
Study strength medium
n=905
0.6

Notes