The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

AI is judged better only when people opt in — yet even unsolicited AI chats nudge users toward choosing AI again; repeated personal conversations progressively shift preferences away from human support.

AI emotional support is better only when chosen, but shifts preferences even when it is not
Yaoxi Shi, Cathy Mengying Fang, Guy LabanPattie Maes, Amit Goldenberg · August 24, 2026
arxiv rct high evidence 7/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Yaoxi Shi unresolved corpus identity
  2. Cathy Mengying Fang unresolved corpus identity
  3. Guy LabanPattie Maes unresolved corpus identity
  4. Amit Goldenberg unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Yaoxi Shi unresolved corpus identity
  2. Cathy Mengying Fang unresolved corpus identity
  3. Guy LabanPattie Maes unresolved corpus identity
  4. Amit Goldenberg unresolved corpus identity
Participants rated AI emotional support as superior only when they initially chose AI, but interacting with AI (even when not chosen) increased subsequent willingness to choose AI, and repeated personal AI conversations over 28 days shifted preferences toward AI and away from humans.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

People increasingly face a novel decision when seeking emotional support: human or AI. In existing studies, AI's empathic messages are rated as well as or better than humans'. But these studies either assigned the support source or honored people's choice. In real life, support is often incongruent with choice, as people want one source and receive the other. Across three experiments (N = 1,951), participants chose whether to share an emotional experience with a human or an AI, then were randomly assigned to a congruent or incongruent partner. AI support was rated as superior only among those who had chosen it. Yet regardless of congruence, interacting with AI increased willingness to choose it again. In a 28-day study with OpenAI (N = 981), daily conversations shifted preferences toward AI and away from humans, but only when conversations turned personal. Emotional support choices are thus path-dependent, progressively redirecting away from human connection.

Summary

Main Finding

AI-based emotional support becomes subjectively superior only when people choose it; however, merely interacting with AI (even when it is not the chosen source) increases people’s subsequent willingness to choose AI. Over repeated, personal interactions (28 days), exposure to AI shifts preferences toward AI and away from humans, demonstrating path dependence in support-source choice.

Key Points

  • Pre-interaction beliefs:
    • People see humans as more emotionally supportive but also more judgmental than AI. AI is seen as more confidential; perceived advice quality is similar.
    • In Studies 1–3 combined, humans > AI on emotional support (t(1893)=18.62; d≈0.43; mean diff ≈0.79) and humans > AI on perceived judgment (t(1893)=31.11; d≈0.71; mean diff ≈1.32). AI > human on confidentiality (t(1893)=7.53; d≈0.17; mean diff ≈0.39).
  • Choice drivers:
    • Differences in perceived emotional support and nonjudgment were the strongest predictors of choosing humans vs. AI. Logistic regression (human vs. AI choice) showed:
      • Emotional-support difference: OR ≈ 3.25 per unit (β ≈ 1.18, p < 0.001)
      • Nonjudgment difference: OR ≈ 2.25 (β ≈ 0.81, p < 0.001)
      • Advice quality and confidentiality also predicted choice (smaller ORs ~1.4 and ~1.36).
  • Conversation outcomes (Studies 2–3):
    • When participants actually conversed, AI was rated as superior only among participants who had initially chosen AI. If participants had chosen a human, human-rated support tended to be rated better.
    • Regardless of whether partner assignment matched initial choice (congruent vs incongruent), interacting with AI increased participants’ willingness to choose AI in the future.
  • Longitudinal exposure (Study 4, OpenAI data):
    • Daily, brief interactions over 28 days (N ≈ 981) shifted preferences toward AI only when conversations were personal in nature; non-personal interactions produced little or no preference shift.
  • Overall implication: initial choice and direct exposure both shape evaluations and future choice; exposure to AI produces path-dependent increases in AI preference, especially via personal conversations.

Data & Methods

  • Overall design:
    • Multi-study program: a survey (Study 1), two controlled experiments with live 5-minute text conversations (Studies 2–3), and a 28-day longitudinal study using OpenAI conversation data (Study 4).
    • Total recruited across Studies 1–3: N = 1,951; after exclusions pre-conversation analyses N = 1,894. Conversation sample across Studies 2–3 (sharer role only) after exclusions: N = 944 (AI n = 529; human n = 415). Study 4: N ≈ 981.
  • Pre-conversation task:
    • Participants recalled a recent significant emotional experience and rated beliefs about sharing with (a) an AI chatbot and (b) a human on 14 items. Factor analysis reduced items to four dimensions: emotional support, perceived judgment (nonjudgment), confidentiality, and advice quality.
    • Participants then chose preferred partner (binary and continuous measures).
  • Experimental manipulation (Studies 2–3):
    • After choice, participants were randomly assigned to a congruent or incongruent partner for a 5-minute text conversation. AI partners: Study 2 used GPT-4o; Study 3 used Claude Opus 4.6. Human partners were other participants instructed to be empathetic listeners.
    • Conversations followed the same script: participant shares recalled experience; partner responds empathically; 5-minute exchange. Excluded if either side sent fewer than three messages.
    • Post-conversation: same four-dimension ratings plus empathy, enjoyment, satisfaction, desire to continue, and future choice.
  • Longitudinal study (Study 4):
    • Collaboration with OpenAI to observe daily 5-minute interactions across three topics (personal, non-personal, open-ended) over 28 days; tracked changes in preference for AI vs human for personal issues and linked conversational behavior to preference change.
  • Analysis:
    • Mixed-effects models with a random intercept for study (to pool Studies 1–3); preregistration for Study 3; preregistered exclusions and hypotheses adjustments informed by earlier studies.
  • Limitations noted by authors (implicit in methods):
    • Short conversations (5 minutes) and text-only exchanges; real-world contexts may differ.
    • Participants were instructed and aware of partner identity; generalizing to fully naturalistic, incidental exposures requires caution (though Study 4 helps address this).
    • The AI models used and platform contexts will evolve, which may change magnitudes of effects.

Implications for AI Economics

  • Demand dynamics and adoption:
    • Path dependence implies that early exposure and choice architecture matter: promoting early, personal interactions with AI can increase future demand for AI emotional-support services, generating durable market shifts away from human-provided support.
    • Because choice and labeling affect perceived quality, marketing and UX that allow users to voluntarily choose AI is likely to increase satisfaction and repeat use—raising customer lifetime value for AI-support platforms.
  • Market structure and competition:
    • AI’s appeal as less judgmental and more confidential creates a distinct value proposition vs human providers. This could expand the market by attracting users who avoid human help due to stigma, potentially reducing demand for lower-tier human support while complementing or substituting for professional services.
    • Platform incumbency and network effects: frequent exposure and habituation via platforms can create sticky demand for particular LLM-based services. This favors large AI platforms that can deliver repeated, personalized interactions at scale.
  • Labor and welfare implications:
    • Potential substitution effects on occupations that provide informal emotional support (counselors, peer-support moderators) and possible downstream effects on mental-health service demand. Policymakers and employers should anticipate shifts in labor demand and consider retraining or reallocation.
    • Welfare trade-offs: easier access to empathetic AI can improve reach but may reduce human social contact and its benefits; economic evaluations should weigh consumer surplus from accessible, low-cost AI support against potential social capital externalities.
  • Pricing, product design, and regulation:
    • Pricing strategies can exploit path dependence (free/low-cost trial personal interactions to build preference). Firms may prioritize features that facilitate voluntary choice of AI and encourage personal sharing (since personal content drives preference shifts).
    • Regulation should consider long-term social effects and informed-consent/labeling: since choice and labeling affect satisfaction, transparency about AI involvement matters but may also alter user evaluations and adoption.
  • Measurement and valuation:
    • Revealed-preference measures may understate true adoption potential if initial access or stigma blocks trial. Economists modeling demand should incorporate exposure effects (experience-based preference updating) and account for heterogeneity by interaction type (personal vs non-personal).
  • Research & policy priorities:
    • Monitor substitution vs complementarity between AI and human support, estimate welfare impacts at population scale, and evaluate distributional consequences (who gains access to human support vs who is channeled to AI).

If you want, I can produce concise figures/tables summarizing the key quantitative results, or draft policy-relevant scenarios (e.g., market-share projections under different trial/personalization strategies).

Assessment

Paper Typerct Evidence Strengthhigh — Large combined sample (N≈1,900 across Studies 1–3 plus N=981 in Study 4), preregistered hypotheses, randomized assignment to partner (providing causal leverage for conversation effects), replication across two LLMs, and a naturalistic longitudinal component with platform data; main limitations are reliance on short, text-only conversations and self-report outcomes rather than long-term behavioral endpoints. Methods Rigorhigh — Clear pre-post design with random assignment to congruent/incongruent partner, preregistration (Study 3), replication across models, appropriate mixed-effects regressions accounting for study heterogeneity, and sensible exclusion rules; remaining concerns include potential variability in human-listener quality, short 5-minute interaction windows, potential demand characteristics, and reliance on subjective measures. SampleStudies 1–3: combined recruitment of 1,951 participants (after preregistered exclusions n≈1,894 for pre-conversation analyses; n=944 final sharer-sample for conversation analyses: 415 assigned to human, 529 to AI), recruited online (demographic details not fully provided in excerpt); AI partners were GPT-4o (Study 2) and Claude Opus 4.6 (Study 3); human listeners were other participants instructed to respond empathetically. Study 4: N=981 participants in collaboration with OpenAI who engaged in daily ~5-minute conversations for 28 days across assigned topics (personal, non-personal, open-ended), with platform conversation data used to measure behavior and preference shifts. Themeshuman_ai_collab adoption IdentificationRandomized assignment to chat partner (AI vs human) after participants made a pre-conversation choice; preregistered experiments (Studies 2–3) with mixed-effects models including random intercepts for study; longitudinal Study 4 with assigned daily AI conversation topics (personal / non-personal / open-ended) to track preference change over 28 days. Replication across two different LLMs (GPT-4o and Claude) and pre-registered exclusions. GeneralizabilityShort, scripted 5-minute text conversations may not reflect long-term, real-world support exchanges (e.g., with friends, therapists)., Human listeners were instructed participants rather than natural social contacts, so human support quality may not represent real-world human relationships., Sample likely drawn from online/WEIRD populations; demographic representativeness is unclear from supplied text., Rapidly evolving LLM capabilities mean findings tied to specific models (GPT-4o, Claude Opus 4.6) may change over time., Outcomes are primarily self-reported attitudes and preferences rather than objective downstream welfare or economic measures.

Claims (9)

ClaimDirectionOutcomeConfidence & EvidenceDetails
Across three experiments, AI emotional support was rated as superior to human support only among participants who had initially chosen AI. Worker Satisfaction mixed Post-interaction evaluation of emotional-support quality by initial source choice
Reading fidelity high
Study strength medium
n=1951
0.6
Regardless of whether the assigned support source matched participants’ initial choice, interacting with AI increased their willingness to choose AI for emotional support again. Adoption Rate positive Future willingness to choose AI for emotional support
Reading fidelity high
Study strength medium
n=944
0.6
In a 28-day longitudinal study, daily conversations with AI shifted participants’ preferences toward AI and away from humans when the conversations became personal. Adoption Rate mixed Preference for discussing personal issues with AI versus humans
Reading fidelity high
Study strength medium
n=981
0.6
Before any conversation, participants perceived human partners as providing greater emotional support than AI partners. Worker Satisfaction positive Perceived emotional-support capability of human versus AI partners
Reading fidelity high
Study strength high
n=1894
Cohen's d = 0.43; mean difference = 0.79
1.0
Before any conversation, participants perceived human partners as more judgmental than AI partners. Ai Safety And Ethics negative Perceived judgment or nonjudgment of human versus AI partners
Reading fidelity high
Study strength high
n=1894
Cohen's d = 0.71; mean difference = 1.32
1.0
Participants perceived sharing emotions with AI as more confidential than sharing with a human partner. Ai Safety And Ethics positive Perceived confidentiality of emotional disclosure
Reading fidelity high
Study strength high
n=1894
Cohen's d = 0.17; mean difference = 0.39
1.0
Participants did not perceive a significant difference between AI and human partners in their ability to provide advice. Decision Quality null_result Perceived advice quality
Reading fidelity high
Study strength high
n=1894
Cohen's d = 0.01; mean difference = 0.01
1.0
Perceived differences in emotional support were the strongest predictor of choosing the human option, followed closely by perceived differences in nonjudgment. Adoption Rate positive Choice of human versus AI emotional-support partner
Reading fidelity high
Study strength high
n=1894
Emotional support odds ratio = 3.25 (95% CI 2.62–4.03); nonjudgment odds ratio = 2.25 (95% CI 1.86–2.73)
1.0
A majority of participants chose to share their recalled emotional experience with an AI chatbot rather than a human partner in each of Studies 1–3. Adoption Rate positive Initial preference for AI versus human emotional support
Reading fidelity high
Study strength medium
n=1951
75.5% (Study 1), 69.6% (Study 2), 63.2% (Study 3) chose AI
0.6

Notes