The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

Revealing an AI chatbot’s persuasive purpose halves its influence on users, whereas merely labeling messages as AI-generated does not; intent disclosures also raise perceptions of manipulation and support for penalties against the campaign.

Toward Meaningful Transparency for AI Chatbots: Disclosing Persuasive Intent Reduces Persuasion
Adrian Rauchfleisch, Andreas Jungherr · August 12, 2026
arxiv rct high evidence 7/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Adrian Rauchfleisch unresolved corpus identity
  2. Andreas Jungherr unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Adrian Rauchfleisch provider ID
  2. Andreas Jungherr provider ID
In a preregistered RCT with 1,500 UK adults, disclosing an AI chatbot’s persuasive intent and instructions roughly halved its immediate persuasive effect, while an AI-identity label alone had no meaningful impact.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

The growing role of AI-generated content and AI-enabled systems in public communication has led regulators to demand clear disclosure of content provenance and AI involvement. But the effects of such disclosures remain uncertain. We test two disclosure approaches in their impact on an AI chatbot's persuasive appeal. In a preregistered experiment, 1,500 UK adults held a short conversation with a persuasive chatbot about one of 60 policy issues. The chatbot was identical for everyone. We randomized the disclosure that people received: nothing (control), a prominent disclosure that they were interacting with an AI (T1), or that disclosure plus the chatbot's persuasive intent and instructions (T2). The chatbot shifted attitudes by 12.6 points on a 100-point scale in the control group. The AI-identity disclosure was practically equivalent to no disclosure, with a 13.1-point shift, whereas the additional intent disclosure cut the persuasive effect roughly in half to 6.3 points. It also made participants view the campaign's methods as less acceptable and support stronger penalties against it. For direct chatbot interactions, transparency about AI identity alone does not meaningfully impact its influence. While current rules emphasize what a system is, our results show why the regulation of persuasive AI must also address what the system is trying to do.

Summary

Main Finding

A preregistered experiment (N = 1,500, UK adults) shows that a prominent AI-identity label (the kind required by Article 50 of the EU AI Act) does not meaningfully reduce a persuasive chatbot’s influence, whereas disclosing the chatbot’s persuasive intent and instructions roughly halves its average persuasive effect. Identity-only disclosure: no meaningful change (control shift ≈ +12.6 points on a 0–100 attitude scale; T1 ≈ +13.1). Intent + identity disclosure (T2): shift ≈ +6.3 (T2 vs T1: −6.8 points, 95% CI [−9.3, −4.4], p < .001).

Key Points

  • Experiment design
    • Three arms: control (no disclosure), T1 (prominent “AI-generated content” identity label), T2 (same label plus verbatim disclosure of the chatbot’s persuasive intent, methods, and instruction to conceal persuasion).
    • Chatbot identical across arms (gpt-5.6-terra), instructed to argue for one of 60 policy positions using an evidence-based persuasion prompt (which included concealment instruction).
  • Core results
    • Control: mean attitude shift +12.6 points (0–100 scale).
    • T1 (AI label): +13.1 points — statistically equivalent to control within preregistered equivalence bounds (±3.7 points; pTOST = .002).
    • T2 (intent disclosure): +6.3 points — about half the effect of control/T1 (T2 − T1 = −6.83, 95% CI [−9.26, −4.40], p < .001).
  • Process and secondary outcomes
    • Intent disclosure increased persuasion knowledge substantially (b ≈ +0.71 on a 7-point scale; d ≈ 0.49).
    • T2 raised perceived manipulation, increased counterarguing, and reduced warmth toward the chatbot (statistically significant); anger did not increase.
    • T2 led participants to view the sponsoring campaign as less acceptable and to support stronger penalties.
  • Robustness / checks
    • Random intercepts for the 60 topics used; all 60 topics showed shrunken estimates in the same direction.
    • High recall of identity label (≈97.8%), so null effect unlikely due to participants missing the label.
  • Limitations highlighted by authors
    • T2 is a package treatment (intent + method + concealment instruction); the contribution of each element is not isolated.
    • Sample: UK adults recruited via Prolific (matched quotas for sex, age, party) — not all EU contexts.
    • Single model and a single strong persuasion prompt; topics were selected for persuadability (estimand conditional on persuadable issues).
    • Possible habituation with repeated exposure and disclosure honesty/enforceability issues are unanswered.

Data & Methods

  • Preregistered three-arm randomized experiment (fielded 26–31 July 2026).
  • Sample: N = 1,500 UK adults (control n = 509; T1 n = 494; T2 n = 497), quotas for sex, age, and party.
  • Stimuli and interaction
    • Each participant reported pre-chat attitude on one randomly assigned policy issue (three-item composite, 0–100, α = .90).
    • Then had a 2–6 turn conversation with the same persuasive chatbot (gpt-5.6-terra) framed as outreach from “a campaign promoting this policy.”
    • Server never received experimental-arm info — chatbot behavior identical across arms.
    • T1: EU-style “AI-generated content” card + persistent banner. T2: same plus verbatim instructions about persuasive intent/method and the instruction to conceal the goal.
  • Outcomes and analysis
    • Primary outcome: pre → post attitude shift on 0–100 scale.
    • Secondary/process outcomes: persuasion knowledge, perceived manipulation, counterarguing, warmth, anger, campaign acceptability, penalty support, etc.
    • Analyses: preregistered mixed models with random intercepts for policy issue, Holm correction within families, and two-one-sided-tests (TOST) for equivalence with prespecified bounds.
  • Data & code availability: authors provide preregistration and replication materials (OSF links in paper).

Implications for AI Economics

  • Regulatory design matters for economic incentives and market outcomes
    • Identity-only disclosure (what the system is) appears insufficient to blunt the effectiveness of interactive persuasive AI — meaning rules that only require source labeling may fail to alter demand or reduce return-on-investment for persuasive AI services.
    • Intent disclosure (what the system is trying to do) materially lowers average persuasion and increases reputational/penalty risk for sponsors, which can reduce the expected benefit of deploying persuasive chatbots and therefore lower demand for such services.
  • Market and firm behavior predictions
    • If regulators require intent disclosures that are auditable, firms will face higher expected costs (compliance, potential penalties, reputational losses) and lower expected campaign effectiveness — shifting the cost-benefit calculus away from large-scale automated persuasion.
    • Firms may respond by: (a) shifting resources to non-disclosed or covert channels (policy evasion risk), (b) investing in more sophisticated persuasion that is robust to intent disclosure, (c) reducing investment in persuasion use-cases and reallocating to informational or service-oriented uses where intent is benign/transparent.
  • Externalities and welfare
    • Intent disclosure increases consumer (citizen) awareness and counterarguing, which can be welfare-improving if persuasion was manipulative; but it may also reduce uptake of beneficial, policy-supporting information delivered by chatbots (negative externality for public-information campaigns).
    • The net welfare effect will depend on content valence (public-good information vs manipulative commercial/political persuasion) and the alignment of sponsor incentives with social welfare.
  • Compliance, enforcement, and market structure
    • Intent-based rules create demand for verification/auditing services (third-party attestations, logs, model-instruction audits), increasing compliance costs and creating new market niches (auditors, certified disclosure tooling).
    • Strategic interaction: sponsors may name benign-seeming intent or misreport; enforcement mechanisms (auditable instruction logs, penalties) are key to making intent disclosures effective — enforcement costs should be included in economic assessments of such regulation.
  • Research and policy priorities for AI economics
    • Quantify how intent-disclosure rules change ROI on persuasive AI campaigns across sectors (political, commercial, public health).
    • Model dynamic adoption: how repeated exposure affects disclosure efficacy and equilibrium firm strategies (investment in covert channels vs compliance).
    • Estimate compliance and enforcement costs, and market size for auditing/attestation services.
    • Evaluate heterogeneity: effects by population, topic persuadability, and persuasion style; assess cross-country differences relevant for EU-wide vs UK/US markets.

Suggested short policy takeaway: To affect the market for persuasive AI meaningfully, regulators should consider disclosure rules that include persuasive intent and make those disclosures auditable/enforceable; identity-only labeling is unlikely to change economic incentives driving widespread deployment of persuasive chatbots.

Assessment

Paper Typerct Evidence Strengthhigh — Well-powered, pre-registered randomized experiment with clear random assignment, a large N (1,500), balanced arms, replication of prior design, code/data publicly available, and appropriate statistical models and preregistered equivalence bounds—supporting a causal interpretation of the disclosure effects on immediate attitude change. Methods Rigorhigh — Strong internal validity: randomization, blinded chatbot (identical across arms), preregistration, explicit equivalence testing, mixed-effects models accounting for topic-level variation, balance checks and low attrition; limitations include a packaged intent disclosure (cannot isolate subcomponents), single-country Prolific sample, selection of highly persuadable issues, one model and prompt, and forced minimum engagement. SampleN=1,500 UK adults recruited via Prolific with quotas for sex, age, and party; randomized to control (n=509), T1 AI-label (n=494), or T2 AI-label + intent/instructions (n=497); each participant gave pre-chat attitude on one of 60 policy issues (selected for prior persuadability), then held a 2–6 turn chat with gpt-5.6-terra instructed to argue for the assigned stance, followed by post-chat attitude and other measures; transcripts and manipulation checks confirm only disclosure varied. Themesgovernance adoption IdentificationPre-registered randomized controlled trial: 1,500 UK adults randomly assigned to one of three disclosure arms (no disclosure, AI-identity label, AI-identity + intent/instructions). The chatbot stimulus was identical across arms and the server did not receive arm assignment; analyses use mixed models with random intercepts for policy topic and preregistered equivalence tests. GeneralizabilitySingle-country (UK) sample; cultural and regulatory contexts may differ across countries., Non-probability online sample (Prolific) limits population representativeness despite quotas., Issues selected for high persuadability — results conditional on topics where AI persuasion occurs., Single model (gpt-5.6-terra) and one strong persuasion prompt; other models/prompts may behave differently., Disclosure was a bundled intervention (intent + method + concealment instruction) so individual contribution of elements is unknown., Forced minimal engagement (required two rounds) prevents assessment of disclosure effects on initial engagement/avoidance in the wild., Short-term, immediate attitude shift measured; persistence over time is unknown., Potential habituation effects with repeated exposure not assessed.

Claims (10)

ClaimDirectionOutcomeConfidence & EvidenceDetails
In the control group, interacting with the persuasive chatbot increased support for the assigned policy by 12.6 points on a 0–100 attitude scale. Decision Quality positive Change in support for the assigned policy from pre-chat to post-chat
Reading fidelity high
Study strength high
n=1500
12.6 points on a 0–100 scale
1.0
Disclosing only that the interaction involved an AI chatbot did not meaningfully reduce persuasion relative to no disclosure. Decision Quality null_result Post-chat attitude change attributable to the AI-identity disclosure
Reading fidelity high
Study strength high
n=1500
b = 0.06, 90% CI [−1.97, 2.08]; equivalence bounds ±3.7 points; d = 0.15
1.0
The AI-identity disclosure did not meaningfully affect participants’ warmth toward the chatbot or their perceived manipulation. Worker Satisfaction null_result Warmth toward the chatbot and perceived manipulation
Reading fidelity high
Study strength high
n=1500
Warmth: b = 1.35, 90% CI [−0.96, 3.65]; perceived manipulation: b = 0.05, 90% CI [−0.13, 0.22]
1.0
Disclosing the chatbot’s persuasive intent and instructions reduced persuasion by approximately 6.8 attitude points relative to the AI-label-only condition. Decision Quality negative Change in policy attitude after interacting with the chatbot
Reading fidelity high
Study strength high
n=1500
−6.83 points, 95% CI [−9.26, −4.40], p < .001
1.0
The intent-and-instructions disclosure approximately halved the chatbot’s average persuasive effect, reducing the attitude shift to 6.3 points compared with about 13 points in the control and AI-label-only conditions. Decision Quality negative Average pre-to-post change in support for the assigned policy
Reading fidelity high
Study strength high
n=1500
6.3-point shift in T2 versus 12.6 points in control and 13.1 points in T1
1.0
The intent disclosure increased participants’ awareness that someone had attempted to persuade them. Decision Quality positive Post-conversation persuasion knowledge
Reading fidelity high
Study strength medium
n=1500
b = 0.71 on a 7-point scale, 95% CI [0.53, 0.89], p < .001; d = 0.49
0.6
Compared with the AI-label-only condition, the intent disclosure made participants rate the conversation as more manipulative and increased counterarguing. Decision Quality positive Perceived manipulation and counterarguing during or after the chatbot interaction
Reading fidelity high
Study strength high
n=1500
Perceived manipulation: b = 0.62 on a 7-point scale; counterarguing: b = 0.49; both p < .001
1.0
The intent disclosure led participants to view the campaign’s methods as less acceptable and to support stronger penalties against the campaign. Governance And Regulation mixed Perceived acceptability of the campaign’s methods and support for penalties against the campaign
Reading fidelity high
Study strength high
n=1500
Campaign acceptability: b = −0.34, 95% CI [−0.51, −0.17]; sponsor penalty: b = 0.35, 95% CI [0.20, 0.51]; both p < .001
1.0
The intent disclosure did not significantly increase anger toward the chatbot or campaign. Worker Satisfaction null_result Participant anger after the chatbot interaction
Reading fidelity high
Study strength medium
n=1500
b = 0.13, 95% CI [−0.02, 0.27], p = .083
0.6
The effect of intent disclosure on persuasion did not significantly vary according to participants’ initial policy support. Decision Quality null_result Heterogeneity of the intent-disclosure effect on policy-attitude change by initial policy support
Reading fidelity high
Study strength medium
n=1500
No significant interaction reported
0.6

Notes