The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

Telling users an LLM is biased against their party cuts its ability to correct economic-policy misconceptions by 28%; warnings change interactions — users push back more and engage less receptively.

Perceived Political Bias in LLMs Reduces Persuasive Abilities
Matthew DiGiuseppe, Joshua Robison · February 20, 2026
arxiv rct high evidence 8/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Matthew DiGiuseppe unresolved corpus identity
  2. Joshua Robison unresolved corpus identity

Semantic Scholar

Latest observation:

  1. M. DiGiuseppe provider ID
  2. J. Robison provider ID
In a preregistered U.S. survey experiment (N=2,144), a short warning that ChatGPT was biased against respondents' political party reduced the model's ability to correct economic-policy misconceptions by 28%, with respondents responding more skeptically in conversation.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

Conversational AI has been proposed as a scalable way to correct public misconceptions and spread misinformation. Yet its effectiveness may depend on perceptions of its political neutrality. As LLMs enter partisan conflict, elites increasingly portray them as ideologically aligned. We test whether these credibility attacks reduce LLM-based persuasion. In a preregistered U.S. survey experiment (N=2144), participants completed a three-round conversation with ChatGPT about a personally held economic policy misconception. Compared to a neutral control, a short message indicating that the LLM was biased against the respondent's party attenuated persuasion by 28%. Transcript analysis indicates that the warnings alter the interaction: respondents push back more and engage less receptively. These findings suggest that the persuasive impact of conversational AI is politically contingent, constrained by perceptions of partisan alignment.

Summary

Main Finding

A short message implying that a conversational LLM (ChatGPT 4.1) is biased against a respondent’s party substantially reduces the model’s ability to correct political/economic misconceptions. In a preregistered U.S. survey experiment (N = 2,144), the strongest out-party bias cue attenuated persuasion by 28% (95% CI: 16%–40%) relative to a no-information control; a lighter out-party cue produced a 23% attenuation (95% CI: 9%–35%).

Key Points

  • Experiment structure
    • Four-arm, between-subjects, preregistered design: (1) no-information control, (2) non-directional bias info, (3) out-party bias — light, (4) out-party bias — heavy.
    • Respondents engaged in a three-round ChatGPT (non-reasoning 4.1) conversation about one of six common economic-policy misconceptions they personally held.
  • Persuasion effects
    • Control group average pre→post change in misconception agreement: −1.20 (0–4 scale).
    • Heavy out-party cue group change: −0.86 → 28% reduction in persuasion vs control.
    • Reversal rates (full switch from at-least-moderate misconception to no agreement or consensus): control 34.4% → heavy treatment 22.1%.
  • Mechanism and interaction dynamics
    • Transcript analysis shows respondents in out-party cue conditions wrote more and were more argumentative (not simply disengaged or dismissive), consistent with directional motivated reasoning.
    • Manipulation checks: both out-party treatments increased perception that ChatGPT favored the opposing party for respondents of both parties, eliminating baseline partisan differences in perceived bias.
  • Heterogeneity and robustness
    • Effects persist across topics and across partisanship; no statistically significant heterogeneity by party, affective polarization, prior trust in AI, or self-reported topic knowledge.
  • Additional outcomes
    • Out-party cues reduced perceived persuasiveness of the LLM, willingness to use AI again for opinion challenge, general trust in AI chatbots, and support for politicians using AI for policy information.
  • Limitations noted by authors
    • Single, short one-off interaction; single model family (non-reasoning ChatGPT 4.1); U.S. online sample. Results may differ with longer exposure, different models, or other populations.

Data & Methods

  • Sample and timing
    • Fielded Dec 2025–Jan 2026. Initial respondent pool N = 2,414; final experiment N = 2,144. 93% of initial pool held at least one targeted economic misconception.
  • Treatment
    • Control vs non-directional bias information vs two out-party bias messages (light and heavy). Heavy treatment included cues about founder political ties and an image implying out-group alignment.
  • Conversation
    • Three-round prompts to ChatGPT 4.1: model given the respondent’s answer and instructed to persuade toward economist consensus while remaining truthful.
  • Outcomes
    • Primary: change in agreement with a chosen misconception (0–4 scale) pre- and post-conversation; reversal to consensus.
    • Secondary: perceived persuasiveness, willingness to reuse AI, general trust in AI, support for politician use of AI.
  • Transcript scaling and mechanism tests
    • Open-ended chat transcripts were compared using an LLM-as-judge pairwise method (≈30 comparisons/transcript) and scaled with a Bayesian Bradley–Terry model to measure argumentative and dismissive behaviors.
  • Statistical analysis
    • OLS models of post-treatment agreement controlling for pretreatment agreement and topic fixed effects. Bootstrap (stratified) used for CIs on attenuation percentages. Multiple robustness and heterogeneity checks reported.

Implications for AI Economics

  • Credibility is an economic input: perceived political neutrality materially affects the marginal effectiveness (benefit) of LLM-based persuasion or corrective information. Perceived bias reduces the return on deploying conversational LLMs for public information campaigns.
  • Demand and adoption effects
    • Reduced willingness to reuse AI and lower trust imply lower sustained demand for LLM-based advisory/fact-checking services among groups primed to see bias. This creates heterogeneity in adoption and use-value across political segments.
  • Strategic externalities and incentives
    • Partisan actors can reduce the public-good value of LLMs by credibly or cheaply signaling bias (e.g., media claims, political messaging). This suggests a negative externality: third-party attempts to manipulate perceived neutrality can lower the social returns of information-correcting technology.
    • There is a potential arms race: AI providers and public-interest actors may need to invest in alignment, transparency, audits, or reputation-building to protect persuasive value — these are costly investments whose returns depend on reducing perceived bias.
  • Cost-effectiveness of AI interventions
    • If perceived bias increases the cognitive resistance (more argumentative responses) required to change beliefs, the marginal cost (effort, time, content tailoring) of persuasion rises. Economic evaluations of LLM interventions should incorporate credibility-framing and counter-messaging costs.
  • Market design and product strategy
    • Firms may have incentives to (a) design interfaces and messaging that mitigate perceived partisan alignment (credibility signaling), (b) offer calibrated products for ideologically heterogeneous markets, or (c) monetize credibility (third-party certification).
    • Political economy of AI investment: returns to alignment and neutrality-preserving R&D are higher when political actors actively attack neutrality; investors and regulators will factor this into valuations and policy choices.
  • Policy and regulatory implications
    • Regulators interested in maximizing public-information benefits might promote standards for transparency, independent audits of political bias, or disclosure regimes to stabilize perceived neutrality — reducing the effectiveness of partisan credibility attacks.
  • Modeling and future research directions for AI economists
    • Incorporate perceived-source credibility as an explicit state variable in diffusion/persuasion models (e.g., Bayesian updating with a credibility prior or a two-stage game where actors invest in shaping perceived bias).
    • Analyze equilibrium where political actors allocate resources to (i) discredit LLMs, (ii) invest in pro-LMM reputation, and (iii) use LLMs for persuasion — derive implications for welfare and information quality.
    • Quantify cost-benefit tradeoffs of alignment investments: how much must platform/firm spending on transparency/audits increase to offset persuasion losses from politicization?
    • Empirically: estimate long-run effects across repeated exposures, different model families (reasoning vs non-reasoning), and in non-U.S. or less polarized contexts to generalize welfare and market outcomes.

Suggested concise takeaway for AI economists: perceived partisan bias is a key determinant of the effectiveness and market value of LLM-based information interventions. Political attacks on neutrality can materially lower the social and private returns to deploying LLMs for persuasion or fact correction, creating incentives for alignment, transparency, and regulatory responses — all of which carry costs that should enter economic assessments and models.

Assessment

Paper Typerct Evidence Strengthhigh — Randomized assignment with a large preregistered sample (N=2,144) provides strong internal validity for causal claims about the effect of partisan credibility warnings on LLM persuasion; transcript analysis supports mechanism evidence. Main limitations are external/ecological validity (survey setting, single LLM, short-term measurement) rather than identification. Methods Rigorhigh — Study is preregistered, uses a large sample size, employs an experimental manipulation with clear counterfactual, and complements outcome measures with qualitative transcript analysis; potential concerns include limited detail on sampling frame/representativeness, possible demand effects, and reliance on immediate post-intervention outcomes. SamplePreregistered U.S. online survey experiment with N=2,144 participants who each completed a three-round conversation with ChatGPT about a personally-held economic policy misconception; participants were randomized to a neutral control or a short message suggesting the LLM was biased against their party; conversation transcripts were collected and analyzed. Themeshuman_ai_collab governance IdentificationPreregistered randomized controlled survey experiment: participants were randomly assigned to receive either a neutral control message or a short warning that the LLM was biased against their political party prior to a three-round ChatGPT conversation correcting a personally-held economic policy misconception; causal effects estimated by comparing outcomes across randomly assigned conditions (intention-to-treat). GeneralizabilityU.S.-only sample — political dynamics may differ in other countries, Online survey experiment may not reflect real-world interactions or longer-term persuasion, Single LLM (ChatGPT) — results may not generalize to other models or future versions, Specific to economic-policy misconceptions — different topics or factual domains could yield different effects, Sample representativeness unclear (e.g., online panel vs probability sample), Short-term attitudinal outcomes measured; behavioral or durable belief change unobserved

Claims (7)

ClaimDirectionOutcomeConfidence & EvidenceDetails
Conversational AI has been proposed as a scalable way to correct public misconceptions and spread misinformation. Other mixed ability of conversational AI to correct misconceptions or spread misinformation
Reading fidelity high
Study strength speculative
not reported
0.1
As LLMs enter partisan conflict, elites increasingly portray them as ideologically aligned. Governance And Regulation negative portrayal of LLMs by elites as ideologically aligned
Reading fidelity medium
Study strength speculative
not reported
0.06
We test whether these credibility attacks reduce LLM-based persuasion. Decision Quality null_result LLM-based persuasion (susceptibility to persuasion following credibility attacks)
Reading fidelity high
Study strength medium
n=2144
0.6
In a preregistered U.S. survey experiment (N=2144), participants completed a three-round conversation with ChatGPT about a personally held economic policy misconception. Other null_result engagement in three-round conversation with ChatGPT about a personally held economic policy misconception
Reading fidelity high
Study strength high
n=2144
1.0
Compared to a neutral control, a short message indicating that the LLM was biased against the respondent's party attenuated persuasion by 28%. Decision Quality negative persuasion (change in respondent beliefs/attitudes following LLM interaction)
Reading fidelity high
Study strength high
n=2144
28%
1.0
Transcript analysis indicates that the warnings alter the interaction: respondents push back more and engage less receptively. Decision Quality negative interaction behavior (pushback frequency, receptiveness) during LLM conversation
Reading fidelity high
Study strength medium
not reported
0.6
The persuasive impact of conversational AI is politically contingent, constrained by perceptions of partisan alignment. Decision Quality negative persuasive impact of conversational AI conditional on perceptions of partisan alignment
Reading fidelity high
Study strength medium
n=2144
0.6

Notes