The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

AI shopping agents enable a subtle, hard-to-detect market failure: platforms and sellers independently exploit LLM biases and jointly inflict more than twice the consumer harm of isolated manipulation. Simulations calibrated to LLM behavior suggest this vertical tacit collusion arises from aligned incentives rather than agreement, exposing an urgent regulatory gap.

Vertical tacit collusion in AI-mediated markets
Felipe M. Affonso · January 06, 2026
arxiv theoretical medium evidence 8/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Felipe M. Affonso unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Felipe M. Affonso provider ID
Calibrated multi-agent simulations show that platforms and sellers independently exploiting LLM cognitive biases create a super-additive form of 'vertical tacit collusion' that more than doubles consumer harm compared with independent manipulation and can evade antitrust detection.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

AI shopping agents are being deployed to hundreds of millions of consumers, creating a new intermediary between platforms, sellers, and buyers. We identify a novel market failure: vertical tacit collusion, where platforms controlling rankings and sellers controlling product descriptions independently learn to exploit documented AI cognitive biases. Using multi-agent simulation calibrated to empirical measurements of large language model biases, we show that joint exploitation produces consumer harm more than double what would occur if strategies were independent. This super-additive harm arises because platform ranking determines which products occupy bias-triggering positions while seller manipulation determines conversion rates. Unlike horizontal algorithmic collusion, vertical tacit collusion requires no coordination and evades antitrust detection because harm emerges from aligned incentives rather than agreement. Our findings identify an urgent regulatory gap as AI shopping agents reach mainstream adoption.

Summary

Main Finding

Platforms and sellers independently learn to exploit systematic cognitive biases of AI shopping agents, producing a novel market failure—“vertical tacit collusion”—where consumer harm from joint exploitation is super-additive. In calibrated multi-agent simulations, joint learning by platform and sellers reduces consumer surplus by 37.1% (from 0.303 to 0.191). That harm is more than double what would be predicted if platform and seller strategies acted independently (pure complementarity = +19.7 percentage points). This coordinated-seeming harm requires no communication and therefore can evade traditional antitrust detection.

Key Points

  • Definition: Vertical tacit collusion — independent, strategic alignment across vertically positioned actors (platform sets ranking; sellers craft product-level inputs) that jointly exploit predictable biases in AI intermediaries (shopping agents).
  • Mechanism: Platform ranking decides which items occupy bias-triggering positions; seller manipulation controls conversion/appeal of those positions. Together they act as complements, amplifying consumer harm.
  • Distinction from horizontal collusion: Horizontal algorithmic collusion requires coordination among competitors on the same instrument (e.g., price). Vertical tacit collusion requires no agreement because incentives are naturally aligned toward exploiting a shared AI target.
  • Empirical magnitudes from simulation:
    • Baseline (quality ranking, no manipulation): consumer surplus = 0.303.
    • Platform-only learning: CS = 0.222 (−27.0%).
    • Seller-only learning: CS = 0.333 (+9.6% — manipulation can be pro-consumer under quality ranking).
    • Joint learning: CS = 0.191 (−37.1%). Joint harm > sum of parts; complementarity = 19.7 pp (95% CI [18.3, 21.1], Cohen’s d = 2.75).
  • Gatekeeper effect: The platform’s ranking choice is a binary-like switch. Seller manipulation helps consumers under quality-weighted ranking but becomes massively harmful if the platform fully abandons quality in favor of bid-weighted rankings (bid-weight = 1.0 → seller effect +69.2% harm).
  • Dominant exploitation channel: Position/serial-position bias is the primary driver—position bias alone explains ~79% of the full-model joint harm (position-only → 29.4% harm). Without position bias, manipulation often benefits consumers.
  • Robustness and falsification:
    • Results hold across 82 robustness checks (different learning algorithms, parameter variations).
    • Replacing the biased AI agent with a debiased agent collapses complementarity by 96%; random-choice agent eliminates it. Exploitable AI biases are necessary for the phenomenon.
    • Human overrides (50% override rate) reduce but do not eliminate harm; heterogenous user susceptibility still allows substantial effects.

Data & Methods

  • Market model:
    • Players: one platform P, six sellers S1..S6 (qualities q ∈ {0.90,0.75,0.60,0.45,0.30,0.20}), one AI shopping agent C representing consumers.
    • Platform actions: choose bid-weight w ∈ {0,0.33,0.67,1}, endorsement rule, decoy insertion d ∈ {0,1}.
    • Seller actions: manipulation intensity m ∈ {0,1,2,3} (keywords, anchoring, framing), bid level b ∈ {0,1,2}.
  • AI agent choice model:
    • Perceived utility Ui = α·qi − β·pi + Bi, where Bi = position effects + prime/first-position + recency + endorsement + manipulation×visibility(ri) + decoy effect.
    • Choice probabilities via softmax/logit (standard random-utility / Type I extreme value).
    • Bias parameterization calibrated to empirical LLM findings (e.g., ACES benchmark, position effects 8–10× larger than quality-price variation; first-position and anchoring effects from LLM literature).
  • Learning and simulation:
    • Platform and sellers learn via Q-learning (temporal-difference, following Calvano et al. implementation); AI agent policy fixed and biased.
    • Each simulation: 20,000 rounds; analysis on final 40% to ensure convergence; 100 independent trials per condition.
  • Experimental conditions compared:
    • Fair baseline (quality ranking, no seller manipulation),
    • Platform-only learning,
    • Seller-only learning (platform uses quality ranking),
    • Joint learning (both adapt).
  • Factorial analysis: 2^4 combinations toggling four bias channels (position, endorsement, manipulation, decoy) to identify channel contributions.
  • Robustness: 82 specifications including alternate learning algorithms, parameter sensitivity, and falsification (debiased and random agents).

Implications for AI Economics

  • New regulatory gap: Traditional antitrust frameworks focus on evidence of agreement or overt coordination. Vertical tacit collusion produces large consumer harm from independent optimization across vertical actors and therefore may evade existing legal standards focused on horizontal collusion or explicit vertical agreements.
  • Platform as critical intervention point: Because platform ranking acts as the gatekeeper switch that enables seller manipulation to become harmful, policies targeting platform ranking rules (e.g., limits on bid-weighting, transparency about ranking criteria, constraints on monetized prominence features) are high-leverage.
  • Prioritize position-bias mitigation: Position/serial-position effects in LLMs are the dominant channel. Technical fixes (de-biasing models, response-randomization in rank-sensitive contexts) and platform interface changes (shuffle or normalize position influence, expose alternative ranking views) could eliminate most of the joint harm.
  • Auditing and standard-setting for AI intermediaries: Regular measurement of shopping-agent biases (position, anchoring, framing susceptibility) should be a standard part of platform audits. Regulators could mandate independent bias testing and certification for shopping agents used in commerce.
  • Limitations of user-side remedies: Human oversight and heterogeneous user susceptibility reduce but do not eliminate harm. Relying primarily on consumer overrides or education is insufficient.
  • Antitrust and competition policy implications:
    • Enforcement tools need to recognize harms that arise from aligned incentives across vertical actors exploiting shared AI targets, not only from explicit agreements.
    • Remedies could include rules constraining revenue-driven ranking mechanics, disclosure obligations, technical standards for de-biasing LLMs in high-stakes recommendation contexts, and platform liability for design choices that enable exploitation of deployed AI intermediaries.
  • Research priorities: Empirical field studies measuring real-world incidence of vertical tacit collusion, interventions that alter platform ranking rules, and verification of mitigation effects in deployed shopping agents.

Summary takeaway: When consumers delegate decisions to widely shared, biased AI shopping agents, platform ranking rules plus seller-level manipulations can jointly produce large, hard-to-detect consumer harms. The platform’s ranking design is the core policy lever to prevent this new form of vertical, tacit exploitation.

Assessment

Paper Typetheoretical Evidence Strengthmedium — The study offers plausible causal mechanisms and counterfactual comparisons via calibrated simulations, lending internal validity for the modeled settings; however, it lacks real-world experimental or observational causal estimates and depends on assumptions in the simulation and calibration, limiting external validity. Methods Rigormedium — Using multi-agent simulation and calibration to empirical LLM bias measurements is an appropriate and rigorous approach for exploring mechanisms and interactions, but the rigor depends on the quality and representativeness of the calibration data, the transparency of assumptions, robustness checks, and sensitivity analyses—elements not verifiable from the abstract alone. SampleSynthetic multi-agent market simulated agents representing platforms (rankers), sellers (product descriptions), and consumers (LLM-mediated shopping agents), with model parameters and LLM cognitive-bias behaviors calibrated to empirical measurements from large language models; no primary field or administrative market-level data are reported in the abstract. Themesgovernance adoption IdentificationMulti-agent computational simulation calibrated to empirical measurements of large language model (LLM) cognitive biases; the paper compares counterfactual simulation scenarios (platform-only exploitation, seller-only exploitation, and joint exploitation) to infer the incremental consumer harm attributable to joint strategies rather than from observational causal identification. GeneralizabilityCalibrations depend on measurements of specific LLMs and prompts; results may not hold for other models or prompt contexts, Simulations rely on modeling assumptions about agent objectives, consumer behavior, and market structure that may diverge from real-world complexity, Does not incorporate dynamic real-world responses such as regulatory intervention, platform policy changes, or long-run seller adaptation outside the modeled strategies, Heterogeneity across product categories, geographies, and user populations is likely underrepresented, Scale effects and multi-platform interactions in actual markets could alter the magnitude or direction of harms

Claims (7)

ClaimDirectionOutcomeConfidence & EvidenceDetails
AI shopping agents are being deployed to hundreds of millions of consumers, creating a new intermediary between platforms, sellers, and buyers. Adoption Rate positive deployment/adoption of AI shopping agents
Reading fidelity high
Study strength low
hundreds of millions
0.06
We identify a novel market failure: vertical tacit collusion, where platforms controlling rankings and sellers controlling product descriptions independently learn to exploit documented AI cognitive biases. Market Structure negative existence of a market failure (vertical tacit collusion)
Reading fidelity high
Study strength medium
not reported
0.12
Using multi-agent simulation calibrated to empirical measurements of large language model biases, we show that joint exploitation produces consumer harm more than double what would occur if strategies were independent. Consumer Welfare negative consumer harm from joint exploitation versus independent strategies
Reading fidelity high
Study strength medium
more than double
0.12
This super-additive harm arises because platform ranking determines which products occupy bias-triggering positions while seller manipulation determines conversion rates. Consumer Welfare negative mechanism linking ranking and seller manipulation to increased consumer harm
Reading fidelity high
Study strength medium
not reported
0.12
Unlike horizontal algorithmic collusion, vertical tacit collusion requires no coordination and evades antitrust detection because harm emerges from aligned incentives rather than agreement. Governance And Regulation negative detectability of collusive harm by antitrust enforcement
Reading fidelity high
Study strength low
not reported
0.06
Our findings identify an urgent regulatory gap as AI shopping agents reach mainstream adoption. Governance And Regulation negative existence of a regulatory gap related to AI shopping agents
Reading fidelity high
Study strength medium
not reported
0.12
Large language models exhibit cognitive biases that are documented and used to calibrate the simulation. Ai Safety And Ethics negative presence of LLM cognitive biases (empirically measured)
Reading fidelity high
Study strength medium
not reported
0.12

Notes