The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

Google's AI Overview feature cuts English Wikipedia pageviews by about 15%, substituting for many informational pages; culture-related articles suffer the largest traffic losses while STEM topics are less affected.

Impact of AI Search Summaries on Website Traffic: Evidence from Google AI Overviews and Wikipedia
Mehrzad Khosravi, Hema Yoganarasimhan · February 05, 2026
arxiv quasi_experimental medium evidence 8/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Mehrzad Khosravi unresolved corpus identity
  2. Hema Yoganarasimhan unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Mehrzad Khosravi provider ID
  2. Hema Yoganarasimhan provider ID
Exposure to Google's AI Overview reduces daily English Wikipedia traffic by about 15%, with the largest declines in Culture topics and much smaller effects for STEM articles.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

Search engines increasingly display LLM-generated answers shown above organic links, shifting search from link lists to answer-first summaries. Publishers contend these summaries substitute for source pages and cannibalize traffic, while platforms argue they are complementary by directing users through included links. We estimate the causal impact of Google's AI Overview (AIO) on Wikipedia traffic by leveraging the feature's staggered geographic rollout and Wikipedia's multilingual structure. Using a difference-in-differences design, we compare English Wikipedia articles exposed to AIO to the same underlying articles in language editions (Hindi, Indonesian, Japanese, and Portuguese) that were not exposed to AIO during the observation period. Across 161,382 matched article-language pairs, AIO exposure reduces daily traffic to English articles by approximately 15%. Effects are heterogeneous: relative declines are largest for Culture articles and substantially smaller for STEM, consistent with stronger substitution when short synthesized answers satisfy informational intent. These findings provide early causal evidence that generative-answer features in search engines can materially reallocate attention away from informational publishers, with implications for content monetization, search platform design, and policy.

Summary

Main Finding

Google’s AI Overviews (AIO) causally reduced English Wikipedia article pageviews by about 15% after AIO was rolled out in the U.S. (March 22, 2024 → Aug 14, 2024). In the authors’ sample this implies roughly 11.5 million fewer daily visits (≈4.21 billion fewer visits per year, scaled to the sample). Effects are heterogeneous: declines are largest for Culture articles and much smaller for STEM topics, consistent with substitution being strongest when a short synthesized answer can satisfy the user’s informational intent.

Key Points

  • Estimated effect: ≈15% average decline in daily pageviews to English Wikipedia articles following AIO exposure (DiD estimate).
  • Absolute effect: ~−220 daily pageviews per English article (Model with article×language and date fixed effects).
  • Sample: 52,262 English articles with matched versions in Hindi, Japanese, Portuguese, and Indonesian → 161,382 article–language pairs; ~46.5 million article–language–day observations (Oct 28, 2023 – Aug 14, 2024).
  • Identification strategy: difference-in-differences exploiting Google AIO’s staggered U.S. rollout and Wikipedia’s multilingual editions (English treated; other languages as controls).
  • Robustness: similar estimates with alternative post-period definitions, weekly aggregation, and alternative model specifications (PPML, weighted log) reported in appendices.
  • Standard errors: two-way clustering by underlying article and calendar date.
  • Important caveat: Wikimedia pageviews are not reported by reader country → the authors proxy exposure using language edition (English ≈ U.S. exposure during test/early rollout).
  • Context: AIOs appeared on an estimated ~18% of searches according to third-party audits; Google expanded AIO internationally after the study window.

Data & Methods

  • Data source: Wikimedia public pageview API (metrics/pageviews/per-article and metrics/pageviews/top) for daily pageviews, all-access types (desktop + mobile).
  • Time window: Oct 28, 2023 – Aug 14, 2024. Pre-treatment: Oct 28, 2023 – Mar 21, 2024. Post-treatment: Mar 22, 2024 – Aug 14, 2024.
  • Treatment definition: English Wikipedia articles treated (proxying exposure to U.S. AIO rollout); matched same underlying articles in hi/ja/pt/id as controls.
  • Article selection: union of pages that appeared at least once in top-1,000 most-viewed lists (per-language) during July 1, 2023 – Mar 21, 2024; then restrict to only those with English + at least one control-language version.
  • Empirical model (levels): viewsi,l,t = αi,l + δt + β · (Englishl × Postt) + εi,l,t. Authors use levels (not logs) to keep publisher-relevant additive interpretation (impressions → revenue), to avoid undue influence of low-traffic pages, and to avoid sign issues when treated/control magnitudes differ.
  • Inference: article×language fixed effects absorb time-invariant cross-language differences; date fixed effects absorb common shocks; two-way clustering (article and date).
  • Heterogeneity: analyzed by topical categories (Culture vs STEM etc.); substitution strongest in categories where short synthesized answers suffice.

Implications for AI Economics

  • Attention and revenue reallocation: Generative-answer features embedded in SERPs can materially reallocate attention away from third-party informational publishers. For ad-supported publishers, the measured reduction in pageviews implies economically meaningful revenue losses (even though Wikipedia itself is non-advertising in the sample).
  • Differential effects by content type: The stronger declines for culture/soft-information pages versus STEM indicate that AIO’s economic impact depends on query/intent type — search features that provide short factual summaries will substitute more for publishers than those requiring richer content.
  • Platform design and incentives: The finding fuels the policy debate over whether platforms should share ad revenues or otherwise compensate content producers whose content is synthesized into on-platform answers. The authors note potential attribution-based approaches (e.g., Shapley valuation) as one candidate.
  • Long-run content supply risk: If publishers’ revenues fall persistently, this may reduce the incentive to produce or maintain high-quality informational content, creating a negative feedback loop that could degrade the source material search engines rely on.
  • Measurement and regulation: Results point to the need for finer measurement (country-level exposure to AIO, query-level AIO incidence) and possibly regulatory scrutiny of platform interface changes that materially shift ad impressions from publishers to the platform.
  • Research agenda: Further work should (i) obtain user-level/country-level exposure data to refine causal estimates, (ii) quantify advertiser and publisher revenue transfers, (iii) study publisher responses (SEO, paywalls, API usage), and (iv) evaluate long-run effects on content investment and quality.

Limitations to keep in mind: the identification relies on language as a proxy for geographic exposure (Wikimedia does not provide country-disaggregated pageviews), the setting is Wikipedia (highly cited, open encyclopedia) so external validity to other publisher types may vary, and the study window predates some later AIO ad integrations and broader international rollout.

Assessment

Paper Typequasi_experimental Evidence Strengthmedium — The paper leverages a plausible natural experiment (staggered rollout) and a large sample of matched multilingual article pairs, producing a clear negative effect estimate; however, causal interpretation depends on parallel trends across language editions, potential cross-language/search behavior spillovers, and unobserved concurrent changes in search or Wikipedia that could bias estimates, reducing confidence from 'high' to 'medium'. Methods Rigormedium — Uses a large-scale difference-in-differences design with article-level matching across languages and heterogeneous analysis by topic, indicating careful empirical work; but the write-up (as summarized) does not document all identification checks (e.g., pre-trend tests, robustness to geographic spillovers, user-language switching, or alternative control groups) and faces measurement threats (bots, platform UI changes) that leave some methodological concerns. Sample161,382 matched article-language pairs comparing English Wikipedia articles to the same articles in Hindi, Indonesian, Japanese, and Portuguese language editions; daily pageview (traffic) data for these articles around the period of AIO's staggered geographic rollout (exact dates not specified in summary). Themesadoption innovation IdentificationStaggered geographic rollout of Google's AI Overview (AIO) is used as an exogenous shock; a difference-in-differences design compares English Wikipedia articles exposed to AIO in treated geographies to the same underlying articles in other language editions (Hindi, Indonesian, Japanese, Portuguese) that were not exposed during the observation period, thereby controlling for article-level time trends and isolating the effect of AIO exposure on English pageviews. GeneralizabilityFindings are limited to Wikipedia and may not generalize to other publishers or paywalled/content-monetized sites, Only five non-English language editions are used; results may differ in languages/countries not observed, Early-stage/staggered rollout behavior may differ from long-run user behavior after product maturation, Specific to Google's AIO UI and implementation; other search engines or answer formats may have different effects, Does not directly measure revenue or ad monetization effects, only pageviews (attention), Possible user-language switching and cross-border search behavior could limit external validity across geographies

Claims (7)

ClaimDirectionOutcomeConfidence & EvidenceDetails
Search engines increasingly display LLM-generated answers shown above organic links, shifting search from link lists to answer-first summaries. Adoption Rate mixed presence of LLM-generated answer-first summaries in search engine results pages
Reading fidelity high
Study strength low
not reported
0.24
Publishers contend these summaries substitute for source pages and cannibalize traffic, while platforms argue they are complementary by directing users through included links. Firm Revenue mixed claimed impact on publisher traffic and referral behavior
Reading fidelity high
Study strength speculative
not reported
0.08
We estimate the causal impact of Google's AI Overview (AIO) on Wikipedia traffic by leveraging the feature's staggered geographic rollout and Wikipedia's multilingual structure. Firm Revenue null_result causal impact on Wikipedia traffic
Reading fidelity high
Study strength medium
not reported
0.48
Using a difference-in-differences design, we compare English Wikipedia articles exposed to AIO to the same underlying articles in language editions (Hindi, Indonesian, Japanese, and Portuguese) that were not exposed to AIO during the observation period. Firm Revenue null_result comparative change in pageviews between exposed and unexposed language editions
Reading fidelity high
Study strength medium
not reported
0.48
Across 161,382 matched article-language pairs, AIO exposure reduces daily traffic to English articles by approximately 15%. Firm Revenue negative daily traffic (pageviews) to English Wikipedia articles
Reading fidelity high
Study strength high
n=161382
approximately 15%
0.8
Effects are heterogeneous: relative declines are largest for Culture articles and substantially smaller for STEM, consistent with stronger substitution when short synthesized answers satisfy informational intent. Firm Revenue negative relative decline in daily traffic by article category (Culture vs STEM)
Reading fidelity high
Study strength medium
not reported
0.48
Generative-answer features in search engines can materially reallocate attention away from informational publishers, with implications for content monetization, search platform design, and policy. Firm Revenue negative reallocation of user attention away from informational publisher pages (as proxied by reduced pageviews)
Reading fidelity high
Study strength medium
n=161382
0.48

Notes