The Commonplace
Home Three-study pilot Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

ChatGPT did not meaningfully reduce Reddit help-seeking: informational subreddits saw no post-volume decline and AI-generated posts rose only a few percentage points—insufficient to account for earlier reports of large drops, which appear driven in part by pre-existing platform drift.

Informational Help-Seeking on Reddit Did Not Decline After ChatGPT
Hazem Ibrahim, Yasir Zaki · September 11, 2026
arxiv quasi_experimental high evidence 7/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Hazem Ibrahim unresolved corpus identity
  2. Yasir Zaki unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Hazem Ibrahim provider ID
  2. Yasir Zaki provider ID
Using contemporaneous hobby-community controls, placebo-date diagnostics, and calibrated AI-text detectors on hundreds of thousands of documents, the authors find no meaningful decline in informational help-seeking on Reddit after ChatGPT's launch and only a small rise in AI-like posts insufficient to explain previously reported drops.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

Did people stop asking other people for advice online once generative AI could answer their questions? Prior work on ChatGPT's effect on online help-seeking disagrees in both size and sign, in part because no study has compared affected communities against similar communities that AI cannot easily substitute for, over the same months. In this paper, we track monthly post counts in 26 Reddit informational communities against 90 size-comparable hobby communities over the same six calendar months before and after the launch of ChatGPT. We also repeat the entire analysis at 66 earlier dates, before ChatGPT existed, to see what our method reports when no ChatGPT-effect exists. We find that informational help-seeking did not decline. Our results rule out any decline in posting larger than 3.4%, far smaller than the 8% to 25% declines documented in prior work. Steady post counts could still be misleading if AI-written posts had replaced human ones. We test this possibility by scoring 274,411 posts and 223,775 comments with AI-text detectors, compared in a way that cancels out detector false-positives on human-written text. AI-written posts rose only 2-3 percentage points more in informational communities than in hobby communities, short of the 5.1 points that would be needed to hide even the smallest decline previously reported for Reddit. In addition, the comments people receive show no such rise at all. Why, then, do published studies disagree? Reddit community types were already drifting apart before ChatGPT existed, at rates comparable to every published estimate, and without same-time controls, that drift can look like an effect of generative AI. Our own largest estimate, an 18% fall in posts to low-stakes curiosity communities, matches its pre-existing trend. Humans still ask humans for help, and, as far as detection can tell, humans still answer them.

Summary

Main Finding

Informational help-seeking on Reddit did not decline after ChatGPT’s launch. Using contemporaneous hobby subreddits as controls and multiple robustness checks (including 66 placebo “launch” dates before ChatGPT), the authors estimate a 5.1% increase in informational posts and rule out any true decline larger than 3.4%. Detectors find only a small relative rise in LLM-like post text (≈2–3 percentage points) in informational communities versus controls, and no detectable rise in LLM-like content among comments (answers).

Key Points

  • Sample: 151 Reddit communities meeting volume and age filters, classified into:
    • I (informational advice): 26 subreddits
    • H (human-anchored support): 29 subreddits
    • Q (low-stakes curiosity): 6 subreddits
    • C (controls — hobby communities): 90 subreddits
  • Primary comparison window: Dec 2021–May 2022 (pre) vs Dec 2022–May 2023 (post), matching calendar months to remove seasonality.
  • Main estimation: two-way fixed-effects difference-in-differences on log(1 + monthly posts), with community and month fixed effects, cluster-robust SEs and wild-cluster bootstrap p-values.
  • Main result: informational communities showed a +5.1% change relative to hobby controls after ChatGPT; confidence bounds exclude declines >3.4% (i.e., a meaningful decline of the magnitudes reported in many prior studies is ruled out).
  • Detector analysis:
    • Scored 274,411 posts and 223,775 comments with Fast-DetectGPT; added Binoculars for comments.
    • Calibrated detector responses using four LLMs (GPT-4o, Claude Sonnet 4, DeepSeek-V3, and GLM-5.2) to convert score shifts into implied shares of machine-generated text.
    • Result: AI-like posts rose only ~2–3 percentage points more in informational communities than in hobby controls (insufficient to conceal the previously reported declines); comments showed no relative increase in AI-like content.
  • Placebo tests: rerunning the entire analysis at 66 pre-ChatGPT placebo dates produced nonzero “effects” as large as many published estimates, demonstrating that pre-existing drift between community types can mimic an AI impact if contemporaneous controls are not used.
  • A notable exception: low-stakes curiosity communities (Q) show an 18.3% decline relative to controls after the launch (p = 0.002), but identical pre-existing trends explain this decline (they were shrinking ~18%/year before ChatGPT), so the decline is not attributable cleanly to ChatGPT.

Data & Methods

  • Data sources:
    • Monthly post counts per subreddit from Arctic Shift (Pushshift successor) for the count-based DID analysis.
    • Text samples for detector analysis drawn from Arctic Shift (up to 100 posts and 100 comments per community-month), two windows: Jan 2021–Nov 2022 (pre-ChatGPT baseline) and Jan 2024–Nov 2025 (post period).
  • Regression specification:
    • log(1 + y_st) = α_s + γ_t + Σ_g δ_g · 1[s ∈ g]·Post_t + ε_st
    • δ_g are group-by-post indicators (difference-in-differences relative to control group C).
    • Community fixed effects (α_s) and calendar-month fixed effects (γ_t) control for time-invariant community differences and platform-wide shocks.
    • SEs clustered by community; wild-cluster bootstrap used for p-values.
  • Power & sensitivity:
    • Noise dominated by slow community-level drift; with 26 treated informational communities the detectable decline at 80% power was ~14% (5% α) for the primary window. Despite this, empirical confidence intervals still exclude declines >3.4%.
  • Detector approach:
    • Used Fast-DetectGPT (curvature-based scoring using GPT-Neo-1.3B) as the primary detector; Binoculars added for comments.
    • Outcome for detector DID: change in mean detector score per community-period, which cancels time-invariant detector biases (e.g., against non-native English) because the comparison is relative to controls and baseline period.
    • Calibration: sampled real pre-ChatGPT posts and had four LLMs generate matching documents; compared detector responses to translate score shifts into plausible percentage-point increases in LLM-generated documents.
  • Robustness & diagnostics:
    • 66 placebo dates (same calendar-month design) to quantify pre-existing drift and estimate false positives the design would produce in the absence of ChatGPT.
    • Longer window (2021–22 vs 2024–25) also reported but overlaps other platform events (e.g., Google–Reddit licensing, Google AI Overviews), so treated cautiously.

Implications for AI Economics

  • Substitution magnitude: On Reddit, generative AI (ChatGPT-era models) does not appear to have substantially substituted away public informational help-seeking at the community level. Any displacement in posted questions would have to be smaller than a few percent on average, at least in the six-month post-launch window studied.
  • Complementarity and persistence of human-mediated help: The stability of post volumes and the small relative rise in LLM-like posts suggest that people continue to ask humans in public forums, and — as detectable — humans still provide most answers (comments). This points to persistent complementarity or frictions to substitution (trust, norms, desire for human validation, or privacy preferences).
  • Methodological lesson for AI-economics research: contemporaneous controls that are plausibly unaffected (or less affected) by the technology are crucial. Pre-existing trends between groups can generate spurious large estimated effects if not accounted for; placebo-date diagnostics are a practical tool to reveal such drift.
  • Measurement of AI content: detector-based, score-differencing approaches calibrated on platform-relevant text and multiple LLMs provide a defensible, bias-cancelling way to measure relative change in LLM-like content without binary labeling. Still, detector limitations (bias vs non-native English, document-length sensitivity) mean conclusions about exact shares should be cautious.
  • Policy and market forecasting:
    • Forecasts of labor or platform disruption from LLMs should allow for limited immediate displacement in low-cost, public Q&A contexts; large-scale substitution may require overcoming additional behavioral, privacy, or trust barriers (or will show up more in private product usage rather than public posts).
    • Firms and platforms whose business models assume rapid erosion of user-generated informational content should reassess timing and magnitude assumptions; replacement dynamics may be heterogeneous across domains and platforms (e.g., differences between Reddit and Stack Overflow reported in other studies).
  • Directions for further research:
    • Account-level analyses (with careful privacy protections) to track whether individual human posters are replaced by LLM-written posts.
    • Cross-platform studies (e.g., Stack Overflow, private chat apps, search engines) to map where substitution is stronger.
    • Quality, usefulness, and welfare impacts of LLM-generated answers in community settings (not just volume).
    • Longer-term studies that account for later adoption phases and concurrent platform changes (search- and licensing-driven shifts).

Limitations to keep in mind: detector imperfectness (and known biases), inability to follow individual accounts for privacy reasons (limits tests of who left and who replaced them), small treated sample for Q (n=6), and later-window confounding events (Google–Reddit deals). Overall, the study provides evidence that public, community-based informational help-seeking on Reddit remained largely resilient to ChatGPT in the early adoption period.

Assessment

Paper Typequasi_experimental Evidence Strengthhigh — The paper uses contemporaneous controls on the same platform, carefully matches calendar months, runs extensive placebo tests across 66 pre-event dates to detect pre-existing drift, clusters and bootstraps inference, and triangulates posting-volume results with detector-based evidence on authorship—together providing strong, robust evidence for the stated null/near-null effect; remaining caveats (platform scope, detector limits, inability to follow individual accounts) lower external validity but do not substantially weaken the internal identification. Methods Rigorhigh — Appropriate two-way FE DID with community and month fixed effects, clustered SE and wild-cluster bootstrap, calendar-month matching to remove seasonality, systematic placebo-date checks to diagnose pre-trends, pre-registered-like robustness checks and detector calibration versus multiple LLMs; limitations include detectable-power constraints for small effects, exclusion of short documents, detector biases (non-native English), and no account-level panel. Sample151 Reddit communities selected from an initial 157 after mechanical filters (founded by Dec 2017, median ≥100 posts/month, ≥50 posts in 90% of months, no missing months); groups: 26 informational (I), 29 human-anchored support (H), 6 low-stakes curiosity (Q), 90 hobby controls (C). Monthly post counts from the Arctic Shift Reddit archive across multiple windows (primary: Dec 2021–May 2022 vs Dec 2022–May 2023). Detector analysis sampled up to 100 posts and 100 comments per community-month across Jan 2021–Nov 2022 (pre) and Jan 2024–Nov 2025 (post), yielding large raw pools (~680k posts, ~689k comments) and final scored samples reported in the paper (e.g., 274,411 posts and 223,775 comments after preprocessing/length filters). Themeshuman_ai_collab adoption IdentificationDifference-in-differences (two-way fixed effects) comparing changes in monthly post counts in 26 informational Reddit communities (treated) to 90 size-comparable hobby communities (contemporaneous controls) over matched calendar-month windows before vs after ChatGPT's launch; clustering standard errors by community with wild-cluster bootstrap p-values; 66 placebo 'launch' dates rerun across a 59-month pre-period to measure pre-existing drift; complementary DID on mean AI-detector scores of sampled posts/comments (detector calibration using four LLMs) to test for replacement by machine-generated text. GeneralizabilityFindings restricted to Reddit communities and may not generalize to other platforms (Stack Overflow, Google Search, private messaging, paid support)., Primarily English-language / culturally specific community norms—non-English communities or global populations may behave differently., Selected community sample excludes very small subreddits and communities founded after 2017., Early-adoption window: primary comparison captures early post-launch behavior and may not reflect long-run adoption dynamics or later model improvements., Detector-based inferences limited by detector biases (e.g., false positives for non-native English), document-length exclusions, and inability to attribute authorship at the account level.

Claims (10)

ClaimDirectionOutcomeConfidence & EvidenceDetails
Informational help-seeking on Reddit did not decline after ChatGPT was released. Task Allocation null_result Monthly number of posts in informational Reddit communities relative to hobby-community controls
Reading fidelity high
Study strength high
n=116
5.1% increase
0.8
The study rules out an informational-post decline larger than 3.4% after ChatGPT's launch. Task Allocation null_result Relative change in monthly informational-community post volume
Reading fidelity high
Study strength high
n=116
decline larger than 3.4% ruled out
0.8
AI-written posts increased only 2–3 percentage points more in informational communities than in hobby communities after ChatGPT's release. Automation Exposure positive Relative increase in machine-like or AI-generated text in sampled Reddit posts
Reading fidelity high
Study strength medium
n=274411
2–3 percentage points
0.48
The estimated increase in AI-written posts was insufficient to conceal even the smallest previously reported Reddit decline in posting. Task Allocation null_result Whether increased AI-generated content could offset a decline in human-authored post volume
Reading fidelity high
Study strength medium
n=274411
2–3 percentage points versus 5.1 percentage points needed
0.48
Comments received by posts in informational communities did not show an additional accumulation of AI-generated content relative to hobby communities. Automation Exposure null_result Relative change in machine-like or AI-generated text in Reddit comments
Reading fidelity high
Study strength medium
n=223775
0.48
Post volume in low-stakes curiosity communities fell by 18.3% relative to hobby controls after ChatGPT's launch. Task Allocation negative Monthly post volume in low-stakes curiosity Reddit communities relative to controls
Reading fidelity high
Study strength medium
n=96
18.3% fall, p = 0.002
0.48
The 18.3% post-volume decline in low-stakes curiosity communities is consistent with a pre-existing trend rather than a ChatGPT effect. Task Allocation mixed Pre-existing relative trend in monthly post volume of low-stakes curiosity communities
Reading fidelity high
Study strength medium
n=6
roughly 18% a year
0.48
Community types exhibited pre-existing trends large enough to resemble published estimates of ChatGPT's effect on online help-seeking. Task Allocation mixed Relative change in monthly post volume between treated community types and controls at pre-ChatGPT placebo dates
Reading fidelity high
Study strength medium
n=66
effects as large as the published estimates
0.48
The surviving Reddit sample contained 151 communities: 26 informational, 29 human-anchored support, 6 low-stakes curiosity, and 90 hobby control communities. Other null_result Study sample composition
Reading fidelity high
Study strength high
n=151
0.8
The primary comparison used the same six calendar months before and after ChatGPT's release to remove seasonal differences. Organizational Efficiency positive Validity of estimated change in monthly Reddit post counts
Reading fidelity high
Study strength high
n=151
approximately 4% potential bias from an unmatched window
0.8

Notes