The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

Guaranteeing initial exploration views raises creator activity—about 8.6% more videos per creator and 7.1% more creators posting—yet short-window viewer A/B tests can miss the long-run benefit because the shared content corpus evolves slowly and requires long, isolated experiments to measure total value.

Content Exploration Beyond the Feed: Creator Supply and the Shared Corpus
Yuanyuan Shen, Yiren Yan, Wenjie Li, Chunhui Zhu · August 29, 2026
arxiv rct high evidence 9/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Yuanyuan Shen unresolved corpus identity
  2. Yiren Yan unresolved corpus identity
  3. Wenjie Li unresolved corpus identity
  4. Chunhui Zhu unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Yuanyuan Shen provider ID
  2. Yiren Yan provider ID
  3. Wen-Jie Li unresolved corpus identity
  4. Chunhui Zhu provider ID
Budgeted exploration meaningfully raises creator supply (≈8.55% more videos per creator and 7.10% more creators posting) while producing modest and mixed immediate viewer effects, and because exploration feeds a shared corpus that turns slowly, short-run viewer-side A/B tests cannot recover the mechanism’s full long-run value.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

No provider observation is available for this paper.

Missing data, not a zero citation count.

Industrial recommenders give new content initial views through budgeted exploration, then use early performance to decide further delivery. On many short-video platforms, exploration is the primary way new videos reach viewers. Viewer-side tests measure consumption; the published budget objectives we review omit creator response. We analyze four experiments on a major short-video platform. An eight-month creator ablation finds production exploration raises videos posted per creator by 8.55% and creators posting at least once by 7.10% relative to a minimal floor. A budget-matched reallocation raises creator participation with no detectable short-run viewer-side change. A year-long viewer ablation finds 1.74% more video views but 2.13% less view time. A delivered view creates immediate feed value, can trigger organic take-up, and can induce creator supply. Take-up and supply replenish a shared corpus, creating two measurement limits. Viewer-side A/B tests cancel the corpus effect when both arms consume the same corpus. Giving each arm its own corpus avoids cancellation, but turnover still controls the horizon. If the corpus turns over at rate w per posting cycle, a t-cycle experiment expresses at most wt of the eventual corpus effect. More users reduce noise but do not speed turnover. Before the corpus path visibly bends, data cannot distinguish a modest fast effect from an arbitrarily large slow one, so a valid confidence interval may lack a finite upper endpoint. As predicted, the three-week co-diverted experiment cannot determine the sign of the eventual corpus effect. Within the window, it identifies the direct feed effect, and an exploratory cohort analysis detects organic lift after exploration ends. The experiments establish a positive creator response, measure the gross corpus flow visible within three weeks, and show the design and duration needed to identify total value.

Summary

Main Finding

Exploration in short-video recommenders not only affects immediate viewer metrics but meaningfully increases creator supply and expands the discoverable corpus. The paper shows a signed positive creator response to production exploration (8.55% more videos per creator; 7.10% more creators posting at least once), quantifies short-run viewer trade-offs, and proves that the shared corpus and its turnover impose strict limits on what A/B tests can observe within finite horizons. Identification of the corpus-mediated long-run value requires tailored randomizations (co-diversion / isolated cells) and sufficiently long experiments matched to corpus turnover.

Key Points

  • Empirical magnitudes
    • Eight-month creator ablation: +8.55% videos posted per creator; +7.10% creators posting at least once (relative to a minimal floor).
    • Budget-matched reallocation: raises creator participation with no detectable short-run viewer-side cost.
    • Year-long viewer ablation (disablement vs enabled): +1.74% video views but −2.13% view time (viewer-side trade-off); no sustained buildup over 12 months visible via one-sided viewer test.
    • Three-week co-diverted probe: within that window, the experiment identifies the direct feed effect and detects organic lift after exploration ends, but cannot sign the long-run corpus asymptote.
  • Conceptual decomposition of a delivered exploration view (Proposition 2):
    • Direct creator channel: induced future postings (marginal posting response δ times downstream value).
    • The view itself: immediate consumption value η_e.
    • Take-up channel: increased probability of organic distribution q′(B) times downstream organic reach V_org and its lifetime value.
  • Supply-loop multiplier (Proposition 1): when each submission induces further submissions with expected factor R (<1 for stability), the cumulative posting amplification is 1/(1 − R). As R → 1 the multiplier blows up.
  • Measurement limits from shared corpus and turnover:
    • Viewer-side (one-sided) A/B tests share the same corpus across arms, so corpus-mediated effects cancel and are unobserved.
    • Co-diversion (isolated arms each with their own corpus) is required to retain the corpus channel in contrasts, but the corpus has mechanical turnover rate ω. If an experiment runs t posting cycles, it can reveal at most ω t of the eventual corpus effect — i.e., slow stock dynamics attenuate what a finite-horizon experiment can express.
    • Before the corpus trajectory shows curvature, data cannot distinguish a small, fast effect from a large, slow one; confidence sets can therefore fail to have finite upper bounds (weak identification).
  • Design and inference results:
    • Theorem 1 and related propositions formalize horizon bounds and identification limits.
    • Detection/estimation scaling laws for slow-expressing effects are derived (e.g., N^{-1/3} detection, N^{-1/5} estimation rates in the exponential-approach family).
  • Practical observation: increasing sample size reduces noise but does not accelerate the corpus turnover; longer duration (matched to ω) or structural priors/creator experiments are needed to identify the long-run corpus value.

Data & Methods

  • Setting: a major short-video platform (industrial recommender with budgeted exploration as practiced; authors worked with Snap Inc. data).
  • Mechanism formalization:
    • Budgeted exploration: each new item in a pool receives a per-item budget B(s) with minimum B_min and maximum B_max; escalation above B_min is conditioned on early viewer response within a TTL window.
    • Ecosystem model (discrete submission–feedback cycles): states are submission rate Λ_t (new posts per period) and organic corpus K_t (items remaining in organic distribution). Key primitives: baseline Λ0, creator response g(·) with marginal δ = g′(·), take-up probability q(B) with derivative q′(B), organic reach V_org, corpus turnover ω, and per-unit viewer values η.
    • Linearized dynamics (mean-field tangent at steady state):
    • Λ_{t+1} = Λ0 + R Λ_t (R = g( E[V|B] ))
    • K_{t+1} = (1 − ω) K_t + q(B) Λ_t
  • Experiments (four total):
  • Eight-month creator ablation: disables production exploration in some creator cells (compared to minimal floor) to estimate posting response.
  • Budget-matched reallocation: reassigns budgets at fixed nominal aggregate capacity to test whether shifting allocation increases creator participation absent viewer-side change.
  • Year-long viewer ablation: one-sided disablement vs enabled to measure immediate feed trade-offs over a long horizon (measures views and view time).
  • Three-week co-diverted A/B: matched creator and viewer submarkets randomized so each arm develops its own corpus (isolated cells); intended to capture corpus flow visible within a short horizon and probe the estimation floor.
  • Estimation strategy:
    • One-sided designs identify immediate viewer effects and supply responses that express within the chosen horizon, but cancel shared-corpus channel.
    • Co-diverted/isolated designs retain corpus effects but require matching the horizon to corpus turnover ω to reveal asymptotic effects.
    • Cohort analyses and long-running ablations used to separate gross corpus inflow from net long-run value.

Implications for AI Economics

  • Valuation of exploration must account for two-sided externalities: views produce immediate consumer value, induce organic take-up, and trigger creator supply. Failing to account for creator response understates long-run value of exposure policies.
  • Mechanism design: cold-start budget objectives that ignore creator supply are incomplete. Designers should incorporate priors for creator response (δ, R) and corpus-turnover (ω) when setting per-item budgets, escalation rules, and floors/ceilings.
  • Experimental practice on platforms:
    • Standard viewer-side A/B tests can miss corpus-mediated benefits; co-diversion (isolated arms) is necessary to measure corpus effects but requires long horizons matched to corpus turnover.
    • Short-duration tests (weeks) may only reveal the direct feed effect and a fraction (≤ ω t) of the eventual corpus benefit; policy decisions based on such tests risk mis-evaluating long-run welfare.
    • Sample size increases reduce variance but cannot shorten the time-scale for stock dynamics; thus experimentation budgets should trade off duration vs. sample size differently for slow-expressing effects.
  • Inference and policy caution:
    • Slow-expressing two-sided feedback loops can produce weak identification: data within practical horizons may yield confidence intervals without finite upper endpoints for long-run effects. Policymakers and platform economists must combine experiments with structural priors, horizon-matched creator experiments, or very long-term isolation to obtain credible bounds.
  • Broader two-sided market lessons:
    • Small changes in allocation (raising R closer to 1) can produce large multipliers for creator production — beneficial or harmful depending on downstream quality and displacement dynamics. Close-to-critical regimes require careful control because of large amplified externalities.
    • Evaluations and regulations focused solely on short-run user metrics may misjudge platform decisions that trade short-run consumption for long-run supply and diversity benefits.
  • Practical recommendations:
    • Include creator-side priors and estimates (from horizon-matched creator experiments) into cold-start budget objectives.
    • Use budget-matched reallocations to isolate creator participation effects without altering nominal capacity.
    • Run co-diverted experiments over durations informed by measured or plausible ω; if impractical, supplement with structural modeling and prior information to bound the long-run effects.

Limitations noted by authors - The linear/mean-field model and local multiplier apply near steady state; opposing-sign feedbacks (crowd-out, quality dilution) could alter conclusions and require additional modeling. - Some heterogeneity (δ_c, q′_s) is not identified by these designs; within-pool re-ranking by creator responsiveness remains an open challenge. - Practical identification of the asymptotic net corpus value may require experiments much longer than typical operational windows.

Assessment

Paper Typerct Evidence Strengthhigh — Multiple large-scale randomized experiments on a major production platform run over long horizons (weeks to a year) provide direct causal estimates of creator posting responses and immediate viewer effects; theoretical results clarify which experimental contrasts identify which channels and why short horizons miss corpus value. Methods Rigorhigh — Careful combination of long-running randomized ablations, budget-matched reallocations, and co-diversion designs, plus a formal dynamic model that derives multipliers, decomposition of marginal view value, and limits on observability; authors explicitly address identification failures, horizon constraints, and report cohort analyses to probe slow channels. SamplePlatform data from a major short-video service (authors at Snap Inc.); experiments randomize submarkets/users and/or creator-facing exposure policies and observe creator-level outcomes (videos posted per creator, share of creators posting at least once) and viewer-side metrics (views, view time) across millions of impressions over windows ranging from three weeks to one year (eight-month creator ablation, one-year viewer ablation, budget-matched reallocation, three-week co-diverted probe). Themeshuman_ai_collab adoption IdentificationRandomized field experiments on a large short-video platform: (1) an eight-month creator-side ablation (disablement vs minimal floor), (2) a budget-matched reallocation, (3) a year-long viewer-side ablation (disablement vs enablement), and (4) a three-week co-diverted (cohort/isolation) experiment; random assignment of submarkets/segments isolates direct viewer feed effects, creator posting responses, and (with co-diversion) corpus-mediated effects, combined with theoretical identification of which contrasts recover which channels. GeneralizabilityResults are specific to short-video platforms with budgeted per-item exploration windows and TTL-based escalation/withdrawal policies, and may not generalize to platforms using different cold-start policies (e.g., pure bandits or non-budgeted ranking)., Platform-specific factors (Snap’s user base, ranking architecture, organic reach dynamics) may limit external validity to other geographies, content types, or long-form video., Linear mean-field model and local multiplier apply near observed steady states; settings with strong quality dilution, crowd-out, or very different creator heterogeneity may behave differently., Co-diversion and long-horizon demands mean operational constraints (ability to isolate corpora, willingness to run long experiments) may limit replication elsewhere.

Claims (9)

ClaimDirectionOutcomeConfidence & EvidenceDetails
Production exploration raises the number of videos posted per creator by 8.55% relative to a minimal exploration floor. Innovation Output positive Videos posted per creator
Reading fidelity high
Study strength medium
8.55%
0.6
Production exploration increases the share of creators posting at least once by 7.10% relative to a minimal exploration floor. Adoption Rate positive Creators posting at least once
Reading fidelity high
Study strength medium
7.10%
0.6
A budget-matched reallocation increases creator participation without producing a detectable short-run viewer-side change. Task Allocation mixed Creator participation and short-run viewer-side outcomes
Reading fidelity high
Study strength medium
not reported
0.6
A year-long viewer ablation produces 1.74% more video views but 2.13% less view time relative to disabling exploration. Consumer Welfare mixed Video views and total view time
Reading fidelity high
Study strength medium
1.74% more video views; 2.13% less view time
0.6
Within the three-week co-diverted experiment, the direct feed effect is identifiable, but the eventual shared-corpus effect cannot be determined to be positive or negative. Consumer Welfare null_result Eventual corpus-mediated viewer value and immediate feed effects
Reading fidelity high
Study strength low
not reported
0.3
An exploratory cohort analysis detects organic lift after the exploration period ends. Consumer Welfare positive Organic content take-up or post-exploration distribution
Reading fidelity high
Study strength low
not reported
0.3
In a one-sided viewer-side A/B test, the shared-corpus contribution cancels because both experimental arms consume the same corpus. Consumer Welfare null_result Measured corpus-mediated treatment contrast
Reading fidelity high
Study strength high
not reported
1.0
If the corpus turns over at rate ω per posting cycle, a t-cycle experiment expresses at most ωt of the eventual corpus effect. Organizational Efficiency positive Fraction of the eventual corpus-mediated effect expressed during an experiment
Reading fidelity high
Study strength high
at most ωt
1.0
Under the linear subcritical ecosystem model, one additional delivered view induces δ/(1−R) cumulative additional submissions, where R is the supply-loop gain and R<1. Innovation Output positive Cumulative additional creator submissions induced by one delivered view
Reading fidelity high
Study strength speculative
δ/(1−R) cumulative additional submissions
0.1

Notes