The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

A lightweight LLM assistant that gives gist-level feedback improves decision accuracy and confidence under uncertainty and curbs unnecessary oversampling; verbatim feedback, by contrast, drives more extensive exploration, suggesting designers should match feedback granularity to the uncertainty profile of the task.

Conversational Decision Support for Information Search Under Uncertainty: Effects of Gist and Verbatim Feedback
Kexin Quan, Jessie Chin · February 16, 2026
arxiv rct medium evidence 7/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Kexin Quan unresolved corpus identity
  2. Jessie Chin unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Kexin Quan provider ID
  2. Jessie Chin provider ID
An LLM-based decision-support assistant that provides gist summaries improves decision accuracy and confidence—especially under higher uncertainty—and reduces excessive sampling, whereas verbatim feedback encourages broader exploration.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

Many real-world decisions rely on information search, where people sample evidence and decide when to stop under uncertainty. The uncertainty in the environment, particularly how diagnostic evidence is distributed, causes complexities in information search, further leading to suboptimal decision-making outcomes. Yet AI decision support often targets outcome optimization, and less is known about how to scaffold search without increasing cognitive load. We introduce SERA, an LLM-based assistant that provides either gist or verbatim feedback during search. Across two experiments (N1=54, N2=54), we examined decision-making outcomes and information search in SERA-Gist, SERA-Verbatim, and a no-feedback baseline across three environments varying in uncertainty. The uncertainty in environment is operationalized by the perceived gain of information across the course of sampling, which individuals may experience diminishing return of information gain (decremental; low-uncertainty), or a local drop of information gain (local optimum; medium-uncertainty), or no patterns in information gain (high-uncertainty), as they search more. Individuals show more accurate decision outcomes and are more confident with SERA support, especially under higher uncertainty. Gist feedback was associated with more efficient integration and showed a descriptive pattern of reduced oversampling, while verbatim feedback promoted more extensive exploration. These findings establish feedback representation as a design lever when search matters, motivating adaptive systems that match feedback granularity to uncertainty.

Summary

Main Finding

An LLM-based conversational assistant (SERA) that provides on-demand summaries during sequential information search improves decision accuracy and confidence—particularly in high-uncertainty environments. The format of feedback matters: gist (high-level) summaries encourage more efficient integration and reduced oversampling, while verbatim (detail-preserving) summaries encourage more extensive exploration. These effects imply that matching feedback granularity to the environment’s uncertainty structure can improve exploration–exploitation regulation.

Key Points

  • Problem framed: many real-world choices require sequential information search and stopping under uncertainty; AI support typically targets final outcomes but may better serve users by scaffolding the search process.
  • SERA: a conversational, LLM-driven assistant that (a) summarizes user-recorded notes on demand and (b) prompts brief monitoring (self-regulatory) reflections.
  • Two feedback formats tested:
    • Gist: compressed, high-level meanings emphasizing what matters now.
    • Verbatim: literal, detailed recaps preserving attribute-level content.
  • Environments (operationalized by how perceived information gain evolves during sampling):
    • Low-uncertainty (Decremental): predictable, steadily diminishing information gain.
    • Medium-uncertainty (Local Optimum): a local drop then rise in information gain (risk of stopping too early).
    • High-uncertainty (Random): no predictable pattern in information gain.
  • Main behavioral patterns:
    • SERA (either format) → higher decision accuracy and greater confidence than no-feedback, with larger gains in higher-uncertainty settings.
    • Gist → more efficient integration, descriptive reduction in oversampling (i.e., fewer redundant samples; better stopping).
    • Verbatim → greater exploration (more sampling), which can help avoid local optima but may increase cognitive cost.
  • Design implication emphasized by authors: feedback representation is a manipulable design lever; adaptive interfaces should match summary granularity to environmental predictability.

Data & Methods

  • Overall design: two experiments (N1 = 54, N2 = 54) using a web-based search-and-stop decision task adapted from decisions-from-experience paradigms.
  • Conditions: SERA-Gist, SERA-Verbatim, and a no-feedback baseline crossed with three uncertainty environments (decremental, local-optimum, random).
  • Interaction flow: participants click to reveal sequential descriptive information for two options, record notes, request summaries from SERA, respond to a monitoring prompt, then continue sampling or stop and choose.
  • SERA implementation:
    • Frontend: React.js; backend: Flask; data logged to Firebase.
    • LLM: OpenAI gpt-4o-mini; prompt framed the model as a neutral assistant with concise outputs. Parameters reported (temperature = 1, max tokens = 150, top_p = 0).
    • Summaries generated from participant-entered notes; format (gist vs verbatim) controlled via prompt framing.
  • Outcomes measured:
    • Outcome effectiveness: decision accuracy, decision confidence.
    • Process regulation: number of samples, stopping behavior, integration patterns (efficiency vs oversampling), exploratory behavior.
    • Secondary: user perceptions and individual differences (to assess reliance patterns).
  • Main results (reported descriptively in manuscript):
    • SERA improves accuracy and confidence relative to baseline; effect sizes increase with environmental uncertainty.
    • Gist reduces redundant sampling and supports efficient stopping; verbatim increases exploration.

Implications for AI Economics

  • Design of decision-support tools in economic settings should go beyond final recommendations to actively scaffold the information search process; the representation of feedback materially affects exploration costs and welfare.
  • Match feedback granularity to environment structure:
    • Predictable/diminishing-return environments (e.g., routine product comparisons, well-known financial instruments): favor gist summaries to reduce sampling costs, speed decisions, and reduce overload.
    • Environments with local optima or sparse but informative late signals (e.g., complex investment research, procurement with hidden features): favor verbatim or detail-rich summaries to encourage deeper exploration and avoid premature stopping.
    • Highly unpredictable markets: SERA-style prompts and summaries increase confidence and accuracy overall; a hybrid/adaptive policy that responds to observed information-gain patterns can balance exploration costs and decision quality.
  • Platform and market applications:
    • E‑commerce search: gist summaries can lower search costs and abandonment; verbatim summaries can be used in expert-aimed workflows where exhaustive evaluation matters.
    • Financial advising and consumer finance: adaptive feedback could reduce under-search or overreliance and improve calibration of confidence in costly uncertainty.
    • Labor markets and procurement: conversational summaries can speed screening while verbatim support can be enabled when the environment indicates the risk of missing late-arriving important signals.
  • Policy and welfare considerations:
    • Designers must consider cognitive costs and risk of overreliance; fluent conversational summaries may be persuasive and increase confidence—even when model outputs are uncertain—so transparency about system limits remains important.
    • Adaptive feedback that detects the payoff-structure and switches granularity can improve aggregate welfare but requires careful evaluation to avoid nudging users toward prematurely low-effort choices.
  • Research & modeling opportunities for AI economics:
    • Formalize the value of information under different information-gain temporal structures and compute optimal feedback policies (when to provide gist vs verbatim).
    • Integrate user search-cost heterogeneity and attention limits into market-design models to predict impacts of different feedback strategies on social welfare, search duration, and market liquidity.
    • Evaluate long-run behavioral effects (habit formation, calibration of trust, exploitation of persuasive summaries) and the strategic consequences for platforms that choose different feedback policies.

Limitations to note for economic application: relatively small experimental Ns, lab-style descriptive information (not monetary gambles), single LLM and prompt family, and potential domain differences; results should be validated in larger, domain-specific field settings before widescale deployment.

Assessment

Paper Typerct Evidence Strengthmedium — Randomized experiments provide strong internal validity for the effect of feedback format on search behavior and decision accuracy, and the study reports two independent experiments. However, sample sizes are modest (N=54 per experiment), tasks are laboratory/online simulations rather than field settings, and ecological validity and robustness across populations, tasks, and LLM implementations are not established. Methods Rigormedium — The design manipulates both feedback type and uncertainty, allowing clean tests of interaction effects; outcome measures include decision accuracy, confidence, and sampling behavior. Missing or unclear details (from the summary) include participant recruitment and representativeness, power calculations, pre-registration, exact randomization/unit of assignment, model specifications, and LLM prompt/architecture transparency, which limit confidence in replication and external validity. SampleTwo experiments with N1=54 and N2=54 human participants performing simulated information-search decision tasks under three experimentally controlled uncertainty environments; participants experienced one of three feedback conditions (SERA-Gist, SERA-Verbatim, no-feedback); outcomes measured: decision accuracy, confidence, and sampling behavior (amount and pattern of evidence sampled). (Recruitment source and demographics not specified in summary.) Themeshuman_ai_collab productivity IdentificationRandomized controlled experiments in which participants were assigned to one of three feedback conditions (SERA-Gist, SERA-Verbatim, no-feedback) while performing information-search tasks in experimentally manipulated uncertainty environments (decremental/local optimum/no-pattern). Causal effects are identified via random assignment and experimental control over feedback and environmental uncertainty. GeneralizabilitySmall, convenience-style samples limit population representativeness (demographics not reported)., Laboratory/online simulated tasks may not capture complexity of real-world, high-stakes decision contexts., Results may depend on specific LLM, prompts, and feedback implementations used; other models or UIs could differ., Only three stylized uncertainty environments considered; real-world information landscapes may have richer structure., Short-term experiments do not speak to longer-run learning, habituation, or organizational adoption dynamics.

Claims (7)

ClaimDirectionOutcomeConfidence & EvidenceDetails
We introduce SERA, an LLM-based assistant that provides either gist or verbatim feedback during search. Other null_result description of system design (feedback modality: gist vs verbatim)
Reading fidelity high
Study strength high
not reported
1.0
Across two experiments (N1=54, N2=54), we examined decision-making outcomes and information search in SERA-Gist, SERA-Verbatim, and a no-feedback baseline across three environments varying in uncertainty. Other null_result experimental comparison of decision-making outcomes and information search across conditions
Reading fidelity high
Study strength high
n=108
1.0
The uncertainty in environment is operationalized by the perceived gain of information across the course of sampling: decremental (low-uncertainty), local drop/local optimum (medium-uncertainty), or no patterns (high-uncertainty). Other null_result operationalization/manipulation of environmental uncertainty (patterns of information gain)
Reading fidelity high
Study strength high
n=108
1.0
Individuals show more accurate decision outcomes and are more confident with SERA support, especially under higher uncertainty. Decision Quality positive decision accuracy and self-reported confidence
Reading fidelity high
Study strength medium
n=108
0.6
Gist feedback was associated with more efficient integration and showed a descriptive pattern of reduced oversampling. Task Allocation positive sampling efficiency / extent of oversampling
Reading fidelity high
Study strength speculative
n=108
0.1
Verbatim feedback promoted more extensive exploration. Task Allocation positive extent of exploration / amount of sampling
Reading fidelity high
Study strength speculative
n=108
0.1
These findings establish feedback representation as a design lever when search matters, motivating adaptive systems that match feedback granularity to uncertainty. Organizational Efficiency positive implication for system design (matching feedback granularity to environmental uncertainty)
Reading fidelity high
Study strength speculative
n=108
0.1

Notes