0 cumulative citations
View corpus contextA lightweight LLM assistant that gives gist-level feedback improves decision accuracy and confidence under uncertainty and curbs unnecessary oversampling; verbatim feedback, by contrast, drives more extensive exploration, suggesting designers should match feedback granularity to the uncertainty profile of the task.
Citation observations
Cumulative provider counts captured on specific dates; providers are never combined.
Many real-world decisions rely on information search, where people sample evidence and decide when to stop under uncertainty. The uncertainty in the environment, particularly how diagnostic evidence is distributed, causes complexities in information search, further leading to suboptimal decision-making outcomes. Yet AI decision support often targets outcome optimization, and less is known about how to scaffold search without increasing cognitive load. We introduce SERA, an LLM-based assistant that provides either gist or verbatim feedback during search. Across two experiments (N1=54, N2=54), we examined decision-making outcomes and information search in SERA-Gist, SERA-Verbatim, and a no-feedback baseline across three environments varying in uncertainty. The uncertainty in environment is operationalized by the perceived gain of information across the course of sampling, which individuals may experience diminishing return of information gain (decremental; low-uncertainty), or a local drop of information gain (local optimum; medium-uncertainty), or no patterns in information gain (high-uncertainty), as they search more. Individuals show more accurate decision outcomes and are more confident with SERA support, especially under higher uncertainty. Gist feedback was associated with more efficient integration and showed a descriptive pattern of reduced oversampling, while verbatim feedback promoted more extensive exploration. These findings establish feedback representation as a design lever when search matters, motivating adaptive systems that match feedback granularity to uncertainty.
Summary
Main Finding
An LLM-based conversational assistant (SERA) that provides on-demand summaries during sequential information search improves decision accuracy and confidence—particularly in high-uncertainty environments. The format of feedback matters: gist (high-level) summaries encourage more efficient integration and reduced oversampling, while verbatim (detail-preserving) summaries encourage more extensive exploration. These effects imply that matching feedback granularity to the environment’s uncertainty structure can improve exploration–exploitation regulation.
Key Points
- Problem framed: many real-world choices require sequential information search and stopping under uncertainty; AI support typically targets final outcomes but may better serve users by scaffolding the search process.
- SERA: a conversational, LLM-driven assistant that (a) summarizes user-recorded notes on demand and (b) prompts brief monitoring (self-regulatory) reflections.
- Two feedback formats tested:
- Gist: compressed, high-level meanings emphasizing what matters now.
- Verbatim: literal, detailed recaps preserving attribute-level content.
- Environments (operationalized by how perceived information gain evolves during sampling):
- Low-uncertainty (Decremental): predictable, steadily diminishing information gain.
- Medium-uncertainty (Local Optimum): a local drop then rise in information gain (risk of stopping too early).
- High-uncertainty (Random): no predictable pattern in information gain.
- Main behavioral patterns:
- SERA (either format) → higher decision accuracy and greater confidence than no-feedback, with larger gains in higher-uncertainty settings.
- Gist → more efficient integration, descriptive reduction in oversampling (i.e., fewer redundant samples; better stopping).
- Verbatim → greater exploration (more sampling), which can help avoid local optima but may increase cognitive cost.
- Design implication emphasized by authors: feedback representation is a manipulable design lever; adaptive interfaces should match summary granularity to environmental predictability.
Data & Methods
- Overall design: two experiments (N1 = 54, N2 = 54) using a web-based search-and-stop decision task adapted from decisions-from-experience paradigms.
- Conditions: SERA-Gist, SERA-Verbatim, and a no-feedback baseline crossed with three uncertainty environments (decremental, local-optimum, random).
- Interaction flow: participants click to reveal sequential descriptive information for two options, record notes, request summaries from SERA, respond to a monitoring prompt, then continue sampling or stop and choose.
- SERA implementation:
- Frontend: React.js; backend: Flask; data logged to Firebase.
- LLM: OpenAI gpt-4o-mini; prompt framed the model as a neutral assistant with concise outputs. Parameters reported (temperature = 1, max tokens = 150, top_p = 0).
- Summaries generated from participant-entered notes; format (gist vs verbatim) controlled via prompt framing.
- Outcomes measured:
- Outcome effectiveness: decision accuracy, decision confidence.
- Process regulation: number of samples, stopping behavior, integration patterns (efficiency vs oversampling), exploratory behavior.
- Secondary: user perceptions and individual differences (to assess reliance patterns).
- Main results (reported descriptively in manuscript):
- SERA improves accuracy and confidence relative to baseline; effect sizes increase with environmental uncertainty.
- Gist reduces redundant sampling and supports efficient stopping; verbatim increases exploration.
Implications for AI Economics
- Design of decision-support tools in economic settings should go beyond final recommendations to actively scaffold the information search process; the representation of feedback materially affects exploration costs and welfare.
- Match feedback granularity to environment structure:
- Predictable/diminishing-return environments (e.g., routine product comparisons, well-known financial instruments): favor gist summaries to reduce sampling costs, speed decisions, and reduce overload.
- Environments with local optima or sparse but informative late signals (e.g., complex investment research, procurement with hidden features): favor verbatim or detail-rich summaries to encourage deeper exploration and avoid premature stopping.
- Highly unpredictable markets: SERA-style prompts and summaries increase confidence and accuracy overall; a hybrid/adaptive policy that responds to observed information-gain patterns can balance exploration costs and decision quality.
- Platform and market applications:
- E‑commerce search: gist summaries can lower search costs and abandonment; verbatim summaries can be used in expert-aimed workflows where exhaustive evaluation matters.
- Financial advising and consumer finance: adaptive feedback could reduce under-search or overreliance and improve calibration of confidence in costly uncertainty.
- Labor markets and procurement: conversational summaries can speed screening while verbatim support can be enabled when the environment indicates the risk of missing late-arriving important signals.
- Policy and welfare considerations:
- Designers must consider cognitive costs and risk of overreliance; fluent conversational summaries may be persuasive and increase confidence—even when model outputs are uncertain—so transparency about system limits remains important.
- Adaptive feedback that detects the payoff-structure and switches granularity can improve aggregate welfare but requires careful evaluation to avoid nudging users toward prematurely low-effort choices.
- Research & modeling opportunities for AI economics:
- Formalize the value of information under different information-gain temporal structures and compute optimal feedback policies (when to provide gist vs verbatim).
- Integrate user search-cost heterogeneity and attention limits into market-design models to predict impacts of different feedback strategies on social welfare, search duration, and market liquidity.
- Evaluate long-run behavioral effects (habit formation, calibration of trust, exploitation of persuasive summaries) and the strategic consequences for platforms that choose different feedback policies.
Limitations to note for economic application: relatively small experimental Ns, lab-style descriptive information (not monetary gambles), single LLM and prompt family, and potential domain differences; results should be validated in larger, domain-specific field settings before widescale deployment.
Assessment
Claims (7)
| Claim | Direction | Outcome | Confidence & Evidence | Details |
|---|---|---|---|---|
| We introduce SERA, an LLM-based assistant that provides either gist or verbatim feedback during search. Other | null_result | description of system design (feedback modality: gist vs verbatim) |
Reading fidelity
high
Study strength
high
|
not reported
|
| Across two experiments (N1=54, N2=54), we examined decision-making outcomes and information search in SERA-Gist, SERA-Verbatim, and a no-feedback baseline across three environments varying in uncertainty. Other | null_result | experimental comparison of decision-making outcomes and information search across conditions |
Reading fidelity
high
Study strength
high
|
n=108
|
| The uncertainty in environment is operationalized by the perceived gain of information across the course of sampling: decremental (low-uncertainty), local drop/local optimum (medium-uncertainty), or no patterns (high-uncertainty). Other | null_result | operationalization/manipulation of environmental uncertainty (patterns of information gain) |
Reading fidelity
high
Study strength
high
|
n=108
|
| Individuals show more accurate decision outcomes and are more confident with SERA support, especially under higher uncertainty. Decision Quality | positive | decision accuracy and self-reported confidence |
Reading fidelity
high
Study strength
medium
|
n=108
|
| Gist feedback was associated with more efficient integration and showed a descriptive pattern of reduced oversampling. Task Allocation | positive | sampling efficiency / extent of oversampling |
Reading fidelity
high
Study strength
speculative
|
n=108
|
| Verbatim feedback promoted more extensive exploration. Task Allocation | positive | extent of exploration / amount of sampling |
Reading fidelity
high
Study strength
speculative
|
n=108
|
| These findings establish feedback representation as a design lever when search matters, motivating adaptive systems that match feedback granularity to uncertainty. Organizational Efficiency | positive | implication for system design (matching feedback granularity to environmental uncertainty) |
Reading fidelity
high
Study strength
speculative
|
n=108
|