The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

Designers testing a multi-agent creative tool quickly abandon full agent autonomy and instead orchestrate multiple AIs themselves, suggesting humans prefer hands-on control when coordinating several AI roles.

Understanding Human-Multi-Agent Team Formation for Creative Work
Hyunseung Lim, Dasom Choi, Sooyohn Nam, Bogoan Kim, Hwajung Hong · January 20, 2026
arxiv descriptive low evidence 7/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Hyunseung Lim unresolved corpus identity
  2. Dasom Choi unresolved corpus identity
  3. Sooyohn Nam unresolved corpus identity
  4. Bogoan Kim unresolved corpus identity
  5. Hwajung Hong unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Hyunseung Lim provider ID
  2. Dasom Choi provider ID
  3. Sooyohn Nam provider ID
  4. Bogoan Kim provider ID
  5. Hwajung Hong provider ID
In an exploratory study with 12 designers using the CrafTeam probe, participants shifted from expecting autonomous multi-agent operation to preferring direct human orchestration of multiple AI agents during creative ideation.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

Team-based collaboration is a cornerstone of modern creative work. Recent advances in generative AI open possibilities for humans to collaborate with multiple AI agents in distinct roles to address complex creative workflows. Yet, how to form Human-Multi-Agent Teams (HMATs) is underexplored, especially given that inter-agent interactions increase complexity and the risk of unexpected behaviors. In this exploratory study, we aim to understand how to form HMATs for creative work using CrafTeam, a technology probe that allows users to form and collaborate with their teams. We conducted a study with 12 design practitioners, in which participants iterated through a three-step cycle: forming HMATs, ideating with their teams, and reflecting on their teams' ideation. Our findings reveal that while participants initially attempted autonomous team operations, they ultimately adopted team formations in which they directly orchestrated agents. We discuss design considerations for HMAT formation that humans can effectively orchestrate multiple agents.

Summary

Main Finding

When people form teams that include multiple AI agents for creative work, they initially try to let agents operate autonomously, but quickly converge on formations where the human explicitly orchestrates agents. Autonomous inter-agent interactions often produce unproductive loops or misaligned behavior; human-directed orchestration (clear direction, role boundaries, and oversight) improves ideation outcomes and manageability.

Key Points

  • Problem space
    • Human–Multi-Agent Teams (HMATs) combine one human with two or more autonomous agents; team formation (size, structure, roles, composition, shared mental models) is critical and more complex than single-agent HATs.
    • Inter-agent interactions introduce novel coordination risk: unexpected loops, lack of value judgments, and drift from creative intent.
  • CrafTeam probe
    • A web-based technology probe that lets non-developers specify five formation dimensions and then run iterative ideation cycles with HMATs.
    • Five configurable dimensions: team size, team structure (hierarchical/peer links), role allocation (e.g., Idea Generation, Idea Evaluation, Feedback, Request), member composition (personas/profiles), and shared mental models (explicitized goals/constraints).
  • Study and behavioral findings
    • Participants: 12 design practitioners from IT companies; three-hour sessions; repeated three cycles of forming → ideating → reflecting.
    • Participants iteratively refined HMATs. Early trials favored agent autonomy; subsequent trials shifted to human-orchestration after observing:
      • Agents getting stuck in unproductive back-and-forths.
      • Agents failing to make evaluative/value judgments needed to move creative decisions forward.
      • Difficulty tracing which formation choices produced which outcomes without explicit human control.
    • Participants preferred configurations where the human set direction, orchestrated agents actively, and used agents for complementary tasks (diverse perspectives, drafting, critique) rather than full autonomy.
  • Design takeaways (from participant behavior)
    • Interfaces should enable direct orchestration: high-level directives, authority controls, and the ability to route or sequence agent actions.
    • Explicit role boundaries and visible inter-agent communication reduce ambiguity and emergent loops.
    • Support building shared mental models (shared goals, constraints, examples) and make agent capabilities transparent.
    • Reflection tools (transcripts, interaction visualizations) help users link formation choices to outcomes and iterate formations.

Data & Methods

  • System: CrafTeam, a technology probe implementing multi-agent ideation based on a human-AI co-ideation framework (Shen et al.) with pre-defined agent role types. Users configure team formation and then run collaborative ideation sessions; post-session reflection interfaces show transcripts and interaction patterns.
  • Participants: 12 professional design practitioners working in IT companies.
  • Procedure:
    • Three-hour lab studies per participant.
    • Iterative three-cycle protocol: (1) form HMAT (configure five dimensions), (2) ideate with HMAT, (3) reflect on ideation (review transcripts, decide changes), then repeat (three rounds).
    • Data collected: session transcripts, configuration logs, recorded interactions, and post-study interviews.
  • Analysis: Qualitative analysis of participant behavior, formation changes, and interview data to distill patterns and design considerations. (Implementation/appendix details provided in paper.)
  • Limitations:
    • Small N (12) and domain-limited (design practitioners); exploratory/qualitative rather than statistically generalizable.
    • Probe settings and pre-defined agent role choices influenced the space of possible formations.

Implications for AI Economics

  • Labor complementarities and skill-bundling
    • Human orchestration (coordination, setting objectives, integrating outputs) emerges as a valuable complementary skill to agent capabilities. Demand may rise for workers who can design, manage, and orchestrate HMATs.
    • Task decomposition: routine or specialized sub-tasks may be delegated to agents, while higher-level integrative/ evaluative tasks remain human-heavy—altering the division of labor and potentially increasing non-routine cognitive premium.
  • Productivity and organizational design
    • HMATs can increase productivity in creative workflows only when firms invest in orchestration interfaces, governance, and training. Returns to adopting HMATs hinge on reducing coordination failures (unproductive agent loops) and on human orchestrator effectiveness.
    • Firms that standardize orchestration roles and team-formation protocols may achieve efficiency advantages, leading to scale economies in creative output and possibly winner-take-most dynamics in creative sectors.
  • Cost structures and transaction costs
    • Multi-agent systems change transaction costs: lower marginal cost of “hiring” specialized agent roles but higher coordination/monitoring costs. Investments in tooling (visualization, SMM-support, audit trails) are essential to lower these coordination costs.
    • Pricing of agent services and platforms may reflect not only agent output quality but also the quality of orchestration tools and the ease of integrating agents into human workflows.
  • Labor market and skill-biased demand
    • A premium may develop for “orchestrator” skills (meta-design, cross-agent governance, judgment under ambiguity). Training and credentialing markets could expand for these coordination competencies.
    • Some creative roles may be reshaped rather than replaced: workers move toward supervision, curation, and synthesis roles while agents handle generation and routine critique.
  • Platform and market structure
    • Platforms offering plug-and-play agent modules (specialized personas/roles) plus orchestration UX could capture high value. Vendor differentiation will depend on the richness of role libraries and orchestration affordances.
    • Network effects: firms that accumulate orchestration expertise, templates, and shared mental models may drive lock-in and higher returns.
  • Policy and risk considerations
    • Misalignment and emergent behaviors in inter-agent interactions create operational risks (quality, reputational, compliance). Economic adoption requires governance (auditability, liability allocation) which affects implementation costs and regulation.
  • Empirical research directions
    • Quantify productivity gains from orchestration vs. agent autonomy across tasks and industries.
    • Measure wage and demand shifts for orchestration roles; estimate complementarity magnitude between human orchestration skill and agent capability.
    • Model firm-level returns to investing in orchestration infrastructure and the resulting market concentration dynamics.

Overall, the paper suggests that the value created by multi-agent AI in creative work depends less on scaling agent autonomy and more on designing institutions, interfaces, and human skills to orchestrate agents effectively—an insight with direct implications for labor demand, firm investment, and market structure in the AI economy.

Assessment

Paper Typedescriptive Evidence Strengthlow — Small, exploratory qualitative study (N=12) with no counterfactual or causal design, no objective productivity/wage/firm-level outcomes, and likely self-selection and lab/artificial task effects, so findings are suggestive but not strong causal evidence. Methods Rigormedium — The study uses a structured technology-probe and iterative three-step cycle (form, ideate, reflect) with domain practitioners, which is appropriate for early-stage HCI research; however, the small sample, likely convenience sampling, lack of quantitative validation, and limited reporting of analytic procedures (e.g., coding, interrater reliability) constrain methodological rigor. Sample12 design practitioners participated in an exploratory lab/field study using CrafTeam (a technology probe) to form and collaborate with human-multi-agent teams across iterative cycles of team formation, ideation, and reflection. Themeshuman_ai_collab org_design innovation GeneralizabilitySmall sample (N=12) limits statistical generalizability, Participants were design practitioners only — results may not apply to non-creative or non-design occupations, Short-term interactions with a prototype probe — may not reflect long-term use or organizational deployment, Likely convenience or purposive sampling with limited geographic/sectoral diversity, Task and interface constraints of the CrafTeam probe may bias behaviors compared with other multi-agent systems

Claims (8)

ClaimDirectionOutcomeConfidence & EvidenceDetails
Team-based collaboration is a cornerstone of modern creative work. Creativity positive importance/centrality of team-based collaboration in creative work
Reading fidelity high
Study strength speculative
not reported
0.03
Recent advances in generative AI open possibilities for humans to collaborate with multiple AI agents in distinct roles to address complex creative workflows. Creativity positive feasibility/potential for multi-agent human-AI collaboration in creative workflows
Reading fidelity high
Study strength speculative
not reported
0.03
How to form Human-Multi-Agent Teams (HMATs) is underexplored. Other negative extent of prior research on HMAT formation
Reading fidelity high
Study strength low
not reported
0.09
Inter-agent interactions increase complexity and the risk of unexpected behaviors. Error Rate negative complexity of interactions and risk of unexpected behaviors among multiple agents
Reading fidelity high
Study strength speculative
not reported
0.03
We conducted a study with 12 design practitioners who iterated through a three-step cycle: forming HMATs, ideating with their teams, and reflecting on their teams' ideation. Other null_result process of forming and using HMATs (formation → ideation → reflection)
Reading fidelity high
Study strength high
n=12
0.3
Participants initially attempted autonomous team operations but ultimately adopted team formations in which they directly orchestrated agents. Task Allocation mixed preferred mode of team operation (autonomous vs human-orchestrated)
Reading fidelity high
Study strength medium
n=12
0.18
CrafTeam is a technology probe that allows users to form and collaborate with their teams. Other positive ability to form and collaborate with HMATs using the CrafTeam probe
Reading fidelity high
Study strength high
n=12
0.3
The paper presents design considerations for HMAT formation that enable humans to effectively orchestrate multiple agents. Organizational Efficiency positive design guidance for effective human orchestration of multiple agents
Reading fidelity high
Study strength low
n=12
0.09

Notes