The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

An LLM-based planner, SynthEx, designs convergent, strategy-driven syntheses for over a thousand complex natural products using a template-free reaction representation, and expert chemists judge its key steps on par with published human routes; without wet-lab tests, feasibility remains to be proven experimentally.

Strategy-first synthesis planning for complex natural products
Daniel Armstrong, Xuan-Vu Nguyen, Octavian Susanu, Gabriel Gibberd, Théo A. Neukomm, Taddäus Strunden, Dan Forster, Morgane Delattre, Shawn Teh, Clément Rols, John Federice, Hayden Leatherwood, M. Lavelle Barnes, Maarten R. Dobbelaere, Peter Wipf, Jon T. Njardarson, Jieping Zhu, Philippe Schwaller · August 07, 2026
arxiv descriptive medium evidence 7/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Daniel Armstrong unresolved corpus identity
  2. Xuan-Vu Nguyen unresolved corpus identity
  3. Octavian Susanu unresolved corpus identity
  4. Gabriel Gibberd unresolved corpus identity
  5. Théo A. Neukomm unresolved corpus identity
  6. Taddäus Strunden unresolved corpus identity
  7. Dan Forster unresolved corpus identity
  8. Morgane Delattre unresolved corpus identity
  9. Shawn Teh unresolved corpus identity
  10. Clément Rols unresolved corpus identity
  11. John Federice unresolved corpus identity
  12. Hayden Leatherwood unresolved corpus identity
  13. M. Lavelle Barnes unresolved corpus identity
  14. Maarten R. Dobbelaere unresolved corpus identity
  15. Peter Wipf unresolved corpus identity
  16. Jon T. Njardarson unresolved corpus identity
  17. Jieping Zhu unresolved corpus identity
  18. Philippe Schwaller unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Daniel P. Armstrong provider ID
  2. X. Nguyen provider ID
  3. Octavian Susanu provider ID
  4. G. Gibberd provider ID
  5. Théo A. Neukomm provider ID
  6. Taddäus E N Strunden provider ID
  7. Dan Forster provider ID
  8. Morgane Delattre provider ID
  9. Shawn Teh provider ID
  10. Clément Rols provider ID
  11. John G Federice provider ID
  12. Hayden Leatherwood provider ID
  13. M. L. Barnes provider ID
  14. Maarten R. Dobbelaere provider ID
  15. Peter Wipf provider ID
  16. J. T. Njardarson provider ID
  17. Jieping Zhu provider ID
  18. P. Schwaller provider ID
SynthEx is an LLM-driven, agentic synthesis planner using a template-free ReactionJSON representation that proposes convergent, often novel synthetic routes for complex natural products which expert chemists rate comparable to published human syntheses.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

The total synthesis of a complex molecule is among the most demanding intellectual and experimental feats in chemistry: a chemist must plan many steps ahead for how to assemble simple building blocks into an intricate target, devise backup strategies, and anticipate procedural challenges. It is also a profoundly creative activity. For half a century, efforts to automate the retrosynthetic design of natural products and other complex molecules have drawn on catalogued reactions, and the resulting tools now report near-complete success on benchmarks built from that same source. But these tools were shaped to fit benchmarked chemistry, and they falter on many natural products, the frontier of the field, whose densely functionalized, polycyclic architectures demand precisely the inventive chemistry the record contains least. Whether a machine could reasonably design such syntheses like an expert chemist does has remained unclear. Here, we show that SynthEx, an agentic framework built on large language models, plans routes to complex natural products that lie beyond the reach of conventional design algorithms. SynthEx proposes competing strategies, assembles a sequence of routine and key steps into a cohesive route, and critiques and improves its own design; the chemistry it favours is more convergent than existing tools produce, and spans a region of reaction space that catalogue-based tools cannot match. Most notably, in blinded assessments, expert chemists judged its key steps comparable to those of published human syntheses and engaged with them as genuine synthesis plans, a response algorithmic route prediction has not previously accomplished. We release routes to more than a thousand natural products as SynthAtlas, an open, interactive database, and anticipate it will become a shared resource for a collection of complex target molecules that lack existing literature routes.

Summary

Main Finding

SynthEx, an agentic multi‑agent planner built on large language models and a novel template‑free reaction representation (ReactionJSON), can generate strategic, multi‑step synthetic routes for complex natural products that lie beyond the reach of conventional template‑ or catalogue‑based CASP systems. Its routes are more convergent, occupy a qualitatively distinct reaction space (more bond‑forming ring constructions and cross‑couplings), and—in blinded expert reviews—its key disconnections were judged comparable to human‑published syntheses. The routes and analyses are released as SynthAtlas (1,098 targets; 3,243 routes; 33,145 reactions).

Key Points

  • Architecture and workflow
    • SynthEx is a multi‑agent pipeline of LLM agents: Strategy Generator (proposes competing high‑level strategies keyed on key disconnections), Route Builder (generates full routes using ReactionJSON), Critic (simulates forward reactions and flags infeasible steps), Editor (repairs routes while preserving strategy), and Analyst (scores feasibility, identifies risks/key steps).
    • Emphasis on strategy generation before committing to route search—mirrors human chemists entertaining multiple hypotheses.
  • ReactionJSON
    • Template‑free reaction representation: ordered lists of atom‑level graph edits mapping product → precursors.
    • Enables the LLM to invent/instantiate novel single‑step transformations rather than selecting from a fixed template library; allows deterministic application and in‑place surgical edits without redoing full searches.
  • Performance and qualitative behavior
    • SynthEx opens reaction space that template‑based planners cannot: proposes chemistry common in literature/total‑synthesis knowledge but rare or absent in reaction databases.
    • Produces more convergent syntheses (commonly unifying two independent fragments) and favors constructive bond formations, ring closures, and cascade reactions over simple functional‑group manipulations.
    • Advantage over template approaches increases with target complexity (more rings, stereochemical density).
  • Validation and outputs
    • Benchmarked on NPAtlas-derived set of 1,098 natural‑product targets; produced 3,243 strategy‑annotated routes and 33,145 reactions (up to 3 distinct routes per target).
    • Blinded expert chemist assessments rated SynthEx key steps on par with published human syntheses; highlighted several elegant, mechanistically complex disconnections (e.g., Grob fragmentation cascades, [3+2] dipolar cycloadditions, intramolecular cascade Michael additions).
    • Case studies: SynthEx recovered a literature route published after the model cutoff (Okaramine M), proposed an alternative feasible plan for Melonine, and suggested an elegant Hofmann–Löffler–Freytag disconnection for Chanoclavine→Lysergol.
  • Scope and limitations
    • Focused on “reach, convergence, distinctiveness, and strategy.” Wet‑lab feasibility and experimental validation remain future work; the authors treat feasibility scoring as analytic rather than proof of experimental success.
    • Because the system generates atom‑level edits, it can hallucinate chemically implausible steps; the Critic/Editor loop mitigates but does not eliminate this risk—human validation remains necessary.

Data & Methods

  • Dataset and outputs
    • Targets: 1,098 natural products (from NPAtlas targets not flagged as already synthesized).
    • Outputs: 3,243 strategy‑annotated routes, 33,145 total reaction steps; publicly released as SynthAtlas (interactive).
  • Core methods
    • LLM agents orchestrated in stages: Strategy Generator (multiple strategies per target), Route Builder (creates ReactionJSON for each disconnection), Critic (forward‑simulate and flag infeasible steps), Editor (surgically edit routes without abandoning strategies), Analyst (feasibility scoring, risk/key‑step annotation).
    • ReactionJSON: atom‑mapped, ordered graph edits (add/remove bonds, change atom attributes) that deterministically produce precursors from products when applied; allows fine‑grained edits and route repair.
  • Baselines and comparisons
    • Compared qualitatively and quantitatively to state‑of‑the‑art template/template‑library and trained retrosynthesis planners (which rely on reaction templates or neural policies trained on reaction corpora).
    • Key empirical findings: SynthEx solves many targets that leading template‑based planners fail on; its advantage grows with molecular complexity.
  • Human evaluation
    • Blinded assessments by expert synthetic chemists evaluated key steps vs. published human syntheses; results showed comparable ratings and highlighted several creative, experimentally interesting disconnections.
  • Reproducibility & release
    • Routes and analyses released as SynthAtlas for community inspection, comment, and reuse. The methodology emphasizes interpretability (text descriptions, annotated conditions, strategy explanations).

Implications for AI Economics

  • R&D productivity and value creation
    • Potential to materially accelerate early‑stage route ideation in natural product synthesis, medicinal chemistry, and complex molecule discovery by lowering the time/brainpower cost of high‑quality strategic planning.
    • Greater reach into previously “hard” chemical space may unlock new molecules and intermediates, altering discovery pipelines (faster hypothesis generation, expanded chemical ideas per project).
  • Labor and skill‑composition effects
    • Could shift the division of labor: routine strategic ideation may be automated, increasing demand for chemists who can validate, adapt, and experimentally implement AI‑proposed routes (higher‑skill experimentalists, validation specialists).
    • May reduce the marginal value of traditional retrosynthesis expertise for ideation while increasing demand for interdisciplinary operators (chemists + AI toolchain curators).
  • Market structure and competition
    • Open release (SynthAtlas) lowers entry barriers for academic labs, startups, and smaller firms to access high‑quality strategic planning, potentially decentralizing innovation in complex synthesis.
    • Commercial CASP vendors may be pressured to integrate LLM‑driven, template‑free planning and iterative critique/repair; differentiation may shift to experimental validation, ease‑of‑integration, wet‑lab automation, and IP.
  • Investment and capitalization
    • Tools that increase ideation throughput without proportionate increases in experimental cost can raise the rate of project initiation and broaden portfolios—changing how firms allocate R&D capital and portfolio risk.
    • Conversely, if many projects become easier to propose, experimental validation becomes the bottleneck—raising capital needs for lab resources, automation, and scale‑up.
  • Data, IP, and incentives
    • Because SynthEx leverages broadly learned literature knowledge and releases routes openly, tensions may arise over intellectual property (prior art, patent landscapes) when AI‑proposed routes overlap with unpublished or proprietary methods.
    • Open atlases like SynthAtlas may change the public goods landscape—facilitating cumulative innovation but creating new incentives around curation, commercialization of validated routes, and data‑centric competitive advantages.
  • Benchmarks and measurement
    • The paper illustrates a shift in evaluation metrics for AI systems: saturated retrieval benchmarks no longer probe meaningful progress; economic value derives from out‑of‑distribution reasoning and practical utility (e.g., wet‑lab success rates, time‑to‑prototype).
    • For economics, this suggests evaluation frameworks should incorporate downstream productivity measures (experiments saved, time to viable lead, cost per validated route).
  • Risks and caveats with economic impact
    • Current limitations (lack of guaranteed wet‑lab feasibility) mean that the economic upside is conditional on effective human/experimental validation pipelines.
    • Misestimation of feasibility could lead to wasted experiments and false signals of productivity; firms will need processes for triaging AI proposals.
    • Concentration risk: labs or companies that pair such planners with strong experimental/automation assets could capture disproportionate returns.
  • Policy and workforce implications
    • Education and training should emphasize AI‑assisted experimental design, critical validation skills, and integration of computational and lab workflows.
    • Policymakers and funders may consider supporting public validation efforts (e.g., crowdsourced experimental testing of AI‑proposed routes) to realize broader social returns from open resources.

Summary takeaway for AI economics: SynthEx demonstrates that LLM‑driven, strategy‑first, template‑free planning can expand the productive frontier of synthetic ideation. The main economic effects will depend on how readily the community converts these strategic proposals into validated, scalable experimental outcomes—shifting value from ideation to validation, changing labor demand toward experimental and AI‑integration skills, and reshaping market and IP dynamics through open shared resources.

Assessment

Paper Typedescriptive Evidence Strengthmedium — The paper presents large-scale computational experiments (1,098 NP targets, 3,243 routes, 33,145 steps) and blinded expert evaluations that support its claims, and it recovers at least one post-training literature route; however, no wet-lab experimental validation of proposed routes or measured productivity gains is provided, and some evaluations rely on expert judgment rather than objective experimental outcomes. Methods Rigorhigh — The authors introduce a clear, novel representation (ReactionJSON), a multi-agent LLM architecture that separates strategy generation, route building, criticism and editing, and perform broad benchmark comparisons versus template-based planners plus blinded human expert assessment; they also release an atlas of routes. Limitations include lack of laboratory validation of reaction outcomes and potential dependence on LLM training data/biases. SampleComputational application of SynthEx to 1,098 natural-product targets (drawn from NPAtlas), producing 3,243 distinct synthetic strategies and 33,145 reaction steps; comparisons made against a leading template-based retrosynthesis planner and single-step models; blinded expert chemist assessments of key steps and routes; case studies including recovery of a route published after the model training cut-off and other novel proposals. The system uses LLM agents (Strategy Generator, Route Builder producing ReactionJSON, Critic, Editor, Analyst) but no wet-lab reaction execution data are reported. Themesproductivity human_ai_collab GeneralizabilityNo experimental (wet-lab) validation of predicted routes or yields, so chemical feasibility in practice is uncertain, Expert blinded evaluation is informative but subjective and may not capture experimental difficulty, yields, or scale-up issues, Performance likely depends on LLM pretraining data and may vary with different model families or updates, Benchmarks focus on natural products; results may not fully generalize to other domains of synthetic chemistry (e.g., process or industrial scale synthesis), Reaction condition details are abstracted; absence of reliable condition prediction limits immediate lab adoption

Claims (11)

ClaimDirectionOutcomeConfidence & EvidenceDetails
SynthEx plans routes for the large majority of the natural-product targets for which a leading template-based planner solves only a small fraction. Research Productivity positive Reach or solve rate of synthesis-planning systems across complex natural-product targets
Reading fidelity high
Study strength medium
n=1098
0.18
SynthEx's advantage over conventional synthesis planners widens as target molecules become more complex. Research Productivity positive Synthesis-planning performance as molecular complexity increases
Reading fidelity high
Study strength medium
n=1098
0.18
SynthEx proposes more convergent bond constructions than existing synthesis-planning tools, with most such constructions uniting two independent fragments. Research Productivity positive Convergence of proposed synthetic routes
Reading fidelity high
Study strength medium
n=1098
0.18
SynthEx generates routes occupying a distinct region of reaction space that state-of-the-art single-step models rarely propose. Innovation Output positive Distinctiveness and diversity of proposed chemical transformations
Reading fidelity high
Study strength medium
n=1098
0.18
Compared with template-based output, SynthEx favors constructive bond-forming steps, ring constructions, and cross-couplings over functional-group manipulations. Innovation Output positive Strategic composition of proposed synthetic routes
Reading fidelity high
Study strength medium
n=1098
0.18
In blinded assessments, expert chemists rated SynthEx's key steps as comparable to those in published human syntheses. Output Quality positive Expert-assessed quality of key synthetic steps
Reading fidelity high
Study strength low
not reported
0.09
Expert chemists identified some SynthEx-generated disconnections as genuinely elegant, including Grob fragmentation, [3+2] dipolar cycloaddition, and intramolecular cascade Michael additions. Creativity positive Expert assessment of elegance and strategic quality of proposed disconnections
Reading fidelity high
Study strength low
not reported
0.09
SynthEx independently recovered a route for Okaramine M that was published after the model's training cutoff and that route had been experimentally validated by expert organic chemists. Research Productivity positive Agreement with and recovery of an experimentally validated expert synthesis route
Reading fidelity high
Study strength medium
n=1
0.18
SynthEx produced an alternative route to Melonine that did not match the later published total synthesis but was judged by an author with prior experience on the target to be feasible and worthy of experimental validation. Output Quality mixed Expert-assessed feasibility of a novel alternative synthetic route
Reading fidelity high
Study strength speculative
n=1
0.03
SynthAtlas contains routes for 1,098 natural-product targets, comprising 3,243 synthetic strategies and 33,145 reactions. Research Productivity positive Number of generated synthesis routes and reaction steps made available as a resource
Reading fidelity high
Study strength high
n=1098
1,098 targets; 3,243 strategies; 33,145 reactions
0.3
The paper does not establish wet-lab feasibility for the generated routes; it identifies experimental validation as a next frontier. Output Quality null_result Experimental or wet-laboratory validation of proposed synthetic routes
Reading fidelity high
Study strength high
n=1098
0.3

Notes