The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests About 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

A new governance protocol shows standard ROI analyses can overlook systemic risks from automation and would often automate roles that a systemic-risk-aware test recommends preserving or hybridising; a formal 'automation-debt' metric and five-gate audit produce role-level decisions that remain robust to moderate threshold changes.

When Not to Automate: A Formal Protocol for Human Preservation in AI-Optimized Organizations
Jose Manuel de la Chica Rodriguez, Jairo Rodriguez Arias, Spyridon Chouliaras · July 17, 2026
arxiv theoretical n/a evidence 7/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Semantic Scholar

Latest observation:

  1. J. Rodriguez provider ID
  2. Jairo Rodriguez Arias provider ID
  3. Spyridon Chouliaras provider ID
The paper proposes PHP-AIO, a five-gate decision protocol and a closed-form automation-debt metric ρ(P) that quantify unpriced systemic risks of automation at the role level and produce auditable automate/augment/hybrid/preserve decisions that diverge from standard ROI-based recommendations.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

Standard automation ROI misses four categories of systemic risk -- tacit knowledge erosion, resilience reduction, regulatory exposure, and socio-institutional capital degradation -- that affect long-term organizational performance. PHP-AIO (Protocol for Human Preservation in AI-Optimized Organizations) is a five-gate sequential decision protocol with a final composite check that quantifies these unpriced systemic risks at the role level and produces auditable automation decisions. A closed-form automation-debt measure ($ρ(P)$) formalises how role-level decisions accumulate across multi-step processes; its warning is neutralised only by a regulator-mandated human-in-the-loop anchor. Applied to stylised profiles of representative internal roles, PHP-AIO produces distinct outcomes -- automate, augment, hybrid, and preserve -- for candidates that standard cost-benefit analysis would uniformly automate. Threshold sensitivity analysis confirms the gate decisions are robust to upward perturbations of at least 14% in three of four representative cases. Keywords: AI governance, automation decision, human oversight, tacit knowledge, organizational resilience, financial services

Summary

Main Finding

The paper introduces PHP-AIO (Protocol for Human Preservation in AI-Optimized Organizations), a prescriptive, auditable five-gate decision protocol that augments standard ROI automation analysis by quantifying four categories of unpriced systemic risk (tacit-knowledge erosion, resilience reduction, regulatory exposure, socio‑institutional capital degradation). PHP-AIO produces one of four recommended outcomes (automate / preserve / augment / hybrid), a composite check to catch accumulated moderate risk, and a closed-form process-level automation‑debt measure ρ(P) that warns when high automation density lacks a regulator‑mandated human‑in‑the‑loop (HITL) anchor. Applied to stylised financial‑services roles, the protocol often recommends preserving or augmenting roles that plain cost‑benefit would automate; sensitivity checks show gate decisions are robust to modest perturbations (≥14% in three of four cases).

Key Points

  • Motivation: Standard ROI calculations omit long‑run, systemic costs from automating human roles that are strategic to organizational autonomy and recovery.
  • Four systemic risk dimensions (each normalized to [0,1] and combined from three subcomponents):
    • Tacit‑Knowledge Erosion (TKE): codifiability, irreversibility, criticality. Weighting used: 0.4/0.3/0.3.
    • Resilience Reduction (RR): recovery/fallback capacity, adversarial/novel‑threat detection, correlation of dependent systems. Weighting: 0.35/0.35/0.30.
    • Regulatory Exposure (RE): consequentiality, jurisdictional factors (country multipliers), reversal cost. Weighting: 0.4/0.4/0.2.
    • Socio‑Institutional Capital Degradation (SCD): relational/customer expectation, legitimacy/brand, labor/team effects. Weighting: 0.35/0.35/0.30.
  • Five‑gate sequential protocol (role must pass gates in order):
  • Net benefit (G1): require ≥15% net benefit (cheap pre‑filter).
  • TKE (G2): threshold < 0.70 (failure → preserve).
  • RR (G3): threshold < 0.70 (failure → preserve).
  • RE (G4): threshold < 0.70 (failure → hybrid).
  • SCD (G5): threshold < 0.70 (failure → augment).
    • If all five passed, a composite score PHP‑AIO = weighted average (financial‑services weights wTKE=0.30, wRR=0.25, wRE=0.30, wSCD=0.15) is computed; composite ≥0.60 → hybrid, else automate.
  • Automation‑debt density ρ(P) = time‑weighted share of subtasks automated. Warning triggers if ρ(P) ≥ 0.80 and there is no regulator‑mandated HITL anchor on any leaf task.
  • Time dynamics: each gate score can be projected over horizons H ∈ {1,3,5} with linear drift rates (financial‑services defaults δTKE=+0.03/yr, δRR=−0.02/yr, δRE=+0.05/yr, δSCD=0), producing conditional approvals or scheduled re‑evaluations.
  • Operationalization and auditability:
    • Scoring uses a 40‑field structured input schema (financial‑services v1) mapping inputs to subcomponents.
    • Deterministic scoring path; LLMs allowed for assistive/narrative roles but excluded from scoring to preserve reproducibility and traceability.
    • Full audit trail: inputs, gate scores, config version (weights/thresholds/schema), protocol version.
  • Governance outputs and prescribed actions:
    • automate: proceed; quarterly re‑evaluation.
    • preserve: keep human role; document rationale; annual review.
    • augment: AI assists while human decides; gradual knowledge capture.
    • hybrid: AI executes with human supervision; potential role splitting.
  • Limitations acknowledged by authors: weights and thresholds are expert‑configured (Santander AI Lab defaults tuned to financial services), composite weighting and drift assumptions are configurable and can materially affect outcomes.

Data & Methods

  • Normalization: every raw subcomponent rescaled affinely to [0,1] using its effective input range so scores are comparable across subcomponents.
  • Scoring formulas:
    • Dimension scores are weighted sums of three normalized subcomponents (weights sum to 1); per‑dimension weights provided (see Key Points).
    • Composite score is another weighted sum of the four dimension scores (financial‑services default weights reported).
  • Structured schema: 40 input fields grouped into task/context, net‑benefit, and fields mapping to the three subcomponents per dimension; two qualitative fields used only for narrative justification.
  • Deterministic evaluation: given identical structured inputs and a configuration version, the protocol yields identical gates and decision outputs (designed for auditability and regulatory trace).
  • Automation‑debt metric ρ(P): time‑weighted proportion of subtasks classified as automate; HITL anchors (if regulator‑required) neutralize warnings even at high density.
  • Sensitivity and robustness: threshold sensitivity analysis was performed on illustrative cases; three of four representative roles were robust to ≥14% upward perturbations.
  • Examples and calibration: worked examples in appendices show construction of scores; thresholds (0.70 per gate, composite 0.60) and drift rates are calibrated against stylised internal roles in the paper.

Implications for AI Economics

  • Internalizing externalities: PHP‑AIO operationalizes previously unpriced long‑run externalities of automation (loss of tacit knowledge, reduced resilience, future regulatory reversal costs, trust erosion), enabling firms to incorporate these into project valuation and capital allocation decisions.
  • Investment decision-making: replacing simple ROI thresholds with multi‑dimensional, auditable risk scoring will change which roles/processes are automated, slowing or redirecting automation where systemic risks are high and shifting investments toward augmentation/hybrid governance (human + AI).
  • Path dependence and hysteresis: the automation‑debt concept formalizes path dependence—high process automation density creates recovery costs and potentially irreversible loss of capabilities—so early automation choices have persistent economic consequences. This favors conservative adoption where HITL anchors or regulatory safeguards are absent.
  • Regulatory capital & compliance economics: RE explicitly quantifies regulatory option value; incorporating it into automation decisions can reduce the risk of regulatory penalties and forced reversals, and suggests firms should value flexible architectures (easy reintroduction of human oversight) as real optionality.
  • Labor market and role composition: the protocol produces graded outcomes (preserve/augment/hybrid) that imply varied effects on employment—some roles kept intentionally human (preservation of high‑TKE roles), others redesigned into hybrid supervisory positions—altering the shape of labor demand (more oversight and augmentation roles).
  • Organizational resilience and productivity tradeoffs: optimizing purely for throughput and unit cost risks increasing fragility; PHP‑AIO quantifies the tradeoff allowing firms to choose resilience (and potentially higher long‑run productive capacity) over short‑term efficiency gains.
  • Policy and macro implications: if adopted broadly or encoded into regulation, PHP‑AIO‑type requirements would raise the effective cost of full automation for sensitive functions, possibly slowing productivity gains from AI in regulated sectors but increasing social welfare by preserving institutional capacity, reducing systemic failure risk, and maintaining trust capital.
  • Practical adoption considerations: economic models and stress tests should incorporate automation‑debt and time‑drift parameters; firms will need to invest in structured data collection (the 40‑field schema) and governance to generate reproducible scores. The protocol’s outcomes are sensitive to chosen weights, thresholds, and drift assumptions—requiring sectoral calibration and empirical validation.

Suggested next steps for researchers/economists: - Empirically validate PHP‑AIO components (weights, thresholds, drift rates) across firms and sectors. - Model the macroeconomic effects of widespread adoption, balancing slower automation against avoided systemic failures. - Integrate automation‑debt into firm-level investment models, option valuations, and regulatory capital frameworks.

Assessment

Paper Typetheoretical Evidence Strengthn/a — Paper is a conceptual and formal framework with stylised examples and sensitivity checks rather than empirical or causal identification; it offers no field or observational evidence establishing real-world effects. Methods Rigormedium — The authors provide a formal five-gate protocol and a closed-form automation-debt measure with sensitivity analysis on stylised role profiles, which shows internal logical consistency; however, the approach rests on strong assumptions, uses no real-world data or experiments, and lacks validation against organizational outcomes. SampleNo empirical sample; analysis uses stylised profiles of representative internal roles (illustrative role-level parameters) and simulated threshold perturbations within a financial-services framing. Themesgovernance org_design human_ai_collab productivity GeneralizabilityBased on stylised, illustrative role profiles rather than observed organizational data, Developed and applied in a financial-services framing; sector differences may limit transferability, Relies on assumptions about tacit knowledge, resilience, regulatory enforcement, and socio-institutional capital that may vary across firms, No behavioral or implementation evidence — ignores organizational politics, incentives, and adaptation over time, Regulatory-mandated human-in-the-loop anchor is context-dependent and may not exist in many jurisdictions

Claims (6)

ClaimDirectionOutcomeConfidence & EvidenceDetails
Standard automation ROI misses four categories of systemic risk: tacit knowledge erosion, resilience reduction, regulatory exposure, and socio-institutional capital degradation. Organizational Efficiency negative recognition/valuation of systemic risks in automation ROI
Reading fidelity high
Study strength speculative
not reported
0.02
These unpriced systemic risks affect long-term organizational performance. Organizational Efficiency negative long-term organizational performance
Reading fidelity high
Study strength speculative
not reported
0.02
PHP-AIO (Protocol for Human Preservation in AI-Optimized Organizations) is a five-gate sequential decision protocol with a final composite check that quantifies these unpriced systemic risks at the role level and produces auditable automation decisions. Governance And Regulation positive ability to quantify unpriced systemic risks and produce auditable automation decisions
Reading fidelity high
Study strength speculative
not reported
0.02
A closed-form automation-debt measure (ρ(P)) formalises how role-level decisions accumulate across multi-step processes, and its warning is neutralised only by a regulator-mandated human-in-the-loop anchor. Automation Exposure negative automation-debt accumulation across multi-step processes and conditions for neutralising its warning
Reading fidelity high
Study strength medium
not reported
0.12
Applied to stylised profiles of representative internal roles, PHP-AIO produces distinct outcomes -- automate, augment, hybrid, and preserve -- for candidates that standard cost-benefit analysis would uniformly automate. Task Allocation mixed automation decision category assigned to role candidates (automate/augment/hybrid/preserve)
Reading fidelity high
Study strength medium
n=4
0.12
Threshold sensitivity analysis confirms the gate decisions are robust to upward perturbations of at least 14% in three of four representative cases. Task Allocation positive robustness of gate decisions to parameter perturbations
Reading fidelity high
Study strength medium
n=4
robust to upward perturbations of at least 14% in three of four representative cases
0.12

Notes