The Commonplace
Home Papers Evidence Explore Trends Syntheses Digests References Docs 🎲 Workforce Futures
← Papers
Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review. How this is built →

AGI is not just an engineering problem: structural constraints across technical, social, legal and economic levels mean scaling architectures alone won’t yield general intelligence; the paper maps 23 constraints and issues five benchmarked, falsifiable predictions to redirect research.

What General Intelligence Requires: Non-Reducible Constraints Across Levels of Description
Subhomoy Bakshi · July 21, 2026
arxiv theoretical n/a evidence 7/10 relevance Full text usable extracted full text Source PDF

Structured author observations

Linked only from stored provider relations; the raw author line above is never matched by name.

Arxiv

Latest observation:

  1. Subhomoy Bakshi unresolved corpus identity

Semantic Scholar

Latest observation:

  1. Subhomoy Bakshi provider ID
The paper argues that AGI cannot be produced by architectural advances or scaling alone because structurally distinct, non-reducible constraints across technical, social, legal, and economic levels block single-path solutions, and it maps 23 constraints plus five falsifiable predictions to reframe AGI research priorities.

Citation observations

Cumulative provider counts captured on specific dates; providers are never combined.

General intelligence, of the kind that underwrites the full range of human cognitive achievement, is not a property of computational architecture alone. This paper advances a single thesis: the structural constraints on general intelligence occupy distinct levels of description and are mutually non-reducible, in the sense that the special-sciences tradition gives to that term. It follows that no single architectural advance, and no continuation of the scaling programme by itself, can produce artificial general intelligence (AGI), and that research programmes must be evaluated against the full constraint profile rather than against performance on any one benchmark. The thesis is developed through a method that reads general intelligence through four evidential lenses, AI systems research, anthropology, law, and economics, each anchored to a distinct level of description, supplemented by speculative fiction used as a disciplined heuristic in the context of discovery rather than the context of justification. Applying the method yields a taxonomy of twenty-three structural constraints organised into eight clusters; six are examined in depth and ordered as an ascending ladder of levels, with explicit bridges showing why progress at one level cannot carry to the next. The argument issues in five falsifiable predictions, each stated with a named benchmark family and a disconfirmation condition, converting a descriptive framework into a research programme with a longer horizon than the scaling hypothesis implies.

Summary

Main Finding

General intelligence (the kind that supports the full range of human cognitive achievement) requires a set of structural constraints that sit at distinct levels of description and are mutually non-reducible. Therefore no single architectural advance — and no mere continuation of the current scaling programme (larger models, more data, more compute) — can by itself produce artificial general intelligence (AGI). Research progress toward AGI must be judged against a full, cross-level constraint profile rather than by performance on individual benchmarks.

Key Points

  • Thesis: Structural constraints on general intelligence occupy separate Fodorian special-science levels (computational, evolutionary/developmental, normative/procedural, incentive-theoretic) and are multiply realizable so cannot be reduced to one another.
  • Multi-lens method: The paper reads AGI readiness through four evidential lenses:
    • AI systems (computational level) — what architectures do and cannot do (focus on transformer LLMs and scaling literature).
    • Anthropology (evolutionary & developmental level) — how embodied, social evolution shaped human intelligence (e.g., cumulative cultural learning, embodied conceptual grounding).
    • Law (normative & procedural level) — structural properties required for an agent to be a participant in rule-governed systems (invokes Fuller’s “inner morality of law”).
    • Economics (incentive-theoretic level) — the incentive, coordination, and information-structure constraints that shape agent behavior in multi-agent systems.
  • Taxonomy: The author produces a candidate taxonomy of 23 structural constraints organized into 8 clusters by level; six constraints are treated in depth and used to build a ladder of levels with explicit “bridge” arguments showing why progress at one level does not guarantee progress at the next.
  • Demarcation criteria for retaining constraints: necessity (required for general intelligence, not just human contingencies), independence (not derivable from others), level specificity, and evidential support from at least two independent traditions.
  • Symbol grounding example: Demonstrates convergence across lenses — AI (tokens vs. world), anthropology (sensorimotor concept formation), law (witness/contractual competence), economics (representations necessary for market coordination).
  • Predictions: The framework yields five falsifiable predictions (each tied to benchmark families and disconfirmation conditions), turning the descriptive taxonomy into a testable research programme.
  • Methodological note: Speculative fiction is used as a disciplined heuristic for discovery (to suggest candidate constraints), but evidence must come from the four lenses; the analysis also explicitly addresses the “N = 1” problem (human intelligence as the only observed AGI) via the necessity criterion.
  • Conclusion about scaling: Empirical success of scaling on many benchmarks is acknowledged, but benchmark performance alone does not demonstrate that the structural constraints across the four levels have been satisfied.

Data & Methods

  • Nature of contribution: Conceptual, interdisciplinary synthesis rather than new experimental data. The paper integrates empirical and theoretical literatures across AI systems research, anthropology/evolutionary psychology, legal theory, and economics.
  • Sources drawn on (representative):
    • AI/ML: Transformer/LLM literature and scaling arguments (e.g., Vaswani et al.; Brown et al.; Kaplan et al.; Morris et al.; scaling critiques such as Marcus & Davis).
    • Anthropology/development: Henrich on cumulative cultural learning; Lakoff & Johnson on embodied metaphor and conceptual grounding; developmental work on theory of mind.
    • Law: Jurisprudential tools for interpretation, ambiguity, precedent, Fuller’s eight principles for subjecthood and legal participation.
    • Economics: Coordination, information, and contractual problems (references include classic insights like Akerlof’s information asymmetry).
  • Methodology:
    • Define working operational notion of general intelligence: capacity to acquire and deploy competencies across domains given sufficient time and information (combines Legg & Hutter with Chollet’s skill-acquisition emphasis).
    • Apply four lenses, each anchored to a different level of description (Fodorean special sciences) rather than Marr-style levels.
    • Generate candidate constraints (initially ~30), apply demarcation criteria, merge or discard dependent constraints, produce a final set of 23 candidates (with 6 explored deeply; others catalogued as candidates).
    • Use cross-lens grounding rule: retain constraints only if supported by at least two independent evidential traditions.
    • Produce falsifiable predictions: each associated with benchmark families and explicit disconfirmation conditions to make the framework empirically testable over a longer time horizon than immediate scaling claims.
  • Limitations acknowledged:
    • Not a closed catalogue; taxonomy is a mapping of candidate constraints.
    • Some constraints might later be shown reducible, addressable by architectures, or less fundamental.
    • The analysis works from a single observed instance of general intelligence (humans); the necessity criterion is used to avoid conflating contingency with necessity.

Implications for AI Economics

  • Scaling alone is insufficient for economic coordination: Economic-level constraints (incentives, information, institutions, multi-agent strategic interaction) are distinct from computational capacities. Systems that perform well on predictive benchmarks may still lack representational and institutional features necessary for robust participation in markets, contracting, and other economic interactions.
  • Multi-agent environments matter: AGI-like agents operate in settings with competing goals, asymmetric information, repeated interactions, and institutional rules. Economic models must incorporate how AI systems build and update models of other agents, signal intentions, and respond to sanctions/rewards embedded in legal and market structures.
  • Representations and grounding affect market outcomes: If models lack non-arbitrary grounding of representations, they may miscoordinate, misprice, or misreport states that markets rely on (e.g., valuations, contracts, verifiable claims). Information quality and credibility are fundamental public goods; AI limitations on grounded reference can produce market failures (Akerlof–style adverse selection and coordination problems).
  • Institutional design & regulation become central research topics: Because law and institutions provide normative scaffolding that human intelligence evolved to exploit, economics of AI must study how regulatory rules, liability regimes, certification, and contract design can substitute for or complement missing normative capabilities in AI agents.
  • Evaluation metrics for economic impact: Standard ML benchmarks are insufficient for predicting economic behavior of deployed systems. New benchmark families should measure multi-agent robustness, incentive-responsiveness, reputational dynamics, commitment capacity, and legal/contractual interpretability. The paper's falsifiable predictions linking constraint clusters to benchmarks suggest explicit economic performance tests (e.g., market participation, bargaining outcomes, contract enforcement scenarios).
  • Labor, complementarities, and collective intelligence: The taxonomy highlights that human intelligence depends on cumulative cultural transmission and institutions; similarly, AGI’s economic effects will depend on how systems integrate with human collective intelligence, augment or substitute labor, and interact with institutions that preserve or shift complementarities across tasks and sectors.
  • Research agenda for AI economics (practical directions):
    • Incorporate multi-level constraints into models of AI adoption and impact (not just productivity gains from prediction improvement).
    • Design market and legal mechanisms that internalize externalities from misgrounded or strategically fragile AI behavior (warranties, certification, reputational systems).
    • Build economic benchmarks that test AI systems in strategic, repeated, legally structured interactions (e.g., contracting games, markets with asymmetric information, public-good provisioning).
    • Study distributional and systemic risks that arise when architectures satisfy computational benchmarks but fail at normative/incentive-theoretic levels (e.g., coordination failures, manipulation, regulatory arbitrage).
    • Evaluate policy levers (liability, disclosure, mandatory interpretability/certification) that could close gaps identified in the taxonomy.
  • Broader policy implication: Because constraints span levels, effective governance and economic policy must combine technical standards with institutional design — economic incentives, legal rules, and social practices — to shape safe, reliable, and socially beneficial deployment of advanced AI systems.

If you want, I can: - Extract the six constraints examined in depth and summarize each with its cross-lens evidence and why it’s non-reducible, or - Translate the five falsifiable predictions into concrete benchmark definitions and measurement proposals relevant for economic outcomes.

Assessment

Paper Typetheoretical Evidence Strengthn/a — This is a cross-disciplinary conceptual and theoretical argument rather than an empirical study; it synthesizes literature and speculative heuristics but does not provide empirical tests or causal estimates. Methods Rigormedium — The paper deploys a structured, multi-lens argumentative method (AI systems research, anthropology, law, economics) and produces a clear taxonomy and falsifiable predictions, demonstrating intellectual rigor; however, it relies on interpretive synthesis and speculative fiction rather than systematic empirical validation, which limits methodological rigor. SampleNo empirical sample; the work synthesizes literature and evidence from AI systems research, anthropology, law, and economics, augmented by disciplined use of speculative fiction as a discovery heuristic; produces a taxonomy of 23 structural constraints and five named, benchmark-linked falsifiable predictions. Themesinnovation governance GeneralizabilityNon-empirical conceptual work lacking empirical validation, Conclusions rely on current social, legal, and economic contexts which may change, Findings may not hold if unforeseen architectural or scientific breakthroughs occur, Use of speculative fiction introduces scenario-dependence and interpretive variability, Benchmarks and disconfirmation conditions may be narrow or contestable

Claims (8)

ClaimDirectionOutcomeConfidence & EvidenceDetails
General intelligence, of the kind that underwrites the full range of human cognitive achievement, is not a property of computational architecture alone. Other negative whether general intelligence (human-like AGI) is solely attributable to computational architecture
Reading fidelity high
Study strength speculative
not reported
0.02
The structural constraints on general intelligence occupy distinct levels of description and are mutually non-reducible (in the special-sciences sense). Other positive ontology/structure of constraints on general intelligence (distinctness and non-reducibility of levels)
Reading fidelity high
Study strength speculative
not reported
0.02
No single architectural advance, and no continuation of the scaling programme by itself, can produce artificial general intelligence (AGI). Other negative likelihood that AGI will emerge from a single architectural advance or continued scaling alone
Reading fidelity high
Study strength speculative
not reported
0.02
Research programmes must be evaluated against the full constraint profile rather than against performance on any one benchmark. Other positive appropriate evaluation criteria for AI research programmes (full constraint profile vs single benchmark performance)
Reading fidelity high
Study strength speculative
not reported
0.02
The paper reads general intelligence through four evidential lenses—AI systems research, anthropology, law, and economics—each anchored to a distinct level of description. Other positive method of analysis (use of four evidential lenses anchored to distinct levels)
Reading fidelity high
Study strength speculative
not reported
0.02
The paper produces a taxonomy of twenty-three structural constraints organised into eight clusters; six of these are examined in depth and ordered as an ascending ladder of levels with explicit bridges explaining why progress at one level cannot carry to the next. Other positive existence and organization of a taxonomy of constraints (23 constraints, 8 clusters, 6 examined in depth)
Reading fidelity high
Study strength speculative
not reported
0.02
The argument issues in five falsifiable predictions, each stated with a named benchmark family and a disconfirmation condition, converting a descriptive framework into a research programme with a longer horizon than the scaling hypothesis implies. Other positive presence of five falsifiable predictions linking theory to benchmarks and disconfirmation criteria
Reading fidelity high
Study strength speculative
not reported
0.02
Speculative fiction is used as a disciplined heuristic in the context of discovery rather than the context of justification. Other positive methodological role of speculative fiction in the paper (heuristic for discovery)
Reading fidelity high
Study strength speculative
not reported
0.02

Notes