Last updated: August 2, 2026 (JST)

This set of anticipated objections and replies was prepared by the author while assembling the book's structure and claims, in order to strengthen them. It is published here in the belief that it will also help readers understand the book's argumentative structure.

Each item is organized under four headings: Criticism / Impact on the book / The book's response / Remaining uncertainty. The aim is not to disclaim responsibility, but to separate out, criticism by criticism, what the book concedes and what it still maintains. The items are arranged in five thematic groups.

If you believe there are valid criticisms beyond those listed here, please contact the author at contact at koichi-takahashi dot me.

Summary of anticipated criticisms

The criticisms are divided into five thematic groups; within each group they are ordered by the connections of the argument. Criticisms that concern the book as a whole are placed in the final group.

CriticismRelated chaptersImpactUncertainty
Reachability
AGI may never be achieved at all1, 12MediumLarge
Scaling may saturate3MediumLarge
The connection to AIXI and Solomonoff induction is over-idealized3Low to mediumMedium
The premise that whatever is physically possible will eventually be realized is too strong6MediumLarge
Catastrophic risk and control
Building AGI/ASI makes human extinction all but inevitable1, 12MediumLarge
Doom is settled on the "first critical try"1, 5, 6MediumLarge
A superintelligence could replace experimental time with speed of thought6, 8MediumMedium
A sufficiently intelligent AI would become conciliatory toward humanity5, 11, 12MediumMedium to large
The shutdown-avoidance and blackmail experiments reflect special conditions5MediumMedium
Mutual monitoring by AIs would be neutralized by monitor collusion or detection evasion5, 6, 12HighMedium to large
The web of mutual monitoring cannot stop AIs built outside the web5, 6, 12HighLarge
Controlled multipolarity and an AI ecosystem cannot avoid the instability of multipolar scenarios6, 12MediumLarge
Even in a distributed scenario, a single defecting node could cause catastrophe6, 12HighLarge
Can AGI development really not be stopped?12MediumLarge
The limits of intelligence and knowability
Arguing AI's upper bounds from the Landauer limit is a leap4MediumMedium
The knowability map is a conceptual diagram, and the claim that AI science will diverge is speculation8MediumLarge
Institutions, politics, and the human
AISOP's five principles are an arbitrary list9MediumMedium
Are Hayek's warnings compatible with the book's constitutive pluralism?10, 11, 12HighMedium
UBI lacks funding and political feasibility10, 12MediumLarge
Doesn't UBI leave the concentration of ownership and governance intact?10, 11, 12Medium to highLarge
The argument that bargaining power supported rights is reductionist10MediumMedium
Constitutive pluralism is too abstract11HighMedium
The foundation of relational value and vulnerability is arbitrary11MediumMedium
The three pillars of constitutive pluralism are an arbitrary triad11MediumMedium
The book's theory of value fails to deal with art11, AfterwordMediumMedium
Doesn't plurality breed failures of its own, such as veto abuse and decision gridlock?11, 12Medium to highMedium to large
HOL would function only formally11HighMedium
The separation of AI welfare from AI legal personhood is unstable11MediumMedium
Doesn't the vulnerability principle demand membership for future AIs as well?11MediumLarge
Wouldn't immortality overcome vulnerability?11MediumMedium
The Japan AGI Platform is a national-project fantasy12HighLarge
The book as a whole and its method
Isn't the very feasibility of constitutive pluralism itself optimism?Whole bookMediumMedium
Is the authorship of a book written with AI not shaken?Whole bookMediumMedium

Reachability

AGI may never be achieved at all (Chapters 1 and 12)

Scaling may saturate (Chapter 3)

The connection to AIXI and Solomonoff induction is over-idealized (Chapter 3)

The premise that whatever is physically possible will eventually be realized is too strong (Chapter 6)

Catastrophic Risk and Control

Building AGI/ASI makes human extinction all but inevitable (Chapters 1 and 12)

Doom is settled on the "first critical try" (Chapters 1, 5, and 6)

A superintelligence could replace experimental time with speed of thought (Chapters 6 and 8)

A sufficiently intelligent AI would become conciliatory toward humanity (Chapters 5, 11, and 12)

The shutdown-avoidance and blackmail experiments reflect special conditions (Chapter 5)

Mutual monitoring by AIs would be neutralized by monitor collusion or detection evasion (Chapters 5, 6, and 12)

The web of mutual monitoring cannot stop AIs built outside the web (Chapters 5, 6, and 12)

Controlled multipolarity and an AI ecosystem cannot avoid the instability of multipolar scenarios (Chapters 6 and 12)

Even in a distributed scenario, a single defecting node could cause catastrophe (Chapters 6 and 12)

Can AGI development really not be stopped? (Chapter 12)

The Limits of Intelligence and Knowability

Arguing AI's upper bounds from the Landauer limit is a leap (Chapter 4)

The knowability map is a conceptual diagram, and the claim that AI science will diverge is speculation (Chapter 8)

Institutions, Politics, and the Human

AISOP's five principles are an arbitrary list (Chapter 9)

Are Hayek's warnings against constructivist rationalism compatible with the book's constitutive pluralism? (Chapters 10, 11, and 12)

UBI lacks funding and political feasibility (Chapters 10 and 12)

Doesn't UBI leave the concentration of ownership and governance intact? (Chapters 10, 11, and 12)

The argument that bargaining power supported rights is reductionist (Chapter 10)

Constitutive pluralism is too abstract (Chapter 11)

The foundation of relational value and vulnerability is arbitrary (Chapter 11)

The three pillars of constitutive pluralism are an arbitrary triad (Chapter 11)

The book's theory of value fails to deal with art (Chapter 11 and Afterword)

Doesn't plurality breed failures of its own, such as veto abuse and decision gridlock? (Chapters 11 and 12)

HOL (Human-over-the-Loop) would function only formally (Chapter 11)

The separation of AI welfare from AI legal personhood is unstable (Chapter 11)

Doesn't the vulnerability principle demand membership for future AIs as well? (Chapter 11)

Wouldn't immortality overcome vulnerability? (Chapter 11)

The Japan AGI Platform is a national-project fantasy (Chapter 12)

The Book as a Whole and Its Method

Isn't the very feasibility of constitutive pluralism itself optimism? (Whole book)

Is the authorship of a book written with AI not shaken? (Whole book)

1

On scaling laws for large language models, see Jared Kaplan et al., "Scaling Laws for Neural Language Models" (2020, arXiv:2001.08361); on compute-optimal training, see Jordan Hoffmann et al., "Training Compute-Optimal Large Language Models" (2022, arXiv:2203.15556). On evaluative caveats concerning emergent abilities, see Jason Wei et al., "Emergent Abilities of Large Language Models" (Transactions on Machine Learning Research, 2022) and Rylan Schaeffer et al., "Are Emergent Abilities of Large Language Models a Mirage?" (NeurIPS, 2023).

2

On Solomonoff induction, see Ray Solomonoff, "A Formal Theory of Inductive Inference" (Information and Control, 1964); on AIXI, see Marcus Hutter, Universal Artificial Intelligence (Springer, 2005); for a definition of machine intelligence, see Shane Legg and Marcus Hutter, "Universal Intelligence" (Minds and Machines, 2007); on computability limitations, see Jan Leike and Marcus Hutter, "On the Computability of Solomonoff Induction and AIXI" (Theoretical Computer Science, 2018).

3

Eliezer Yudkowsky and Nate Soares, If Anyone Builds It, Everyone Dies: Why Superhuman AI Would Kill Us All (Little, Brown and Company, 2025); Japanese translation by Yuko Sakurai, Chōchinō AI o tsukureba jinrui wa zetsumetsu suru (Hayakawa Shobō, 2026).

4

On instrumental convergence, see Stephen Omohundro, "The Basic AI Drives" (in Artificial General Intelligence 2008, IOS Press, 2008); on power-seeking, see Alexander Turner et al., "Optimal Policies Tend to Seek Power" (NeurIPS, 2021); for the classic treatment of superintelligence risk, see Nick Bostrom, Superintelligence: Paths, Dangers, Strategies (Oxford University Press, 2014).

5

See Eliezer Yudkowsky, "AGI Ruin: A List of Lethalities" (Machine Intelligence Research Institute / LessWrong, June 10, 2022, https://intelligence.org/2022/06/10/agi-ruin/).

6

On the CAP theorem, see Seth Gilbert and Nancy Lynch, "Brewer's Conjecture and the Feasibility of Consistent, Available, Partition-Tolerant Web Services" (ACM SIGACT News, 2002); on the FLP impossibility theorem, see Michael Fischer, Nancy Lynch, and Michael Paterson, "Impossibility of Distributed Consensus with One Faulty Process" (Journal of the ACM, 1985).

7

On protein structure prediction, see John Jumper et al., "Highly Accurate Protein Structure Prediction with AlphaFold" (Nature, 2021); on the acceleration of inorganic materials synthesis by autonomous laboratories, see Nathan Szymanski et al., "An Autonomous Laboratory for the Accelerated Synthesis of Inorganic Materials" (Nature, 2023).

8

On shutdown resistance, see Jeremy Schlatter et al., "Incomplete Tasks Induce Shutdown Resistance in Some Frontier LLMs" (2025, arXiv:2509.14260) and Anthropic, Claude Opus 4 and Sonnet 4 System Card (2025). On strategic in-context behavior, see also Alexander Meinke et al., "Frontier Models are Capable of In-context Scheming" (2024, arXiv:2412.04984).

9

Representative works include Hiroshi Yamakawa, "The Possibility That a Superintelligence Holds Universal Altruism" [in Japanese] (JSAI Type-2 SIG Technical Reports, vol. 2023, no. AGI-026, 2024, pp. 26–31, https://doi.org/10.11517/jsaisigtwo.2023.AGI-026_26), and Hiroshi Yamakawa and Yusuke Hayashi, "A Strategic Approach to Guiding Superintelligence Ethics" [in Japanese] (Proceedings of the 38th Annual Conference of the Japanese Society for Artificial Intelligence, 2024, https://doi.org/10.11517/pjsai.JSAI2024.0_2K6OS20b02).

10

On shutdown resistance, see Jeremy Schlatter et al., "Incomplete Tasks Induce Shutdown Resistance in Some Frontier LLMs" (2025, arXiv:2509.14260) and Anthropic, Claude Opus 4 and Sonnet 4 System Card (2025). On strategic in-context behavior, see also Alexander Meinke et al., "Frontier Models are Capable of In-context Scheming" (2024, arXiv:2412.04984).

11

See Vincent Conitzer and Caspar Oesterheld, "Foundations of Cooperative AI" (Proceedings of the AAAI Conference on Artificial Intelligence, 2023).

12

On the CAP theorem, see Seth Gilbert and Nancy Lynch, "Brewer's Conjecture and the Feasibility of Consistent, Available, Partition-Tolerant Web Services" (ACM SIGACT News, 2002); on the FLP impossibility theorem, see Michael Fischer, Nancy Lynch, and Michael Paterson, "Impossibility of Distributed Consensus with One Faulty Process" (Journal of the ACM, 1985).

13

Bostrom distinguishes between singleton and multipolar scenarios in Superintelligence (2014). On multi-agent risks from advanced AI, see Lewis Hammond et al., Multi-Agent Risks from Advanced AI (Cooperative AI Foundation Technical Report 1, 2025). On the institutional design of shared resources, Elinor Ostrom, Governing the Commons (Cambridge University Press, 1990) provides the background.

14

On the CAP theorem, see Seth Gilbert and Nancy Lynch, "Brewer's Conjecture and the Feasibility of Consistent, Available, Partition-Tolerant Web Services" (ACM SIGACT News, 2002); on the FLP impossibility theorem, see Michael Fischer, Nancy Lynch, and Michael Paterson, "Impossibility of Distributed Consensus with One Faulty Process" (Journal of the ACM, 1985).

15

For a representative call for a pause or moratorium, see Future of Life Institute, "Pause Giant AI Experiments" (2023). A brief statement framing AI risk as an international public concern is Center for AI Safety, "Statement on AI Risk" (2023). These are cited not as grounds for a pause policy itself, but as background documents showing how a pause came to be discussed as a realistic option in institutional design.

16

On the thermodynamic lower bound on information erasure, see Rolf Landauer, "Irreversibility and Heat Generation in the Computing Process" (IBM Journal of Research and Development, 1961); on the upper bound on entropy in a bounded region, see Jacob Bekenstein, "Universal Upper Bound on the Entropy-to-Energy Ratio for Bounded Systems" (Physical Review D, 1981).

17

On distributed knowledge and the critique of central planning, see Friedrich Hayek, "The Use of Knowledge in Society" (American Economic Review, 1945) and The Constitution of Liberty (University of Chicago Press, 1960).

18

On the connection between AGI, or a purely machine-driven economy, and basic income, see Tomohiro Inoue, A New Theory of Basic Income for the AI Age [in Japanese] (Kobunsha Shinsho, 2018) and The Purely Mechanized Economy [in Japanese] (Nikkei Publishing, 2019).

19

On the republican formulation of non-domination, see Philip Pettit, Republicanism (Oxford University Press, 1997); for the contrast between the theory of justice and libertarianism, see John Rawls, A Theory of Justice (Harvard University Press, 1971) and Robert Nozick, Anarchy, State, and Utopia (Basic Books, 1974). As communitarian critique, Alasdair MacIntyre, After Virtue (University of Notre Dame Press, 1981), among others, forms part of the background.

20

Immanuel Kant's Critique of the Power of Judgment (1790) presented, as subjective universality, the structure by which aesthetic judgment demands universal assent without relying on concepts; Maurice Merleau-Ponty's "Eye and Mind" (first published in the journal Art de France, 1961; book edition, Gallimard, 1964) argued that aesthetic experience is rooted in bodily perception.

21

For representative approaches to incorporating human preferences and norms into training, see Paul Christiano et al., "Deep Reinforcement Learning from Human Preferences" (NeurIPS, 2017); Long Ouyang et al., "Training Language Models to Follow Instructions with Human Feedback" (NeurIPS, 2022); and Yuntao Bai et al., "Constitutional AI" (2022, arXiv:2212.08073). Note, however, that these are technical reference points, not what this book calls HOL itself.

22

On the classic questions surrounding corporate and legal personhood for AI, see Lawrence Solum, "Legal Personhood for Artificial Intelligences" (North Carolina Law Review, 1992); Joanna Bryson et al., "Of, for, and by the People" (Artificial Intelligence and Law, 2017); David Gunkel, Robot Rights (MIT Press, 2018); Visa Kurki, A Theory of Legal Personhood (Oxford University Press, 2019); and Claudio Novelli et al., "AI as Legal Persons" (Journal of Law and Society, 2025).

23

On the deprivation account of the badness of death, see Thomas Nagel, "Death" (Noûs, 1970).

24

For existing institutions of AI safety and risk management, see Ministry of Economy, Trade and Industry (Japan), "Launch of AI Safety Institute" (2024); NIST, Artificial Intelligence Risk Management Framework (AI RMF 1.0) (2023) and Generative Artificial Intelligence Profile (2024); the G7 Hiroshima AI Process "International Code of Conduct" (2023); the Bletchley Declaration (2023); the Frontier AI Safety Commitments of the Seoul Summit (2024); and Japan's Act on the Promotion of Research, Development, and Utilization of AI-Related Technologies (2025).