Version 1.21 (2023). First draft:
2020.
Cambridge Journal of Law, Politics, and Art, 2025,
forthcoming
Recent rapid advances in artificial intelligence makes it timely to start considering what a future society might look like in which humans share the world with digital minds of various kinds and sophistication. Some of those digital minds might be sentient or sapient or possess other bases for claiming degrees of moral and/or political status. At the same time, because their natures may differ in important respects from those of human beings, it would not always be appropriate to simply apply current human norms to such a radically different context. We believe that it is important to begin exploring what shape a broadly cooperative and acceptable framework for harmonious coexistence could take; and to this end we put forward, very tentatively, some propositions concerning digital minds and society that seem to hold some plausibility to us. We are not ready, at this point, to confidently or “officially” endorse them, nor do they give a full picture of our views on these matters; but we put them forward to facilitate feedback and to invite broader discussion.2
“[M]ental states can supervene on any of a broad class of physical substrates. Provided a system implements the right sort of computational structures and processes, it can be associated with conscious experiences. It is not an essential property of consciousness that it is implemented on carbon-based biological neural networks inside a cranium: silicon-based processors inside a computer could in principle do the trick as well.”3
Block, Ned. 1981. “Psychologism and Behaviorism.” The Philosophical Review 90 (1): 5–43.
Bostrom, Nick. 2003. “Are You Living in a Computer Simulation?” Philosophical Quarterly 53 (211): 243–55.
—, 2014. Superintelligence. Oxford: Oxford University Press.
—, 2019. “The Vulnerable World Hypothesis.” Global Policy 10 (4): 455–76.
Bostrom, Nick, and Eliezer Yudkowsky. 2018. “The Ethics of Artificial Intelligence.” In Artificial Intelligence Safety and Security, edited by Roman V. Yampolskiy, 57–69. Boca Raton, FL: Chapman and Hall/CRC.
Brown, Tom, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D. Kaplan, Prafulla Dhariwal, Arvind Neelakantan et al. 2020. “Language models are few-shot learners.” arXiv. https://arxiv.org/abs/2005.14165.
Chalmers, David J. 1996. The Conscious Mind: In Search of a Fundamental Theory. New York: Oxford University Press.
—, 2010. “The Singularity: A Philosophical Analysis.” Journal of Consciousness Studies 17 (9–10): 7–65.
Daswani, Mayank, and Jan Leike. 2015. “A Definition of Happiness for Reinforcement Learning Agents.” arXiv:1505.04497. arXiv. https://arxiv.org/abs/1505.04497.
Evans, Owain, Owen Cotton-Barratt, Lukas Finnveden, Adam Bales, Avital Balwit, Peter Wills, Luca Righetti, and William Saunders. 2021. “Truthful AI: Developing and Governing AI That Does Not Lie.” arXiv. http://arxiv.org/abs/2110.06674.
Garfinkel, Ben. 2021. “A Tour of Emerging Cryptographic Technologies: What They Are and How They Could Matter.” Centre for the Governance of AI, Future of Humanity Institute, University of Oxford.
Hanson, Robin. 2016. The Age of Em: Work, Love, and Life When Robots Rule the Earth. Oxford: Oxford University Press.
Herzog, Michael H., Michael Esfeld, and Wulfram Gerstner. 2007. “Consciousness & the Small Network Argument.” Neural Networks 20 (9): 1054–56.
Jaworska, Agnieszka, and Julie Tannenbaum. 2021. “The Grounds of Moral Status.” In The Stanford Encyclopedia of Philosophy, edited by Edward N. Zalta, Spring 2021. Metaphysics Research Lab, Stanford University.
Kagan, Shelly. 2019. How to Count Animals, More or Less. Oxford: Oxford University Press.
McMahan, Jeff. 2005. “‘Our Fellow Creatures.’” The Journal of Ethics 9 (3–4): 353–80.
Muehlhauser, Luke. 2017. “2017 Report on Consciousness and Moral Patienthood.” Open Philanthropy.
Searle, John R. 1980. “Minds, Brains, and Programs.” Behavioral and Brain Sciences 3 (3): 417–24.
Shulman, Carl. 2010. “Whole Brain Emulation and the Evolution of Superorganisms.” Machine Intelligence Research Institute.
Shulman, Carl, and Nick Bostrom. 2021. “Sharing the World with Digital Minds.” In Rethinking Moral Status, edited by Steve Clarke, Hazem Zohny, and Julian Savulescu, 306–26. Oxford: Oxford University Press.
Sneddon, Lynne U., Robert W. Elwood, Shelley A. Adamo, and Matthew C. Leach. 2014. “Defining and Assessing Animal Pain.” Animal Behaviour 97: 201–12.
Tomasik, Brian. 2014. “Do Artificial Reinforcement-Learning Agents Matter Morally?” arXiv:1410.8233. arXiv. https://arxiv.org/abs/1410.8233.
Warren, Mary Anne. 1997. Moral Status: Obligations to Persons and Other Living Things. Oxford: Clarendon Press.
Future of Humanity Institute, University of Oxford.↩︎
For useful comments we are grateful to Stuart Armstrong, Michael Bailey, Adam Bales, Jake Beck, Asya Bergal, Sam Bowman, Patrick Butlin, Ryan Carey, Joseph Carlsmith, Paul Christiano, Michael Cohen, Teddy Collins, Owen Cotton-Barratt, Wes Cowley, Max Daniel, Eric Drexler, Daniel Eth, Owain Evans, Lukas Finnveden, Iason Gabriel, Aaron Gertler, Katja Grace, Julia Haas, Robin Hanson, Lewis Ho, Michael Huemer, Geoffrey Irving, Deej James, Ramana Kumar, Jan Leike, Robert Long, Vishal Maini, Matthew van der Merwe, Silvia Milano, Ed Moreau-Feldman, Venkat Nettimi, Richard Ngo, Eli Rose, Anders Sandberg, Eric Schwitzgebel, Jonathan Simon, Alex Spies, Nick Teague, Laura Weidinger, Peter Wills, and the participants in several seminars where earlier versions of this work were presented.↩︎
Bostrom 2003, p. 2. For some supporting argumentation, see, e.g., Chalmers 2010, §9; Chalmers 1996, §7. For examples of views that we reject, see, e.g., Searle 1980; Block 1981.↩︎
More precisely: If a conscious experience E supervenes on an implementation of computation C (in some ordinary computer that we have built), then two independent implementations of C (either on the same computer or another similar computer) will subvene “twice as much” experience as E (where the additional experience has exactly the same qualitative character as E).↩︎
Herzog, Esfeld, & Gerstner 2007↩︎
Muehlhauser 2017↩︎
An entity has moral status if and only if it or its interests morally matter to some degree for the entity’s own sake (Jaworska & Tannenbaum 2021).↩︎
Shulman & Bostrom 2021↩︎
Bostrom 2014, pp. 125–126↩︎
Shulman & Bostrom 2021; Bostrom & Yudkowsky 2018↩︎
Warren 1997↩︎
Kagan 2019↩︎
McMahan 2005, pp. 354, 361↩︎
Kagan 2019, pp. 130–137↩︎
Bostrom 2014, pp. 125–126↩︎
Cf. Bostrom 2019↩︎
Hanson 2016, pp. 60–63↩︎
Shulman 2010; Hanson 2016, pp. 171–174↩︎
Garfinkel 2021, §3↩︎
Shulman & Bostrom 2021↩︎
Cf. Bostrom 2019↩︎
Evans et al. 2021↩︎
Tomasik 2014↩︎
See, e.g., Sneddon et al. 2014↩︎
Brown et al. 2020↩︎
Replacement, reduction, and refinement.↩︎
Daswani & Leike 2015↩︎