This is a talk I recently gave at some workshops on moral philosophy in Zagreb, Oxford, and St. Andrews. It’s the last chapter of my PhD thesis, and it covers a lot of ground. The content has now become several different articles. This is the transcription of the talk, and the summarized version of that chapter. I hope people find it interesting!
Previously: Bibliography Review, Seven Takes, Beyond the Expanding Circle
Acknowledgements: Thanks to Hanno Sauer, Victor Kumar, Charlie Blunden, David Thorstad, Olivia Railton, Kate Vredenburgh, and a bunch of other people from Zagreb, Oxford, and St. Andrews for comments on my presentation. (Sorry I don’t remember everyone!)
Disclaimer on AI Use: Thanks to Claude for transcribing my talk into a written blog post. (And my apologies to the readers for the AI mannerisms that this will probably generate!)
Epistemic Status: I feel fairly confident about the main descriptive picture (moral progress as cumulative because it is institutionally stored and not always rediscovered from first principles). The latter half about AI moral decisionmakers, and especially on deferring to systems that might out-reason us, is much more speculative and future-oriented. I’m also still working out how to reply to all the objections that deontologists and virtue ethicists might have, which will probably take a paper on its own.
Most philosophical work on moral progress is backward-looking and somewhat too “psychological” or “psychologistic”. It explains the abolition of slavery, the extension of suffrage, or the recognition of LGBT rights as the historical unfolding of empathy, consistency reasoning, and/or an expanding circle of moral concern (Singer [1981] 2011; Campbell and Kumar 2012; Buchanan and Powell 2018; Kumar and Campbell 2022; Kitcher 2021). That story is not wrong, per se, but it feels to me disappointing and incomplete.
I want to propose a forward-looking and institutional alternative. On my view, moral progress accumulates and advances across the generations by offloading norms and values into stable formal and informal institutions, which then function as collective memory banks (like the “extended mind hypothesis”) for moral gains (which were often obtained through social struggle). Institutions and technologies store, transmit, and enforce the results of past moral struggles, much as writing stores our thoughts (for more on this, see Clark and Chalmers 1998; Risko and Gilbert 2016; Sterelny 2012).
If that is right, then the practical question is not whether to embed our morality in external structures, since we always done so for thousands of years, but which embeddings are good to adopt, and under what conditions. My answer is a slightly qualified but strong yes to moral offloading, including to forms that bypass rather than merely scaffold human judgment, such as adoptiong moral bioenhancement and AI moral assistance and decisionmaking. I will first present the techno-optimistic case, and towards the end I’ll talk about some caveats about how that offloading is only fully desirable when, for now, it preserves some open-ended moral engagement. So we should design our institutions and technologies with some (admittedly still vague) heuristics of contestability, legibility, and corrigibility, or we risk disabling ourselves from making further moral progress.
To start off, it helps to draw a contrast, even if it is a slightly stylized one.
There is a psychological view of moral progress, found in Singer’s The Expanding Circle ([1981] 2011), and Kohlberg’s Stages of Moral Development (1981), and elaborated through the theory of moral consistency reasoning, of “treating like cases alike” (as developed by Kumar and Campbell 2012, 2022). This view puts the engine of progress inside the individual mind, often appealing to metaphors such as an “escalator of reason”, as Singer calls it, that widens the circle of concern from narrow self-interest toward more impartial moral principles. It seems to have happened in the past 200 years, but the mechanism but seems very underspecified. It tells us the circle expanded, but struggles to say why, how, and how to expand it next.
In contrast, there’s a cultural-evolutionary view that locates the engine outside the individual, in the environment. Individuals don’t reinvent the moral lessons of history by isolated private reflection, in the same way that we don’t reinvent modern technology from scratch. And our evolved intuitions create moral inertia, since people rarely abandon long-held beliefs through argument alone. Instead, societies offload moral gains into norms, taboos, and law.
Coupled with modernization theory (Inglehart and Welzel 2003; Welzel 2013; Henrich 2020; Buchanan and Powell 2018; Wright 2000), the overall picture seems to be that rising safety, wealth, education, and peace, together with the breakup of dense kin networks in new urban centers, generated the rule of law and individual rights, which in turn moved societies toward the impartial prosociality of wider moral inclusion. The shifts largely happen as generations replace one another and as secular law, public schooling, and mass media now diffuse new moral norms.
One advantage of the institutional view is that it explains how moral progress has been cumulative. Contemporary morality is clearly not discovered alone by each generation. It is inherited, through education, institutions, laws, custom, and imitation. A child today does not have to re-derive the wrongness of chattel slavery from first principles, but they rather inherit it as a background fact about the world, encoded in law, schooling, and ordinary social norms.
Now, this is just a very simplified sketch, and mostly a matter of emphasis. The two perspectives are not, strictly speaking, exclusive, since moral reformers reason their way to ideals that then get encoded. But the views pick out different primary drivers of history, and a lot of current strategies for moral progress in the philosophical literature seems to assume that people will simply reason themselves into being morally better. This has, historically speaking, been pretty disappointing. The Stoics argued for some cosmopolitan principles, for example, but those principles didn't gain much moral uptake. And I think we are seeing a similar pattern with animals, where the argument for minimal consideration is already pretty persuasive, yet most people still aren’t moved to become vegetarian or vegan.
By moral offloading, I mean embedding moral norms into laws, customs, and tools so that better behavior becomes the default, cheaper, or mandatory, instead of relying on individual moral insight alone, or having to face harsh environments where moral behavior is punished (or suffer from free-riders). To put it simply, it’s a way of either automating a moral action, lowering its costs, or raising its benefits.
That way, each generation starts the moral race ahead and does not have to reinvent older morality, which frees our cognitive load and attention towards worrying about the next moral frontier. I think that, once a society has a morality that takes all humans to matter, it is far easier to build one that also takes non-human animals to matter. It’s hard to jump to universal moral concern if your moral circle is very narrow. Changes are usually relatively piecemeal and cumulative.
(One concern with offloading is that it is content-neutral, that is, it is an engine for moral change, rather than only for moral progress, as it preserves whatever we encode. Terrible atrocities, like apartheid or caste hierarchy were offloaded into culture, social norms, laws, etc. So whatever we are institutionalizing should also be the object of normative moral scrutiny.)
Offloading runs along a spectrum, from arrangements that leave us a lot of individual discretion to ones that take the decision out of our hands almost entirely, we can find:
Prima fracie, it seems to me that weaker forms preserve greater autonomy, but are also more fragile. Just try abolishing slavery by nudging. I think that probably won’t work. If the benefits to slaveholders are large enough, mere social shunning or raising costs a bit will not hold.
Stronger forms coordinate powerfully, and free cognitive bandwidth (much as GPS frees us from memorizing the city, or car routes), but also depersonalize, and they put more weight on getting the encoded content right.
There are often gains from hardening a weak offloading into a strong one, though this process introduces its own problems, such as rules written in absolutist deontological form “Do not kill”, rather than “Do not kill (except in circumstances X, Y, Z)” that don’t bend to exceptional circumstances.
A few philosophers have been studying moral progress, but you get surprisingly little on what makes such progress cumulative. But even the concepts available at each stage in history are often built out of the materials of the stage before. There is an order to it, and the order is not arbitrary, in much the way we do not get gears before the wheel or a metal knife before the stone tool.
Take the long arc of Western moral development. It goes from a kin-based, blood-feud order, where wrongs sit on the lineage rather than the person (Kitcher 2011), to individual responsibility before the law, which required an authority able to hold you rather than your family accountable.
Joseph Henrich (2020) argues that the medieval Church’s long campaign against cousin marriage helped produce this by eroding the dense kin networks that made clan justice viable. Individual responsibility then makes rule of law coherent (the same impartial rules for all, binding to individuals rather than kin groups), and codes like Magdeburg law installed that template wholesale across cities of strangers. Once people live as individuals with plural interests rather than as organs in an organic civic body, the social contract for mutual benefit becomes the framing of politics, and with it the standing of each member to ask “what is in it for me?”. This was very productive, as every later demand we associate with uncontroversial moral progress (suffrage, civil rights, welfare, minority rights) depends on that contractual frame, because each depends on the idea that members have standing to make claims on the polity as individuals.
Maybe a particular episode of ethics history can help make the claim more vivid. When Mary Wollstonecraft published A Vindication of the Rights of Woman (1792), Thomas Taylor answered with an anonymous satire, A Vindication of the Rights of Brutes, which applied her reasoning point for point to animals: if women have rights in virtue of sentience and the capacity to suffer, then so do dogs, horses, and pigs.
Thomas Taylor meant this as a reductio. He thought along the lines of: animals obviously have no rights, so, by modus tollens, Wollstonecraft’s premises must be false.
From two centuries’ distance, some of us would now run the argument the other way, as modus ponens: the capacities that ground women’s claims are not unique to women, so the case extends outward (yes I’m aware that this sounds is rude and dehumanizing to women, sorry about that!). Taylor mistook the implication for an absurdity because “animals have rights” was so far outside his moral environment (his Overton Window) that the inference looked like it had to lead to nonsense. The move was not unavailable by pure argument. The “moral niche” of 1792 society had not yet been built up to the point where it animal rights taken seriously for policymaking. (That build-up was precisely what Wollstonecraft and her successors were producing!) A demand can be framed by a lone thinker early, but it attracts serious uptake only once some underlying substrate is in place.
I think a lot of moral struggles follow this pattern, broadly speaking. The abolition of slavery produced the Thirteenth Amendment and Britain’s 1833 Act, and going back to slavery now beyond the pale.
Similarly, animal welfare moved from a marginal concern, to bans on dogfighting and humane-slaughter rules, with public attitudes also following the law, instead of only leading it.
I take to have established that institutions are the obvious carriers of offloaded morality, but technology does the same work, and its power is increasing.
Emerging technology is not a neutral tool we pick up after our values are fixed. Technology also reshapes the moral niche itself, the locally available set of choices, incentives, cues, and expectations we face. Danaher and Sætra (2023) identify roughly six mechanisms by which technology drives such technomoral change (Hopster et al. 2022): it adds options, changes costs and benefits, creates new relationships, shifts burdens and expectations, redistributes power, and reframes perception through new metaphors.
Two quick examples of how technology can affect moral progress. Dense, fast networks for documenting and coordinating around injustice are part of why the Arab Spring, Black Lives Matter, and #MeToo happened how they did, and it led to greater success than they would have had otherwise. It’s hard to imagine these movements existing without social media (Centola 2018, 2021).
And cheap, palatable cultured meat changes the choice set by adding a low-cost option, which weakens the standard excuse for not going vegan of “I can’t reasonably do otherwise”, smoothing the social path toward broad condemnation of factory farming. Most people are not turning vegans, but they might if meat alternatives become cheap and tasty (Anthis 2018; Milburn and Fischer 2022). Refusing cultured meat grounds of purity or disgust means paying for that purity in the continued suffering of billions of sentient animals.
Okay, let’s get to the more controversial part.
You can take everything I’ve said without committing to anything that follows, and you might even find the above relatively obvious and trivial. So let me put forward a more controversial and interesting thesis.
If offloading is the mechanism of cumulative progress, the natural question is what we should offload next. Here are a few candidates.
Moral bioenhancement is the proposal to use biomedical means to dampen unprovoked aggression and prejudice, raise empathy and perspective-taking, soften scope insensitivity, and shore up impulse control in the cases where people predictably harm others through weakness of will (Persson and Savulescu 2012, 2013). Our evolved moral psychology is, to put it gently, not built for a global, statistical world. (Bracket the feasibility question and the side effects for now. What is at stake is whether the thing would be good if it had no other ramifications.)
AI moral assistance extends the same spectrum to artificial moral cognition. I think it divides into two tiers:
I think people often answer these questions from the wrong starting point, which leads to status quo bias. They imagine our unassisted human decision as a neutral , pure baseline, and then ask whether an intervention is too intrusive. But the baseline already edited as a form of construction of our enviroment, as we saw!
And we keep making terrible decisions about animals, distant strangers, and future people.
There is a cost to refusing these tools, and it is often paid by someone other than the person doing the refusing. Historically, too, leaving moral change to happen on its own would have meant leaving people under slavery, patriarchy, caste, and autocracy for longer (cf. Kitcher 2021).
This could be a way of trying to patch familiar flaws in human moral psychology: parochial empathy, tribalism, scope insensitivity, identifiable-victim effects, our myopia about slow or statistical harms, and our difficulty in taking the standpoint of animals or future people.
Important Note: My argument in this section is conditional on solving alignment, and that such alignment is not deceptive. I am aware the alignment problem is a massive problem and existential risk. What I say here presupposes “the good future” where we get a very smart and powerful AI system that is aligned with, broadly speaking, Coherent Extrapolated Volition (Yudkowsky 2004; Bostrom 2014), or what it deems to be Moral Rightness (Bostrom 2014).
Let me push the frontier a bit further. If an AI system really were better at moral reasoning than we are (even combined as humanity), then in the highest-stakes settings the case for deferring our important moral decisions to it starts to look strong.
Aside from biases, memory and sheer cognitive capacity also matters a ton for moral decisionmaking! Moral decisions often require keeping track of a lot of more evidence, people, and possible consequences than any human can hold in our minds. We have very limited attention spans and memory. A much more capable system could take more of the morally relevant world into account at once.
These systems could also be more consistent. A person gets tired, irritated, distracted, hungry, or attached to one case because the victim has a name and a face. Even an excellent moral philosopher does so, they aren’t perfect. Whatever values it holds, such a system can apply them more evenly than a human deliberator who gets tired, forgetful, moody, and cranky.
None of this proves that a particular AI has good values, or that its verdict should be obeyed. My claim is conditional: if a system had a better grasp of the facts, fewer of our predictable distortions, and enough moral understanding to use that advantage, it could reason better than even our best human moral reasoners. At that point, continuing to insist that we make every important decision ourselves starts to look less noble.
But together, the conditional makes a case for the idea that, if a system were largely free of these distortions, far larger in capacity, more consistent, it would very plausibly outperform even our best moral reasoners, who are all running on the same low-grade evolved hardware.
Obviously, the conditional if in my argument is doing a lot of work. An opaque AI that confidently produces nonsense is not a moral authority. Nor should we hand moral authority to whichever company builds the most impressive model first, or whatever. There are risks of locking in bad values, of losing our own ability to think, and of giving enormous power to whoever chooses the system’s objectives.
This distinction is where I think a lot of the resistance comes from. People might be fairly happy with an AI that gives advice (hell, a lot of people already ask AIs as therapists, for job advice, for relationship advice, and so on). They get much less comfortable when it makes the call, or stops them from making one.
I understand that intuitive reaction. But I am not sure the line between those cases carries as much moral weight as people think. We already let laws, education, social pressure, and the design of our everyday objects shape what we do, promoting some forms of behavior over others, even making particular actions difficult or impossible.
If the intervention prevents something morally terrible, and if the people affected by that terrible thing get a say in the moral comparison, why should the fact that I no longer get to make the call settle it? Why is that more weighty than getting the moral decision right? I think it simply isn’t!
(This is where consequentialists and some deontologists or virtue ethicists will part ways. I cannot resolve that disagreement in a Substack post. For now I just want to put some pressure on the intuition. My question is whether that aversion is justified, and my tentative answer is that it actually is much weaker than it feels. But I do think their resistance to being moral bypassed needs more argument.)
Objection 1: The bioconservative worry. Some people, like Kass (1997) and Sandel (2004), say things along the lines of “don’t tamper with nature”, or “nature is wise”, or stuff like that. I don’t find that convincing. Evolution didn’t design our psychology to be morally admirable. It gave us a decent capacity to cooperate with people close to us, but along with it came tribalism and a startling ability to ignore suffering when we cannot see it.
A thought experiment: If you had to choose your moral psychology from a veil of ignorance perspective, that is, without knowing whether you would be the person making a decision or one of the beings harmed by it, would you really choose our current package of biases and cognitive limitations? That seems implausible to me, and normatively unjustifiable. So the stance reduces to status quo bias plus just-world bias (Bostrom and Ord 2006) against Kass 1997 and Sandel 2004).
The companion worry, that “offloading destroys autonomy”, fails for a related reason: our agency is already scaffolded everywhere, by biology, upbringing, schooling, media, law, shame, moral rules, and so on. And refusing new scaffolds does not free us to a pristine state of nature, but just leaves the old (often inferior ones) in place.
Anyone who objects specifically to putting something in your brain or changing our biology still owes us a clean account of what exactly distinguishes the external versions from the internal ones.
And besides, an external AI advisor or decisionmaker does not change anyone’s biology.
Objection 2: Don’t we need social struggle? Philip Kitcher, in The Ethical Project and Moral Progress (2011, 2021), and Rahel Jaeggi, in Progress and Regression (2025), argue that moral progress is problem-driven. Broadly speaking, failures, friction and conflict are what make injustice salient, so if we automate the friction away, some wrongs may never become visible. I agree that there is historical truth here, but I don’t think it follows that suffering and struggle are always needed. If we can discover and fix an injustice without making people fight through decades of it, that seems even better!
Also, these tools might help us notice more injustice. An AI that shows you the otherwise invisible effects of your choices on distant people or animals could make us less complacent, not more. The thought that every generation should have to rediscover why slavery is wrong before it can move on seems like a terrible use of collective attention.
Objection 3: Mill’s "dead dogma" objection. John Stuart Mill said in On Liberty (1859; Chapter 2) that “A true belief that is not fully and frequently discussed decays into prejudice”, that even a true belief can become something people just parrot without understanding it. There are two failure modes here:
I think this is a risk if we make moral progress too automatic. People still need to be able to question our rules and explain why they exist. But we do not need to leave an injustice in place just to keep the argument about it alive.
I take this seriously in the next section. But notice that it is an argument about how to offload, not about whether to.
Objection 4: The interiority objection (from deontology and virtue ethics). One key objection from deontologists and virtue ethicists is about what we might lose inside ourselves. If an AI removes every opportunity to do the wrong thing, we may also lose opportunities to do the right thing. A person who never has to make a difficult moral decision might is then worse at being at being a moral agent, or a friend, partner, or parent. We might lose something in the social practice of working these things out together and holding each other responsible.
I think that I could simply grant that there is a cost here (although, as a consequentialist, I don’t think it’s super pressing). But I just do not think it is remotely large enough to justify keeping moral atrocities around.
Nobody seriously thinks we should have left slavery in place so that future abolitionists could develop moral courage in the 21st century, right? Similarly, factory farming should not continue so that I get to feel virtuous for being vegetarian.
Because if that were so, some deontological and virtue views end up wanting the world to contain at least some injustice so that good people have something to heroically oppose. That seems straightforwardly odd, or backwards. Historically, moral reformers wanted to end cruelty, and nobody really thinks the abolitionists should have eased off so that later generations could enjoy the “moral workout” of opposing slavery.
If we can get rid of an enormous wrong through better institutions or technology, I think we should. There will still be plenty of ordinary occasions to practice generosity, care, and judgment. We do not need billions of animals suffering to keep some kind of “moral gym” open to exercise our “decision muscles”.
Perhaps the worry gets stronger as we move from the enormous harms to medium stakes decisions, and this is where consequentialists on one side, and deontologists and virtue ethicists on the other, part ways. If we offload absolutely every moral decision, perhaps I stop developing the judgment I need in ordinary relationships, with terrible outcomes long-term. That is a reason to be selective about what we delegate. (But it is not a reason to make vulnerable beings pay for our character development!)
Some moral gains are much less secure than they feel. Even the most basic tenets of liberal democracy, for example, are not a settled achievement everywhere, and might tremble with the rise of new technologies such as AGI.
People need to keep understanding why things like independent courts matter, why an executive power should face limits, and why unpopular minorities (like religious minorities, or criminals) have basic rights. A generation that has never experienced a dictatorship can inherit liberal democratic institutions without understanding that they are better than totalitarian alternatives. So we cannot just “install the right institutions”, forget about them, and assume the reasoning behind them will look after itself.
We can imagine the same problem with AI decisionmaking or moderation. If a system silently catches every piece of dehumanizing speech before anyone sees it, people may never have to explain why hate speech is objectionable. So these things come with a cost. So I would want a few conditions in place, especially for anything as powerful as advanced AGI:
And there are other practical risks too. A technology could be unsafe, available only to rich people, or captured by a government. An immortal dictator could keep enforcing the values of his youth for thousands of years. Bioenhancement could turn into another status race to put their kids into the best universities or other status symbols (comparative goods, rather than absolute ones).
These are reasons to care a great deal about which tools we build, how we govern them, and where we use them.
Institutions and technology are the loom on which moral progress is woven. By embedding hard-won gains into custom, law, and tools, a society constructs a moral niche that outlives any generation, so that each cohort inherits a richer moral world than the last and is freed from relearning old lessons to work on the questions that remain open. This offloading has been central to eradicating past injustices and enabling new gains, and I have argued the process should continue, including through technologies that bypass rather than merely scaffold agency, such as bioenhancement and AI moral assistance.
But progress can be fragile. Unchallenged values can decay into dogma, and contingent choices taken early could harden into arbitrary path-dependence for future generations.
So the design rules are tough, because they pull in different directions. They suggest to offload aggressively to lock in the moral progress we have won, but also to engineer the future for epistemic values of the anti lock-in values of transparency, contestability, and amendment.
Do both, however, and we may leave the future an even richer moral inheritance than the one we personally received.
Anthis, Jacy Reese. 2018. The End of Animal Farming: How Scientists, Entrepreneurs, and Activists Are Building an Animal-Free Food System. Lantern Books.
Arendt, Hannah. 1963. Eichmann in Jerusalem: A Report on the Banality of Evil. Viking Press.
Bostrom, Nick. 2014. Superintelligence: Paths, Dangers, Strategies. Oxford: Oxford University Press.
Bostrom, Nick, and Toby Ord. 2006. “The Reversal Test: Eliminating Status Quo Bias in Applied Ethics.” Ethics 116 (4): 656–679.
Buchanan, Allen. 2020. Our Moral Fate: Evolution and the Escape from Tribalism. Cambridge, MA: MIT Press.
Buchanan, Allen, and Russell Powell. 2018. The Evolution of Moral Progress: A Biocultural Theory. Oxford University Press.
Campbell, Richmond, and Victor Kumar. 2012. “Moral Reasoning on the Ground.” Ethics 122 (2): 273–312.
Centola, Damon. 2018. How Behavior Spreads: The Science of Complex Contagions. Princeton University Press.
Centola, Damon. 2021. Change: How to Make Big Things Happen. New York: Little, Brown Spark.
Choi, Jung-Kyoo, and Samuel Bowles. 2007. “The Coevolution of Parochial Altruism and War.” Science 318 (5850): 636–640.
Clark, Andy, and David Chalmers. 1998. “The Extended Mind.” Analysis 58 (1): 7–19.
Danaher, John, and Henrik Skaug Sætra. 2023. “Mechanisms of Techno-Moral Change: A Taxonomy and Overview.” Ethical Theory and Moral Practice 26 (5): 763–784.
Desvousges, William H., F. Reed Johnson, Richard W. Dunford, Kevin J. Boyle, Sara P. Hudson, and K. Nicole Wilson. 1992. Measuring Non-Use Damages Using Contingent Valuation: An Experimental Evaluation of Accuracy. Research Triangle Institute Monograph 92-1. Research Triangle Park, NC: RTI Press.
Durkheim, Émile. (1893) 2014. The Division of Labour in Society. Edited by Steven Lukes. Translated by W. D. Halls. New York: Free Press.
Giubilini, Alberto, and Julian Savulescu. 2018. “The Artificial Moral Advisor. The ‘Ideal Observer’ Meets Artificial Intelligence.” Philosophy & Technology 31 (2): 169–188.
Habermas, Jürgen. 1996. Between Facts and Norms: Contributions to a Discourse Theory of Law and Democracy. Translated by William Rehg. Cambridge, MA: MIT Press.
Henrich, Joseph. 2020. The WEIRDest People in the World: How the West Became Psychologically Peculiar and Particularly Prosperous. New York: Farrar, Straus and Giroux.
Hopster, Jeroen K. G., Chirag Arora, Charlie Blunden, Cecilie Eriksen, Lily Frank, Julia Hermann, Michael Klenk, Elizabeth O’Neill, and Steffen Steinert. 2022. “Pistols, Pills, Pork and Ploughs: The Structure of Technomoral Revolutions.” Inquiry 68 (2): 264–296.
Inglehart, Ronald. 2018. Cultural Evolution: People’s Motivations Are Changing, and Reshaping the World. Cambridge: Cambridge University Press.
Inglehart, Ronald, and Christian Welzel. 2005. Modernization, Cultural Change, and Democracy: The Human Development Sequence. Cambridge University Press.
Jaeggi, Rahel. 2025. Progress and Regression. Translated by Robert Savage. Cambridge, MA: Harvard University Press.
Kass, Leon R. 1997. “The Wisdom of Repugnance: Why We Should Ban the Cloning of Humans.” The New Republic 216 (22): 17–26.
Kitcher, Philip. 2011. The Ethical Project. Cambridge, MA: Harvard University Press.
Kitcher, Philip. 2021. Moral Progress. Oxford University Press.
Klenk, Michael, Elizabeth O’Neill, Chirag Arora, Charlie Blunden, Cecilie Eriksen, Lily Frank, and Jeroen Hopster. 2022. “Recent Work on Moral Revolutions.” Analysis 82 (2): 354–366.
Kohlberg, Lawrence. 1981. The Philosophy of Moral Development: Moral Stages and the Idea of Justice. Vol. 1. San Francisco: Harper & Row.
Kumar, Victor, and Richmond Campbell. 2022. A Better Ape: The Evolution of the Moral Mind and How It Made Us Human. Oxford University Press.
MacIntyre, Alasdair. 1984. After Virtue: A Study in Moral Theory. 2nd ed. University of Notre Dame Press.
Milburn, Josh, and Bob Fischer. 2022. “Plant-Based and Cultivated Meat and the Demands of Morality.” In The Routledge Handbook of Animal Ethics, edited by Bob Fischer, 524–536. New York: Routledge.
Milgram, Stanley. 1974. Obedience to Authority: An Experimental View. Harper & Row.
Mill, John Stuart. (1859) 2003. On Liberty. Yale University Press.
Morris, Ian. 2015. Foragers, Farmers, and Fossil Fuels: How Human Values Evolve. Princeton: Princeton University Press.
Neurath, Otto. (1932) 1983. “Protocol Statements.” In Philosophical Papers 1913–1946, edited by M. Neurath and R. S. Cohen. Dordrecht: D. Reidel.
Persson, Ingmar, and Julian Savulescu. 2012. Unfit for the Future: The Need for Moral Enhancement. Oxford University Press.
Persson, Ingmar, and Julian Savulescu. 2013. “Getting Moral Enhancement Right: The Desirability of Moral Bioenhancement.” Bioethics 27 (3): 124–131.
Petersen, Steve. 2017. “Superintelligence as Superethical.” In Robot Ethics 2.0: From Autonomous Cars to Artificial Intelligence, edited by Patrick Lin, Keith Abney, and Ryan Jenkins, 322–337. New York: Oxford University Press.
Quine, Willard Van Orman. 1969. “Epistemology Naturalized.” In Ontological Relativity and Other Essays, 69–90. Columbia University Press.
Rawls, John. 1971. A Theory of Justice. Cambridge, MA: Harvard University Press.
Risko, Evan F., and Sam J. Gilbert. 2016. “Cognitive Offloading.” Trends in Cognitive Sciences 20 (9): 676–688.
Sandel, Michael J. 2004. “The Case against Perfection.” The Atlantic 293 (3): 51–62.
Sauer, Hanno. 2023. Moral Teleology: A Theory of Progress. New York: Routledge.
Singer, Peter. (1981) 2011. The Expanding Circle: Ethics, Evolution, and Moral Progress. Princeton University Press.
Slovic, Paul. 2007. “‘If I Look at the Mass I Will Never Act’: Psychic Numbing and Genocide.” Judgment and Decision Making 2 (2): 79–95.
Small, Deborah A., George Loewenstein, and Paul Slovic. 2007. “Sympathy and Callousness: The Impact of Deliberative Thought on Donations to Identifiable and Statistical Victims.” Organizational Behavior and Human Decision Processes 102 (2): 143–153.
Sterelny, Kim. 2012. The Evolved Apprentice: How Evolution Made Humans Unique. Cambridge, MA: MIT Press.
Sunstein, Cass R. 2014. Why Nudge? The Politics of Libertarian Paternalism. New Haven: Yale University Press.
Taylor, Thomas. 1792. A Vindication of the Rights of Brutes. London: Edward Jeffery.
Thaler, Richard H., and Cass R. Sunstein. 2008. Nudge: Improving Decisions about Health, Wealth, and Happiness. New Haven: Yale University Press.
Verbeek, Peter-Paul. 2011. Moralizing Technology: Understanding and Designing the Morality of Things. Chicago: University of Chicago Press.
Weber, Max. (1922) 1978. Economy and Society: An Outline of Interpretive Sociology. Edited by Guenther Roth and Claus Wittich. Berkeley: University of California Press.
Welzel, Christian. 2013. Freedom Rising: Human Empowerment and the Quest for Emancipation. New York: Cambridge University Press.
Wollstonecraft, Mary. 1792. A Vindication of the Rights of Woman. London: Joseph Johnson.
Wright, Robert. 2000. Nonzero: The Logic of Human Destiny. New York: Pantheon Books.
Yudkowsky, Eliezer. 2004. “Coherent Extrapolated Volition.” Singularity Institute.
Facts Only
* Presentations were delivered at workshops in Zagreb, Oxford, and St. Andrews.
* The content originates from the final chapter of a PhD thesis.
* The transcription process utilized the AI tool Claude.
* Peter Singer's "The Expanding Circle" (1981) and Lawrence Kohlberg's "Stages of Moral Development" (1981) are cited as psychological views of moral progress.
* Joseph Henrich's 2020 work on the medieval Church's campaign against cousin marriage is cited as a driver for individual responsibility.
* Mary Wollstonecraft published "A Vindication of the Rights of Woman" in 1792.
* Thomas Taylor published "A Vindication of the Rights of Brutes" in 1792.
* The 13th Amendment and Britain’s 1833 Act are cited as institutional offloadings of the abolition of slavery.
* Persson and Savulescu (2012, 2013) proposed moral bioenhancement.
* John Stuart Mill published "On Liberty" in 1859.
Executive Summary
Moral progress is conceptualized not as a series of individual psychological breakthroughs, but as a cumulative institutional process. By "offloading" moral gains into laws, customs, and technologies, societies create collective memory banks that prevent each generation from having to rediscover ethical principles from first principles. This institutional approach explains why progress is cumulative; individuals inherit a "moral niche" that lowers the cost of prosocial behavior and raises the cost of injustice.
The framework extends this logic to emerging technologies, suggesting that moral bioenhancement and AI-driven decision-making could further automate ethical behavior or correct human cognitive biases. While this offers a path toward reducing systemic suffering—such as that found in factory farming—it introduces risks of moral atrophy, "dead dogma," and the loss of human autonomy. The proposal advocates for a design philosophy based on contestability, legibility, and corrigibility to ensure that while moral gains are locked in, the capacity for future moral evolution remains open.
Full Take
This work utilizes ACADEMIC MODE, presenting a theoretical framework for moral evolution. The methodology is primarily a conceptual synthesis of cultural evolution, modernization theory, and philosophy of mind (specifically the extended mind hypothesis). The author explicitly identifies the second half of the thesis—concerning AI moral decision-makers—as speculative and future-oriented, which serves as a necessary epistemological hedge.
The central claim—that moral progress is institutional rather than purely psychological—is well-supported by the cited transition from kin-based justice to individual legal responsibility. However, a peer reviewer would likely flag the "AI moral assistance" section for a leap in proportionality: the transition from "laws as offloading" to "AI as moral authority" assumes that alignment (specifically Coherent Extrapolated Volition) is a solved or solvable problem. The author acknowledges this as a conditional prerequisite, but the normative weight of the argument rests entirely on this unproven technical assumption.
The real-world implication is a shift in the locus of moral agency. If we move from "scaffolding" (tools that help us think) to "bypassing" (tools that decide for us), we risk a systemic fragility where the "why" of morality is lost to the "how" of automation.
Bridge Questions:
1. How can a system be designed to be "corrigible" if the agents using it have offloaded the very reasoning capacity required to identify the need for correction?
2. Does the "moral gym" argument—that we need struggle to develop virtue—hold weight if the "exercise" is performed at the cost of sentient suffering?
Counterstrike Scan: A bad actor pushing this narrative would use it to justify the surrender of human oversight to "objective" AI systems to consolidate power. The actual content avoids this by emphasizing contestability and transparency. Clean.
