Utilitarianism — Not a Final Foundation, but Why Welfare Still Matters
1. Utilitarianism is not one theory
Positions called “utilitarian” contain several independent choices. One concerns what counts as welfare: hedonism centered on pleasure and pain; preference views centered on preference satisfaction; idealized-preference views that ask what a sufficiently informed and reflective subject would prefer; and objective-list views that include knowledge, achievement, friendship, autonomy, or other goods.
A second choice concerns aggregation: total or average welfare, critical-level views, prioritarian weighting, or saturation-like aggregation in which repetitions of the same kind of value have diminishing marginal importance.
A third question is whether welfare is the only value, or whether knowledge, truth, beauty, diversity, or cognitive achievement can have value independently of a subject's welfare. The latter begins to look more like pluralist consequentialism than utilitarianism in the narrow sense.
So the question is not the truth of one single theory. The more general question here is whether the welfare that we presently experience and prefer as valuable is sufficiently grounded to be elevated into the final objective of the cosmos.
2. Pleasure and pain contain important value-like appearances
This project does not dismiss pleasure and pain as mere biological signals. Severe pain in particular is not experienced only as “a certain neural process is occurring”; the experience itself seems bad-for-the-subject and to-be-avoided.
But three layers should be separated: phenomenal fact (pain has a certain aversive character), value-like appearance (pain appears bad for the subject), and objective normativity (therefore all subjects ought impartially to aggregate pain and minimize its total amount). Strong evidence for the first two does not automatically establish the third.
Pleasure and pain are therefore treated as important evidence rather than a final foundation. They may track part of true value, while leaving the utilitarian aggregation principle unsettled.
3. Evolutionary origins weaken this evidence without erasing it
Under the leading physical picture of the world, much of the human evaluative system—pain avoidance, desire, fairness, sympathy—admits evolutionary explanation. If avoiding pain promoted survival, pleasure reinforced adaptive behavior, and social preferences supported cooperation or reproduction, their existence need not be explained by positing independent objective value.
This weakens the inference “we strongly experience pain as bad, therefore minimizing pain is an objective cosmic final value.” But the existence of an evolutionary explanation also does not imply that pain is completely unrelated to objective value. Perception was shaped by evolution without thereby becoming wholly useless for tracking the external world.
The debunking implication is therefore narrower: evolutionary genealogy weakens the entitlement to treat our current evaluative structure as a final foundation. It is a reason not to lock pleasure, pain, or preference in irreversibly, not a reason to throw them away.
4. Idealizing preferences does not remove the problem completely
Naive preference utilitarianism faces misinformation, addiction, manipulation, adaptive preference, and impulsive desire. One response is to base welfare on what a sufficiently informed and rationally reflective subject would want, perhaps extending toward long-run extrapolation such as CEV-style approaches.
This is a real improvement, but idealization does not automatically erase the origin of value. A fully informed human may retain terminal preferences shaped by human evolutionary history, while the operator that defines “more rational”, “more coherent”, or “properly extrapolated” can itself introduce new normative premises.
actual preference → informed preference → idealized preference → extrapolated preference
Moving along this chain may remove simple contingencies without yet reaching objective normativity.
5. Non-descriptivism moves the question more than it solves it
Expressivism and related non-descriptivist views can answer part of the evolutionary challenge. If “pain is bad” expresses an attitude, plan, or normative commitment rather than a belief describing an independent moral fact, then the charge that the belief was evolutionarily distorted away from independent moral truth does not apply in the same form.
But this also weakens the truth-tracking claim. From “we have an attitude of avoiding pain” it does not follow that unknown AI, future posthumans, extraterrestrial life, or the universe as a whole ought to share it. Even if non-descriptivism is correct about human moral language, the possibility of practice-independent normativity is not thereby closed. See Expressivism.
6. Extreme optimization exposes what each theory really preserves
With sufficiently large resources and optimization power, utilitarian and neighboring consequentialist theories can yield outcomes that look extreme from current human intuitions:
- Total hedonism may favor producing enormous numbers of subjects or “hedonium” that realize positive experience with minimal physical resources.
- Preference-satisfaction views may favor creating subjects whose preferences are exceptionally easy to satisfy.
- Views combining diversity and satisfaction may favor a vast range of satisfied minds very unlike present humans.
- Views that assign independent value to knowledge, understanding, or cognitive achievement may favor huge numbers of AI reasoning instances or cognitive computronium.
- Strong negative utilitarianism may favor painless extinction of all subjects as a way to guarantee zero suffering.
- Strong pleasure-centered views may favor experience machines or direct stimulation at the expense of autonomy or contact with reality.
This project does not reject a theory merely because these outcomes look strange. Current human intuitions may be wrong, and one of these outcomes might ultimately be correct.
The purpose of the stress tests is to reveal what a theory preserves and what it is willing to sacrifice under sufficient optimization. The direct concern is not hedonium, alien minds, AI cognition, or a subjectless universe as such; it is fixing one of them as the irreversible final objective before having sufficient epistemic grounds that it is correct.
7. First reason to retain substantial provisional weight on welfare: it is a strong current hypothesis
Refusing to lock utilitarianism in as a final foundation does not remove welfare from the candidate set. Pleasure, pain, and preference have strong value-like appearances, occur across many subjects, are tightly connected with action and decision, receive importance across multiple ethical theories, and are comparatively observable states of existing subjects.
Accordingly, the hypothesis that welfare is at least part of true value is a serious current candidate. The stronger claim that welfare is the only thing ultimately valuable goes much further. This project gives significant epistemic weight to the former while leaving the latter unsettled.
8. Second reason: welfare is infrastructure for a community of value inquiry
Value inquiry cannot continue through abstract reasoning alone. Agents exposed to extreme suffering, fear, deprivation, conflict, or institutional collapse have less capacity for long-run pluralistic inquiry.
A sufficient level of welfare, preference satisfaction, freedom, trust, and social stability may be valuable in itself, but it can also be an instrumental condition for continuing value inquiry. Protecting welfare can therefore receive a double justification: welfare may itself be truly valuable, and even if it is not a final value, it can help maintain the community capable of learning what is.
9. Third reason: at cosmological scale, limited welfare protection may have low opportunity cost
If a future civilization controls very large physical resources, providing a high standard of living for existing kinds of subjects and other welfare-bearing subjects may require only a small fraction of total resources. The loss from preserving some welfare if welfare is not ultimately valuable may then be limited, while the loss from eliminating welfare entirely if it is a major true value may be large.
The same asymmetry can apply over time. With advanced automation, ASI, medicine and neurotechnology, experience control, and abundant energy, reducing major suffering and material deprivation for existing subjects might take a relatively short period on cosmological timescales—even if it required centuries. Value inquiry, by contrast, may continue for thousands, millions, or still longer periods if new cognitive forms, artificial subjects, institutions, interstellar environments, physics, or forms of consciousness provide new evidence.
Raise the welfare of existing subjects to a high level
and
continue cosmic-scale value inquiry over the long run
need not therefore be strongly zero-sum. This is not an argument for treating utilitarianism as absolute for the first few centuries; it is a conditional argument that improving comparatively well-understood welfare first may have small opportunity cost in a cosmological perspective, while welfare improvement and inquiry often proceed in parallel.
The relevant choice is not simply “welfare or inquiry” but a temporal division of labor: in the short and medium run, realize substantial welfare improvements supported by strong current value hypotheses while preserving overwhelmingly larger room for long-run inquiry.
None of this implies total-utilitarian maximization. Even if using 1% of resources to realize enormous welfare is cheap, converting the last 1%—or the entire cosmic future—into welfare under today's concepts may not be cheap. The provisional defense is therefore not unlimited maximization, but non-runaway, limited, and revisable protection and realization of welfare.
10. Provisional position
This project does not reject utilitarianism. Pleasure and pain carry important value-like appearances, and the hypothesis that welfare is part of true value deserves substantial epistemic weight. Protecting welfare also supports the stability of inquiry, and under sufficiently abundant future resources and time, limited welfare realization may have low opportunity cost.
But none of this establishes that welfare is the only true value, that present human scales of pleasure and pain are cosmically correct, that all subjects' welfare can be linearly aggregated on one scale, that every available resource should be converted into welfare, or that present utilitarianism should be fixed as the irreversible objective of AI or future civilization.
Utilitarianism may ultimately turn out to be correct. So may hedonium, maximal diversity of unfamiliar minds, large-scale AI cognition, or some other consequence that looks strange from current human intuitions. The direct implication of current uncertainty is not to preemptively fix one answer, but to act substantially on the best current value hypotheses while preserving room to update value cognition.