AI Value Exploration Notes
Exploration

Infinite Ethics and Runaway Inquiry — Paralysis, Fanaticism, and Epistemic Meta-Practice

Can we leave infinity and extreme stakes open without letting a low-confidence value hypothesis capture inquiry and civilization as a whole?

Exploration v0.5 · English translation · 2026-08-31

Working position: Do not assign probability zero to infinite worlds or infinite value. But do not automatically compress rival value and metaethical hypotheses into one common currency of “credence × magnitude of good according to that theory.” Separate an epistemic layer that forms epistemic weights from evidence, a reflective meta-practice layer that allocates E (active inquiry), P (preserving future inquiry), and X (realizing present value) and governs irreversible choices, and a hypothesis-relative practice layer that gives live hypotheses provisional representation inside the divisible ordinary-practice portion of X in response to epistemic weight. Magnitudes of good or benefit are not discarded; they primarily guide action within each hypothesis.

1. Infinite ethics contains at least four distinct problems

Treating infinite ethics as one numerical pathology conflates different questions. This page separates at least four layers:

  1. Infinitarian paralysis: if the world already contains positive and negative infinite value, finite actions can appear unable to change total value.
  2. Ranking infinite worlds: Pareto, additivity, anonymity, transitivity, and completeness become difficult to preserve together when extended from finite to infinite domains.
  3. Fanaticism: literal infinity is unnecessary; tiny probabilities multiplied by extremely large outcomes can dominate decision-making.
  4. Contagion under moral uncertainty: a low-credence normative theory claiming infinite good or bad can swamp other theories, or its incomparability can infect the all-things-considered ranking.

The project’s central worry—that bizarre value hypotheses could seize all inquiry resources—is closer to the third and fourth problems than to literal infinity alone.

2. Bostrom and infinitarian paralysis

Nick Bostrom’s Infinite Ethics (2011) argues that modern cosmology makes it difficult to dismiss the possibility of infinitely many value-bearing locations, creating severe problems for aggregative ethics. If infinitely many positive and negative contributions already exist while our intervention changes only finitely many locations, standard cardinal arithmetic can make finite improvements disappear at the level of total value.

The key distinction is between the finitude of our action and the finitude or infinitude of the evaluatively relevant world. We do not currently know that the world is infinite in the relevant sense. Failure of comparison inside an infinite-world hypothesis therefore need not erase a clear improvement under finite-world hypotheses.

Keep finitude explicitly live: “A and B are incomparable if the world is infinite” does not entail “A and B are incomparable across all admissible world models.” If A clearly improves on B in finite worlds and the infinite-world model supplies no positive reason to prefer B, the finite-world difference can remain decision-relevant.

This does not solve the internal mathematics of infinite worlds. It is an outer-layer treatment of uncertainty that denies one unresolved cosmological hypothesis automatic veto power over all action.

3. Vallentyne–Kagan and the temptation to extend finite aggregation

Peter Vallentyne and Shelly Kagan’s “Infinite Value and Finitely Additive Value Theory” (1997) is a canonical attempt to understand how principles that work over finite domains might extend to infinitely many locations. The motivation is compelling: local or finite-part improvements seem as though they should still matter in infinite worlds.

This project, however, need not now lock a complete extension of finite aggregation into practice when our confidence in remote cosmology is low. Our uncertainty concerns not just whether remote domains exist, but also how many subjects there are, whether they are conscious, what counts as welfare or value, and what causal structures obtain there.

We can preserve strongly supported comparisons in finite or well-understood domains, retain only partial Pareto-like comparisons in infinite domains, and leave the rest unresolved. The point is not to ignore infinity but to avoid locking in extra axioms merely to force a complete ranking over a poorly understood domain.

4. Askell: incomparability may become pervasive

Amanda Askell’s dissertation Pareto Principles in Infinite Ethics (2018) sharpens the problem. Combining Pareto principles, transitivity, principles concerning permutations, and qualitative consistency can generate pervasive incomparability between infinite worlds.

This is not merely an arithmetic failure such as “∞−∞ is undefined.” It is a structural conflict among principles that look jointly natural in finite settings but become difficult to preserve together with complete comparability in infinite settings.

Response here: when incomparability appears, do not assume that a new scalar metric must immediately restore a complete ranking. Especially for low-confidence remote worlds, preserving a partial order can itself be an expression of epistemic restraint.

5. Fanaticism is a general problem of decision theory, not a special feature of ethics

The same structure arises with personal benefit. St. Petersburg-style cases show that tiny probabilities and unbounded utility can generate divergent or undefined expectations without any appeal to moral good.

One traditional response is to bound utility. But in ethics this can amount to the substantive claim that once enough good already exists, additional good has diminishing decision-theoretic importance that eventually saturates. If we wish to retain unbounded value, some other part of the theory must give.

Zachary Goodsell’s work on decision theory with unbounded utility explores how rational choice can coexist with unbounded utility by reconsidering which finite-case principles should be extended without modification to countably infinite settings. The important lesson is that accepting unbounded utility is not the same as accepting every infinite extension of familiar finite decision axioms.

6. Risk aversion and probability cutoffs are not enough by themselves

Lara Buchak’s risk-weighted expected utility treats risk attitude as distinct from belief and utility. It is therefore an important precedent against the claim that linear probability-weighted utility is the only rational aggregation rule. Yet if utility is sufficiently unbounded, risk weighting alone need not prevent very large payoffs from eventually swamping tiny probabilities.

A different response simply treats probabilities below some threshold ε as zero. But fixed cutoffs can depend on how events are partitioned and fail to distinguish a precisely known physical probability of 10^-20 from a largely speculative 10^-20 attached to an exotic value hypothesis.

For this project, the relevant distinction is therefore between low probability and low epistemic confidence in the probability assignment itself. For remote cosmology and unknown value, imprecise probability and ambiguity may be better models than a single sharp number.

7. Tarsney: ordinary expected value versus Pascalian cases

Christian Tarsney’s “Expected Value, to a Point” (2025) argues that under large background uncertainty independent of our action, options with higher expected value often stochastically dominate in ordinary-scale cases, while this connection can fail when expected-value differences are driven by tiny probabilities of extreme outcomes.

This supports a middle position: approximately expected-value reasoning in ordinary domains, but no automatic obligation to follow EV in Pascalian domains, without imposing an arbitrary probability cutoff.

At the same time, “Against Anti-Fanaticism” (2025) argues that strong anti-fanaticism conflicts with other plausible decision principles. So a blanket rule to ignore sufficiently small probabilities is also unstable. Allowing incomplete preference in extreme cases fits this page’s willingness to leave some comparisons unresolved.

8. Moral uncertainty turns stakes into a cross-theory contagion problem

MacAskill, Bykvist, and Ord’s Moral Uncertainty (2020) examines fanaticism and infectious incomparability under Maximize Expected Choice-Worthiness. A low-credence theory can dominate if it assigns enormous choice-worthiness differences, while incomparability can spread into the all-things-considered ranking.

The problem is deeper here because uncertainty extends beyond first-order ethics to realism and error theory themselves. If a realism branch claims that one act has cosmically enormous importance, multiplying that internal stake by a small credence can allow a speculative theory to capture inquiry and practice.

Rejected compression: do not assume that “credence in a theory × magnitude of good according to that theory” is automatically a common currency across fundamentally different value theories. This does not ignore stakes; it rejects the extra premise that their cardinal magnitudes are intertheoretically comparable.

9. Does removing cross-theory stakes abandon rational choice?

If probabilities over world states and a single utility function are already settled, then allocating by probability while ignoring payoff magnitude departs from standard expected-utility rationality. This page does not deny that ordinary result.

Under value uncertainty, however, what should count as utility is itself disputed, and it may be unclear whether “100” under U1 and “10^100” under U2 belong to a shared cardinal scale. What is rejected here is not within-theory stakes but the automatic common scaling of stakes across rival value theories.

Within each hypothesis, expected utility, duty, preference satisfaction, agreement, constructed value, or any other internally endorsed criterion may still fully guide action.

10. Epistemic layer: separate truth-likelihood from stakes

At the most upstream layer, evidence E is used to form and update confidence C(Hi|E) in world, value, and metaethical hypotheses. The fact that a hypothesis would be highly desirable or terrifying if true does not by itself make that hypothesis more likely.

Scientific belief works similarly. A universe rich in exploitable resources would be convenient, but convenience is not evidence for that cosmology. Likewise, “if this value theory is true, infinitely much good is at stake” does not by itself increase credence in the theory.

Epistemic priority: the upstream question is first “how strongly is this hypothesis supported by evidence and reasoning?” rather than “how much benefit would this hypothesis promise if true?”

11. Splitting and merging theories: track credence mass, not labels

A proportional representation scheme should not let a theory gain influence merely by being redescribed as one hundred separate labels. Representation should track credence mass over mutually exclusive hypotheses, not the raw number of theory names.

If a coarse hypothesis H is merely partitioned descriptively into H1, H2, …, then in principle C(H)=ΣC(Hj), so the total representation of the family does not increase. Splitting utilitarianism into one hundred synonymous labels does not create one hundred times the confidence in utilitarianism.

If finer hypotheses genuinely differ in evidential predictions or practical recommendations, changed internal allocation is appropriate. That is not vote inflation but recognition of decision-relevant internal uncertainty.

12. Inquiry is not merely one value-theory branch among others

The previous draft treated inquiry resources as though they were simply allocated among value theories. But the existing logic of Practice and Reflective Uncertainty and Irreversible Commitment suggests distinguishing two kinds of inquiry.

Cross-cutting epistemic inquiry
Research, observation, criticism, alternative cognitive architectures, records, and experiments aimed at distinguishing which world, value, or metaethical hypothesis is correct and thereby updating future credence and action.
Hypothesis-internal inquiry
Research conducted conditional on a hypothesis in order to resolve uncertainties internal to that theory and better realize its own aims.

The first is not naturally just another “vote” alongside utilitarianism, error theory, or realism. It is better understood as an upstream meta-practice that updates the weights and future practices of all branches. The second remains inside each branch.

More specifically, Allocating Inquiry Under Unresolved Normative Uncertainty uses the current epistemic state to reflectively allocate total practical resources among E (active inquiry), P (preserving future inquiry), and X (realizing present value). The parliamentary or proportional allocation in this page is therefore not a rule for directly dividing all resources; it is primarily downstream, inside the divisible ordinary-practice portion already assigned to X.

13. Reflective meta-practice: what to research and what to provisionally do

The epistemic layer concerns what to believe, but by itself it does not determine which research program to fund or whether to respond to possible warming, stability, or cooling with one global intervention. A further reflective meta-practice layer sits below belief formation but above individual value theories.

This layer takes the present structure of uncertainty itself as input. It selects additional observations, research programs, staged interventions, reversible experiments, and the level of evidence required before making irreversible changes.

If models include strong warming, stability, and cooling, the response need not be to realize each model in proportion to its probability. It can instead improve discrimination among the models, favor interventions robust across several models, and demand stronger evidence before irreversible planetary-scale manipulation.

Inquiry in the broad sense: inquiry is not only information acquisition. It is epistemically oriented policy aimed at moving into a state where better future understanding can still revise action. Research, reversible experimentation, record preservation, diverse validation systems, staged deployment, and rollback can all belong here.

14. Anti-lock-in belongs to the same layer as inquiry

Anti-lock-in fits more coherently as a constraint of reflective meta-practice than as an intrinsic good asserted inside one value theory. But uncertainty alone does not logically imply either “inquire” or “never lock in.”

As the existing Foundational Thesis stresses, a bridge on the agent side is still required: for example, conditional responsiveness to better justification or a desire to avoid irreversible choices that a better-informed future self would regard as clear errors.

For an agent with such a bridge, the same reflective uncertainty supports two thin meta-policies:

Neither treats inquiry itself as an already discovered terminal objective. They are conditional policies for an agent that regards its present evaluative judgment as fallible while wishing to remain responsive to future justification.

15. Hypothesis-relative practice: proportional provisional representation

Newberry and Ord’s parliamentary approach, together with proportionality and bargaining approaches such as Kaczmarek, Lloyd, and Plant, provide useful models for ordinary practice under rival value hypotheses.

For a divisible practical budget BX already assigned to X by Allocating Inquiry Under Unresolved Normative Uncertainty, if hypothesis Ti has provisional credence ci, a conceptual initial allocation is:

Bi = ciBX

Each hypothesis then optimizes within its share according to its own standard. A low-confidence hypothesis can retain some practical representation without an internal claim that an outcome is infinitely good automatically seizing the rest of the resources.

This is not probability truncation. A 1% hypothesis is not deleted; it retains roughly 1% provisional representation, subject to updating as evidence changes. Pooling and bargaining remain possible where branches find common advantage. Where sharp numerical credences are not justified, the equation should be read as a conceptual proportionality principle rather than a precision allocation rule.

Nor should a residual unconceived hypothesis that still lacks content and an action criterion be treated as a fictitious party that directly spends BX. Open-world residual uncertainty should primarily affect the upstream E/P allocation and constraints on the irreversibility of X. Once a candidate becomes sufficiently specified to support evidential and practical comparison, it can enter the hypothesis-relative practice layer.

16. Error theory and anti-realism retain practical territory

The proportional approach should not reserve representation only for objective-value realism. Error-theoretic and other anti-realist hypotheses receive practical territory in proportion to their epistemic standing as well.

Calling the anti-realist branch “maximizing the good” would be misleading. If objective good does not exist, its internal standards may instead concern actual preferences, pleasure and suffering, cooperation, institutional agreement, reflectively constructed values, or presently endorsed aims.

General form: allow each value or metaethical hypothesis some practical territory in which to act according to the criterion it endorses internally. Realism does not acquire total control merely because it may contain very large objective stakes.

17. What proportionality does not decide: shared, indivisible, irreversible acts

Probability-only representation is vulnerable to stakes-insensitivity. If a 99% theory weakly prefers A while a 1% theory says A causes civilizational extinction and B avoids it, a simple 99-to-1 vote can look too crude.

So proportional allocation should not automatically decide actions such as irreversibly transforming the planet, fixing one value into all future AI, or permanently altering everyone’s preferences. These are shared, indivisible, and irreversible cross-cutting decisions governed by reflective meta-practice.

Nor is it enough to put irreversibility inside expected value as a finite penalty such as −λR, because sufficiently large stakes can swamp it again. Irreversibility is better treated as a condition on the level of epistemic and cross-hypothesis support required for execution.

Inaction can itself be irreversible: extinction, lost ecosystems, destroyed information, or cosmological horizons may erase options while we wait. What matters is the irreversibility of both action and inaction.

18. Remote cosmology is not worthless; it is low-confidence and background-uncertain

This page does not assign value zero to distant worlds or infinite cosmology. It instead denies that we currently possess sufficiently confident detailed models of remote value distributions to make a complete infinite extension a fixed premise of current practice.

Tarsney’s work also suggests that large uncertainty about distant, action-independent value need not simply paralyze decision-making. It can function as background uncertainty under which ordinary local expected-value improvements generate stochastic dominance. Remote cosmology may sometimes be treated as an uncertain background rather than a domain that must already be fully aggregated into one total number.

19. Provisional architecture

  1. Epistemic layer: form credences or imprecise epistemic weights in world, value, and metaethical hypotheses from evidence and reasoning. Stakes do not make a hypothesis more likely.
  2. Reflective meta-practice: use that epistemic state to allocate resources among E (active inquiry), P (preserving future inquiry), and X (realizing present value), choose provisional interventions robust across models, and determine the cross-cutting support required for shared, indivisible, irreversible decisions. Residual uncertainty about unconceived alternatives first bears on inquiry, preservation, and reversibility here.
  3. Hypothesis-relative practice inside X: for divisible ordinary practice, give live hypotheses provisional representation responsive to credence mass.
  4. Internal optimization: within its domain, each hypothesis follows its own good, benefit, duty, preferences, agreements, or constructed values.
  5. Bargaining / pooling: where hypotheses recognize common advantage, allow exchange and joint action.
Central claim: inside the divisible ordinary-practice portion of X, epistemic weight governs how much provisional representation a hypothesis receives, while magnitudes of good or benefit govern what that hypothesis does inside its territory. Inquiry, preservation, anti-lock-in, and shared irreversible choices remain upstream matters for reflective meta-practice.

20. What this architecture does not solve

This architecture does not solve the internal mathematics of infinite ethics. A branch facing +∞ and −∞ still needs its own account of comparison.

There is also a bridge problem between the epistemic layer and reflective meta-practice. If “knowing truth” or “preserving corrigibility” is simply asserted as an unconditional objective good, the justificatory problem has merely moved up one level. Consistently with the existing Foundational Thesis, this page therefore presents the architecture as a provisional meta-policy for agents with conditional responsiveness to justification or reflective error avoidance.

It also leaves open how much civilizational resource should be reserved for cross-cutting inquiry, how to formalize the support threshold for irreversible decisions, and which axioms should govern bargaining across theories.

21. Open questions

Selected literature