www.jimbuhler.site
Also on LessWrong and Substack, with different essays.
Jim Buhler
just getting wide distributions for the cost-effectiveness
A normal Gaussian distribution? If so, then you still think the value in the middle of the curve is uniquely appropriate (even if barely so). To me, that’s the key difference between A) imprecision and B) precision with severe credal fragility. The former assumes you can’t non-arbitrarily pin down a precise credence at all, while the latter assumes you still can.
If VOI is overwhelmingly high, both A and B might recommend research, such that the difference doesn’t matter. But it matters a lot at least in situations where actors want to fund non-research things (because they think VOI is not that high or whatever). Then, A and B deeply disagree on what should be done.
Helpful thanks! Related thoughts from Clifton, here. But you actually do not object to UNSHARP (to some degree) for limited agents like us, then, right?
Do you agree with the other’s (EDIT: authors’) non-endorsement of Uniqueness? My impression was that you’d endorse SHARP because you think your sharp credence is uniquely appropriate. If not, why endorse this one rather than another sharp one that isn’t any less appropriate?
Say I compare different GHD interventions that help the worst off. I know that most such interventions are not crazy effective and that it’s easy to under or overestimate those with little evidence. I have a good reference class. If someone tells me about a new intervention in the area, I don’t expect it to be crazy good. It’s much more likely to be close to the mean. I have a prior expectation. If my math tells me there’s some poorly-studied intervention that beats the one that has consistently proven most effective so far, this weak evidence should not override my prior expectation. The math without accounting for my prior would give an overestimated number, surely.
But now say someone presents me with an x-risk intervention. I don’t have a “cross-cause prior” that says that an intervention from any cause is close to the GHD mean, do I?[1] If I understand your post correctly, you are implicitly assuming we do have such a cross-cause prior. (Otherwise, there wouldn’t be any OC-related reason to downweight the x-risk intervention.)[2] Is that correct?
Hi Vasco. Agreed, I think the relationship between p(sentience) and welfare ranges is very indeterminate.
Also, to be clear, I think p(sentience) is very relevant in theory. It’s just that, in practice, it seems like we have evidence that warrants ruling out assigning tiny probabilities of sentience to, e.g., shrimp and ants. However, I don’t think we have that for very narrow welfare ranges, and this is where uncertainty and disagreements lie.
The empirical evidence that shrimp have small brains detracts from this probability, but not by much.
I very much agree that pretty much whatever our prior should be, the available evidence does not justify substantially updating away from it. I’m just uncertain about what the prior should be (see below).
I think that conditioning on p(sentience) is sufficient to justify a non-negligible probability of similar levels of welfare range in the absence of any further empirical evidence.
Yeah, agreed that’s the crux! :) I think you are applying a principle of indifference (POI) across welfare subjects (or have a significant credence in such a move, at least).[1] While I actually also have sympathy for something of the sort,[2] it is widely criticized in the literature on cluelessness and decision-making under uncertainty. Here’s a list of challenges and possible responses taken from a rough paper draft of mine on this exact topic:
1. Uncertainty about how to individuate welfare subjects within a “welfare-containing” entity or within a bigger welfare-containing entity this one is part of → important instance of the problem of the many. Research could help us non-arbitrarily individuate (see, e.g., Gottlieb 2022; Fischer et al. 2022; McIntyre forthcoming) but this research may face very similar challenges to that on moral weights (on how much we can update away from whatever our prior is) and not bring us far.
But maybe biting the bullet and accepting some arbitrariness here is the least bad option we’ve got?
2. Why apply POI at the level of welfare subjects or brains rather than at the level of, e.g., cells?
Maybe persons (i.e., welfare subjects) are themselves what is morally relevant rather than their experience moments (see, e.g., Bader 2022), but
we’d need a solution to the non-identity problem.
and an argument for why following our intuition on this is fine but not with moral weights.
3. Why endorse any form of POI to start with? In a complex cluelessness context like the one we’re in when estimating moral weights,[3] the plausibility of POI is infamously contested (see, e.g., this, that, and refs therein). Hence, maybe we can’t use POI to justify a precise prior. Maybe we should favor an imprecise one such as each non-human species = (0, X). (where X = 1 or a bit higher.) (which would lead to agnosticism about whether many interspecies tradeoffs we make are justified).
However, to the extent that people want to reject such agnosticism (for whatever reason), even as an uninformed prior, they have to pick a precise-ish alternative prior. In this case, “everyone counts for (~)one” may be more advisable than the other options. (Wager on the possibility that we can apply POI).
A tl;dr from Claude that I like: ignorance about X’s welfare range doesn’t automatically justify treating X’s welfare as if it equals human welfare — it might just justify suspending judgment. The move from “we don’t know the ratio” to “assume the ratio is 1″ needs much more justification.
- ^
See also Dickens and Shepherd et al. (2023), who endorse this move.
- ^
Especially as an alternative to defaulting to our intuitions or “invertebrates don’t matter at all until proven otherwise”.
- ^
One could nitpick that there’s technically no complex cluelessness if we’re truly uninformed and ignore the (conflicting) evidence. But in that case, sure, maybe we can start with POI, but then we update towards agnosticism once we consider evidence, so the POI argument for giving everyone the same moral weight wouldn’t work.
I just had a naive illumination. Say that sentience first appeared in two different simple creatures, independently, at the same time:
Dolores: She’s just like her non-sentient siblings, except that she feels unnecessarily severe pain if she’s about to die of starvation, although not severe in a way that would impair her ability to do what is necessary not to starve (otherwise, she would die, and it’s her non-sentient siblings who would spread their genes.)
Mildred: Same, except that her starving pain is milder, and that’s enough to motivate her to lexically prioritize solving this problem, just like Dolores.
Judging by what you’ve written in the post and comments, you could give two different arguments for why Dolores would have lower fitness than Mildred:
1. Dolores’s pain would override everything else (e.g., she might be so focused on not starving that she forgets about drinking).
But this applies just as much to Mildred, no? No matter how mild her pain is, it will also override everything if that’s the only thing she feels. If she feels some pain when starving and nothing while thirsty, she might forget about drinking just the same. In pure isolation, how bad the pain is changes absolutely nothing in terms of fitness, here, no?
2. Dolores needs a more demanding biology than Mildred in order to feel something worse.
But how would we know this? Why would subjectively worse mean more demanding energy-wise? Why couldn’t it just as well be the more subtle less bad affects that are more demanding?
What am I misunderstanding/missing?
it seems to me to be very unreasonable to be confident that simpler brains most likely have much smaller welfare ranges
I agree, and I absolutely did not mean to defend this. What I defend is that, in the absence of a good argument based on welfare ranges and not p(sentience), we don’t know if the welfare range of simpler animals is below or above the bar above which their welfare would dominate over that of more complex animals (not that it is below!).
But you disagree with my a priori agnosticism because you think we should (roughly) stick to some precise-ish prior welfare ranges in the absence of significant evidence pointing one way or the other, correct? (And this prior would give simpler animals enough weight for them to likely dominate.) This would explain your disagreement with what you quote.[1] I was implicitly assuming that our prior should be an agnostic imprecise one that offers no action-guidance on its own.- ^
If that’s not where the disagreement is, I don’t see how “a presumption of a reasonable probability of a welfare range that is not too small and no significant evidence against it” does not count as “evidence of a welfare range that is not too insignificant.” Maybe you’re just worried my imprecise phrasing will, while technically correct, lead readers to set the bar too high?
- ^
Discussions of (p)sentience of small animals miss the point
Curious what motivated you to spend time assessing the impact of bird-safe glass on arthropods, specifically, then. Were you hoping to find out that bird effects dominated but found and shared the opposite unsatisfying results? Or maybe you think “here’s another example showing how indirect effects on tiny animals may dominate” and that this will convince some people to also prioritize (i) and (ii)? (people who were not convinced by your previous largely-overlapping posts but might by this one?)
Is there any project you think may not impact arthropods and/or soil animals much more than whatever animals are targeted? I feel like exploring this would be far more insightful at this stage.
Most animals are wild animals, so the answer to this question should focus on them.
Even granting that the overwhelming majority are wild animals, this doesn’t necessarily imply we should focus on them. We have to factor in the welfare difference between the two (welfare ranges and quality of life in practice).
Oh good, I have no objection then. Well played.
Are you setting aside wild animals?
this seems to me to imply a greater concern for anthropogenic harm than non-anthropogenic harm. Is that what you meant?
Oh no sorry, increased WAW welfare compared to the “natural” situation counts as impact too.
What I’m saying is: say you help 1 million wild animals out of many or 1 million farmed animals out of fewer. You can’t say the former is better because there are more wild animals. It doesn’t matter how many there are. What matters is how many you help and how much. And there is an asymmetry here where farmed animals are probably 100% helped if humans are disempowered—the problem is totally fixed—whereas, even in the best case scenario, empowered humans will be nowhere near totally fixing wild animal suffering. This asymmetry may compensate for the fact that there are many more wild animals to help.Humans increasing or decreasing the number might be the largest impact
As in (D) is more plausible than (C) (in my typology)? I’d agree. Anyway, my argument holds independently of what people find more likely between (C) and (D).
For example, the regeneration of forest is actively opposed in much of Central Europe, because people have cultural ideas about what the landscape should look like. So there’s a tension there between environmentalists and traditionalists, and I wouldn’t say that the environmentalists are winning.
Oh I didn’t know that, thanks. There, of course, is still the question of the marginal impact WAW advocates would have in such debates, but helpful example!
I wasn’t thinking about promoting/opposing restoration but about influencing how it is done (without necessarily taking a stance on whether no restoration would be better). And I could very well imagine WAI wanting to advise decision-makers on how to conduct restoration.
I think present and future WAW advocates would fiercely disagree about what ecosystems might be net good/bad, and any intervention aimed at making greening more likely would be highly controversial.
Interventions aimed at, at least tentatively, holding off on restoring would be far less controversial, though. And in that case, yes, I doubt that WAW advocates “leveraging conservative valuing of traditional landscapes to oppose it” would successfully prevent any restoration project. Whatever the incentive for restoration is, it seems far stronger than the incentive to please the few detractors who do not want the landscape restored.
Interesting, thanks!
An intervention doesn’t even need to be framed around WAW either—you could just fund an organization to lobby for desert greening (for example) in a particular area, and they could leverage whatever arguments they’ve got.
That’s good only assuming WAW in the ecosystem you create is net positive tho, right?
I was imagining more like:some restoration interventions are and will keep happening anyway.
let’s influence those and push for ecosystems with less suffering.
But I just find it hard to make a difference there. For social/political reasons, yes. Not necessarily because people would be against the idea, but just because there’s no/little incentive for the relevant actors (in the restoration process) to do what we’d want there. Why would they bother? I also feel like WAI would have discussed this more if this were tractable? Haven’t thought about this much tho.
[On my Substack], most of my readers are not as familiar with EA discourse
This surprised me. Where do they come from?
One thing is whoever does not reject UNSHARP might not have severely imprecise credences about everything. I might believe that
intervention 1 has severely indeterminate but astronomically high (positive or negative) EV.
intervention 2 seems overall good, although it has lower EV.
Then, I’d probably prioritize intervention 2. If I instead endorsed SHARP, I might favor intervention 1 (because of a sufficient 51% credence 1 is good). (I’m actually not sure about this, though. One could argue that 1 and 2 remain incomparable and that I have no reason to favor 2 over 1.)
Another thing, assuming there is no 2-like intervention, is that the criterion to pick could be something other than “act straightforwardly as if you were endorsing SHARP”. It could instead be, e.g., some (other) form of bracketing.