TL;DR: I'm releasing a website that ranks philanthropists according to EA principles and research, and allows users to re-rank the list using their own assumptions. I'd like feedback and help making it better. I'd especially like ideas for how to make the results more trustworthy. Funding may be available.
I recently built Impact List (impactlist.xyz), a site which ranks people by their positive impact via donations.
The goal is to make the list popular enough that people care about their ranking on it, so that it influences their decisions about where and how much to donate. A secondary goal is influencing people (whether or not they appear on the list) by making them more aware of the large differences in the cost to save a life depending on where money is donated.
The site uses QALYs as the common currency of value, but this doesn't mean it's limited to considering only effects on health/lifespan. The goal is to consider effects of any type and convert them to human-QALY-equivalents.
I wrote more about the motivation and theory of this project four years ago in this post.
The core of the site is a ranking of (currently 73) wealthy philanthropists by their expected impact via donations.
Each donor has their own page showing all their donations (along with the impact of each donation), and aggregate stats by cause area. Users can toggle between 'donations' and 'lives saved' to visualize differences in effectiveness between cause areas. The internal QALY metric is converted into lives saved by defining one life = 80 QALYs.
I've split all donations into 28 cause areas, each with their own 'cost to save a life' value.
There's also a page for every recipient (charitable organization), which can have its own 'cost per life' value if it's more or less effective than the cause area average. Each cause area and customized recipient has a page with a detailed justification for how the math was done to arrive at the cost per life, and where assumptions are explicitly listed.
The estimates are of course very approximate, and will be controversial. I make heavy use of LLMs to do this research, encouraging them to synthesize existing (often EA) analysis. Users who disagree with any of the default parameter values can edit them.
Users can also edit global parameters specifying their time horizon, assumptions about population growth, and the discount rate.
After entering their own assumptions, users can create shareable links. This allows others to see what Impact List would look like given person X's worldview.
I'd eventually like to add worldviews from notable researchers or organizations, but I'm starting with just a handful of worldviews based on simple tweaks to the defaults:
There's also a calculator feature where people can enter their past donations and/or donations they're considering and see where they'd rank on the list, given those donations.
If you see something below that you want to help with, comment on this post, DM me, or join the Discord.
Few people are going to dive into the details of the effectiveness estimates and verify for themselves that the research is high quality. So it's not enough that the estimates be excellent. People need to be able to easily understand and trust the process that leads to the estimates.
The current process is LLM-driven, opaque (users can't see the prompts that lead to the estimates and they can't see how feedback is processed), and dependent on the judgments of someone (me) without any reputation as a researcher. I've tried to partly compensate for this by making the justifications explicit about which assumptions are being made, and tried to make the reasoning as clear as possible, but I expect it'll be very hard to generate enough trust using the existing process.
A few things that could help:
If you have ideas in this area, please share in the comments. I see this as the most important obstacle to the site becoming popular.
The effectiveness estimates for all cause areas are mostly based on research from LLMs (frontier models from Anthropic and OpenAI). I've also done some manual review and back and forth with the LLMs to refine the estimates. The LLMs rely a lot on existing EA research, but the quality of the analysis is pretty uneven and could probably be improved a lot by putting expert human researchers in the loop.
I'd like these estimates to become the best place to look for a synthesis of all effectiveness research that EAs have done (or found).
If you have expertise in cause area effectiveness research, and especially if you're an expert in some particular cause area, it would be great if you'd be willing to help improve the estimates.
I think the site looks OK now, but not great (especially on mobile). If you have UI skills/taste and want to make the site look better, let me know.
As mentioned above, the site allows anyone to make their own customized set of assumptions and create a shareable link to the rankings using these assumptions.
For notable people/orgs, I'd like to add these links as curated options for all users, so people could see Impact List according to Carl Shulman, Rethink Priorities, etc.
If you're a notable researcher or work in a well-respected EA org, and you also want this to happen, please get in touch.
How should we actually make the site popular? In my post introducing this project I talked about some ideas for how to do this, but I'm pretty uncertain about the right path here.
My sense is that the quality of the site should be significantly higher before I try to popularize this beyond EA/rationalist audiences, but it'd be nice to have a distribution plan soon to guide the other site-improvement decisions.
Everyone currently on the list is well known, but that's not a requirement. Feel free to nominate yourself (if your donations can be proven to the satisfaction of a skeptical reader) or someone else I missed.
See Appendix A below for details on how this works. I've tried to strike a balance between simplicity and expressive power, but I'm not sure I've picked the right tradeoff or whether I'm on the Pareto frontier.
A couple of funders have offered to give me grants for this work. I haven't accepted any yet, but if you want to work on Impact List and you want to be paid, these grantors may be willing to make that happen.
The project has a Discord. You can also submit pull requests via GitHub. Feel free to DM me on this forum.
You can keep up with the project on Twitter.
Above I've listed what I think I most need, but I may be wrong. I'm happy for all sorts of help or feedback.
Thanks to Austin Chen, Ryan Kidd, Kim Korte, Nathan Young, plex, and Laszlo Treszkai for discussion and feedback about the site.
The details are on GitHub. This is a simplified overview.
Each cause area or charitable organization has one or more 'effects', which define how money translates into QALYs. There are two types of effects, standard effects and population effects, and each effect has several parameters.
Standard Effects
The most intuitive type. The more money you spend, the more impact you get. The parameters are:
Population Effects
Effects where money donated changes the probability of some event. This is used to model x-risks, pandemics, etc. The parameters are:
Overrides and multipliers
A recipient organization which is part of some cause area can override the parameters of that cause area, to indicate that the recipient is especially effective or ineffective.
Recipient organizations can also have their parameters defined as a multiple of a parameter of the cause area it belongs to. The site's UI currently hides the ability to express parameters in terms of multiples for simplicity, though I'd like to re-add this to the UI at some point.
Time
The model is capable of handling time-based effects, which allows specifying different levels of effectiveness for a cause/recipient depending on when the donation was made. For instance a donation to MIRI today may have a very different impact than a donation to MIRI in 2012. The current version of the site doesn't make use of time-based effects, but they're available and documented on GitHub.
Motivation for these choices
This model gives each effect a 'shape' over time that we can integrate over to get the total value. This allows assumptions about the future population, discount rate, and time horizon to be separated out from cause-specific assumptions.
I could have made these shapes more complex by having options for exponential or linear decay of effects, but I think it's not worth the extra complexity. Every effect-shape in this model is a simple rectangle.
Thanks to Nathan Young for the suggestion and for discussions around this idea.
This is really cool! If I might make one observation, the default assumptions (counting animal lives + assuming that AI is going to destroy the world) are not going to be understandable to most people. I think most people would assume that a list like this would count human lives that have already been saved, not ones that might be saved in the future. If the goal is to appeal to a broad audience, it might make sense to make more standard assumptions by default.
If you want to highlight animal advocacy and existential risk as cause areas, it might be helpful for the main page to show lives saved broken down by human/animal and certain/uncertain or current/expected future.
Also, I think there are huge issues with the way the population cause area is framed:
Hi Ellie -- thanks for the comments.
RE: the default assumptions, for value judgments I agree that I want the defaults to reflect what is common.
For AI existential risk, I think the crux is mostly not about values (at least for the default assumptions, which only consider the next 100 years) but about how the world works. In those cases I'd like the default rankings to present the best estimates possible without trying to skew towards popular beliefs. But there might be some way that I can make it more clear to people what's happening.
Fair point about the "lives saved" terminology. I could maybe call that column "life equivalents".
RE: whether creating a life is equivalent to saving a life, I agree this is something that lots of people will have objections to. Possibly I should make the user's stance on population ethics an explicit global parameter. I think the most typical view can be simulated with custom assumptions by disabling the Population cause area and possibly shortening the time limit. You can even simulate a longtermist with a person-affecting population ethics view by capping the population at close to the current level and setting the time limit very high (along with disabling the Population cause).
I agree that the motivation behind the rich country phrasing is unclear. I've added some clarification to the site. The estimate focuses on rich countries because the type of charity it's trying to compute a cost per life for tends to be a rich country thing. (There's currently only one recipient in that category.)