Interesting post! My colleague Stephen Clare (at Founders Pledge) is currently doing an investigation into this topic, it will be great to exchange.
Interesting post! My colleague Stephen Clare (at Founders Pledge) is currently doing an investigation into this topic, it will be great to exchange.
Thank you! Sounds great. DM'd!
Feel free to skip this part if you’re familiar with the definition of a global public good and basic concepts from game theory, like the prisoner’s dilemma.
What are public goods? Public goods are those goods that any individual can access (non-excludable), and an individual’s usage of them doesn’t reduce the supply available to other individuals (non-rivalrous). Examples include world peace, a clean atmosphere, and a pandemic-free world.
Global public goods are those public goods which benefit, or at least do not harm, everyone in the world[2]. Note, foreign aid is not a GPG. This is because aid violates the non-rivalrous condition: that country’s consumption of aid reduces the resources available to the donor and other countries.
GPG provision commonly requires international cooperation. However, it is possible for a country to provide a GPG unilaterally, in which case cooperation is not required. For example, if a country had the resources to build a defense system against a catastrophic asteroid impact, that country would have an incentive to do so even if others would not help build it.[3] Despite this possibility, the GPGs I will discuss would benefit from cooperation, so I may use the phrase “GPG provision” interchangeably with “international cooperation” in this post.
One can classify GPGs in terms of aggregator functions[4],[5]. The inputs to these aggregation functions are the contribution levels of the public good from each country, and the outputs are each country’s benefit, given those contribution levels. These aggregator functions are necessarily simplifying. But I think the intuition they provide is nuanced enough to tell us when international cooperation may be especially impactful. Below, I discuss two aggregator functions: “best shot” and “weakest-link”.[6]
What matters most in determining outcomes for a given country? The best effort that any single country makes (asteroid deflection)? Or the efforts of the least well-resourced country (disease eradication)?
Asteroid defense is an example of a best-shot aggregator because if a single country (likely wealthy) can knock the asteroid off course, every country receives the benefit of asteroid defense, even those that do not contribute. The outcomes of all countries depend on the country which invested the most in asteroid deflection. Generally, total provision of a best-shot GPG will depend only on the amount that the largest contributor to the GPG provides.
For smallpox to be eradicated, if even one country still had incidence of the virus, all other countries would’ve had to bear the costs of annual vaccination[7]. Hence, the collective interest depended on the “weakest-link”, or the country which invested the least in eradication[8]. Generally, total provision of a weakest-link GPG will depend primarily on the amount that the smallest contributor to the GPG provides.
Both of these risks also have commonalities with the Type-1 vulnerabilities defined in Bostrom’s Vulnerable World Hypothesis paper[12], since a single actor could cause catastrophic outcomes for large populations. Fortunately, both AGI and engineered pandemics are currently much more difficult to develop than the “easy nukes” in Bostrom’s example.
I’ll continue to focus primarily on AI/EP risks in this post, in order to highlight the more tangible and actionable benefits from international cooperation. However, I believe that the benefits from international cooperation extend far beyond reducing these risks[13], or for that matter, any existential risks we are currently aware of. In (Bostrom, 2019), Bostrom points out that we don’t know how dangerous and accessible future technologies might be[14]. For example, he points out that we were simply lucky that nuclear weapons required a large amount of resources to build. If they could’ve been built with “a piece of glass, a metal object, and a battery”, then their discovery would’ve spelled a global catastrophe. By improving global institutions now, before such technologies are discovered, responses to future risky technologies could be more robust than the uncoordinated actions of ~200 sovereign countries. Since improving international cooperation could reduce future risks larger than even the most dire ones we’re currently facing, the benefits to reducing known existential risks may be just a small fraction of the total value generated by improving international cooperation. For more on the benefits from global governance on yet-unknown existential risks, I refer the reader to (Bostrom, 2019).
If enhancing global cooperation could significantly reduce the risk of just 2 types of existential catastrophe, this cause would probably merit increased efforts from the EA community. Ord estimates a roughly 1-in-6 chance of existential catastrophe over the next 100 years[15]. Two primary contributors[16] to this risk are unaligned AI and engineered pandemics. International cooperation could significantly reduce the likelihood of these two risks materializing, via, for example (1) a UN Security Council resolution establishing AI development norms, with trade sanctions for non-cooperators, and (2) an international treaty to improve monitoring of actors capable of releasing an engineered pandemic.
While I will not assign an exact number to the risk reduction claimed above, I do feel capable of estimating a rough lower bound on this cause’s potential impact, given my intuition of how these risks might unfold. I believe improved cooperation alone could probably prevent at least 1 in 20 AI/EP-related existential catastrophes.
A risk reduction of this order of magnitude seems to be supported by the role of international cooperation in past catastrophic risks. For example, I believe a significant amount of the nuclear war threat in the Cold War was generated by poor international cooperation between the US and the USSR, rather than, say faulty detection systems, human error, or lack of scientific understanding. I think roughly 1/10 of the scenarios which almost led to nuclear war probably could have been prevented if only both countries had committed earlier to mutual arms control. Ord lists several of these close calls in The Precipice [17].
Overall, it seems likely that for every 20 scenarios where existential catastrophes occur from engineered pandemics or unaligned AI, at least one could have been avoided if there had only been stronger international cooperation.
The reasoning above is the most important factor in my recommendation. If international cooperation were demonstrated to have little impact on existential risk, then this cause should not be a priority. I feel approximately 60% confident that improved international cooperation alone could eliminate a significant proportion of AI/EP existential catastrophes.
Finally, despite my focus on existential risk in this post, international coordination could be impactful for non-existential, non-longtermist causes. Improved cooperation on medical research, climate change, and trade could improve our longevity, environment, and wealth in the current generation. Using the 80,000 hours framework, I assign a 14 to the importance factor, as I believe improvements in global cooperation on AI/EP could reduce existential risk by approximately 0.5-1%. This corresponds to a very high importance relative to other priority causes.
I found few organizations which focus exclusively on international cooperation on existential risks. But I believe many institutions devote a minority of their resources to improving cooperation on these issues[23].
While I will list estimates of funding below, I caution that neglectedness in this area is probably not well-captured by a funding number. Rather, prioritization of international cooperation on existential risks by major governments is likely a better measure. So, progress depends more on influencing the opinions of the public and high-ranking government officials, than financial resources alone. That said, more money may be used to indirectly influence opinion, e.g. by launching public awareness campaigns and advertisements on the importance of improved international cooperation.
Nuclear Threat Initiative: upper bound of $19M on EP expenditures, 2019[24]
Johns Hopkins Center for Health Security: Estimated budget: $3M - $15M[25]
Partnership on AI: Revenue of $8M in 2019[26]
Future of Life Institute, $2.4M in 2019 revenue[27]
Future of Humanity Institute, which houses the Centre for the Governance of AI: roughly $1.4M annual budget in 2017.[28]
Centre for the Study of Existential Risk: Estimated budget: $1-$10M[29]
Leverhulme Centre for the Future of Intelligence, roughly $1.4M in annual funding[30]
Global Catastrophic Risk Institute: $0.3M annual budget[31]
The total annual budgets of the above institutions is approximately $57M[32]. This number is likely an overestimate, since these organizations are not entirely devoted to international cooperation on AI/EP risks.
One news story that shows this issue is becoming less neglected are the “track-II” diplomatic talks between Fu Ying, former Chinese ambassador to the U.K., and Brookings President John Allen, which centered on AI safety[35]. These talks indicate that influential individuals outside government place high importance on this issue.
The Asilomar Conference on Beneficial AI also points to increased attention to this cause area. An outcome of the conference was the Asilomar AI principles.
Events that would make me think international cooperation on existential risk is becoming even less neglected include: diplomatic talks on existential risk moving to “track-I” (especially official conversations between China and U.S. governments), a large number of countries endorsing the Asilomar principles, and incorporation of the Asilomar principles into an enforceable international agreement.
Using the 80,000 hours framework, I assign a 6 to the neglectedness factor. This corresponds to a moderate level of neglectedness relative to other priority causes.
International cooperation has usually been successful when (1) benefits to cooperating clearly outweigh costs by a large factor, and (2) all countries benefit from cooperation. I’ve included a few case studies of successes and failures in international cooperation in Appendix 4.
The benefits of reducing risks from engineered pandemics and AI risks are fairly uncertain: unaligned AI has not been developed yet. Bio-warfare has not killed large numbers since World War II[36]. I do not say this to discourage work on these areas or diminish the size of these risks, but to show that governments cannot point to a clear benefit to be gained from cooperation.
The impact of an engineered pandemic would likely be bad for everyone, which means international cooperation is probably more likely on this issue than AI safety. AI might be well-aligned to only one country’s values, so cooperation on AI safety may be less likely if a government believes that not cooperating could give that country a strategic advantage.
This is all to say: Past successes in international cooperation have little in common with reducing AI / EP risk. For example, the costs and benefits of eradicating smallpox were fairly certain, and there was no strategic advantage to be gained by any country not contributing to eradicate smallpox[37].
However, small successes have already occurred in international cooperation on both AI and EP. International membership in the Partnership on AI indicates the potential success of international agreement in this area, although Baidu’s withdrawal is worrying for future US-China cooperation. The near global ratification of the Biological Weapons Convention (BWC) of 1972 shows that collaboration on reducing the risk of engineered pandemics may be tractable, although the convention is not monitored, so compliance levels are uncertain. Certain countries, such as North Korea, are in blatant violation of the BWC[38].
Due to international cooperation’s moderate track record on issues with uncertain benefits, I assigned a 4 to the tractability score. This corresponds to a moderate level of solvability relative to other priority causes.
In this section, I discuss why you might not prioritize international cooperation for AI safety and engineered pandemics.
My argument for international cooperation has at least four areas of potential weakness. I believe objection 1 is the most plausible, as there are at least 3 ways in which it could be true. I am especially uncertain about the propensity of governments to cooperate in the absence of further research and advocacy into international cooperation. I am also uncertain if relevant governments are able to monitor either of these risks with enough scope to reduce existential risk.
This could be true if:
You may believe cooperation is important, but skeptical that the international scale is the best level to focus on.
While central institutions are generally strong in developed countries, I am not sure if even the richest and most technologically advanced nations could monitor AI/EP activities to the degree necessary to reduce existential risk. Since small, poorly funded actors might be able to contribute significant risk to both unaligned AI and EPs in the future, monitoring may need to be extensive. So, even if technically feasible, surveillance may come at the cost of seriously invading individual privacy[39]. If effective surveillance is practically impossible with today’s social norms and technology, investing in the technical capabilities needed to make monitoring more palatable[40] and increasing public awareness of the gravity of existential risks (which may increase support for the necessary surveillance) could be greater priorities than international cooperation.
If you believe change comes from the “bottom-up”, and a more altruistic or engaged citizenry is necessary before leaders will engage in global cooperation, you may instead want to focus on building the EA community.
I framed the reduction of AI/EP risks as weakest-link GPGs. If you believe these risks are better modeled by another aggregator function, especially a “best-shot” GPG, you should want to invest more into technical research instead of international coordination.
What would it mean for these risks to be best modeled by a best-shot aggregator?
For AI, if you are confident that actors in a certain country or project will develop safe AGI first, i.e. there is little risk of a race to develop AGI (in which speed is prioritized ahead of alignment with human values), you may prefer to focus solely on AI technical research in the most promising country/project, to ensure that the first AGI is as safe as possible. This human-aligned AGI could then halt the development of any non-aligned AGI. However, if you believe that projects in multiple countries will compete to reach AGI, as currently appears to be the case[41], then some amount of international cooperation is likely wise. On the other hand, Eliezer Yudkowsky seems to believe that AI risk is best reduced by “best-shot”-like efforts[42].
For engineered pandemic response and prevention, a best-shot aggregator would apply if you believe that some sort of silver-bullet technology could neutralize most viruses. For example, imagine a medical device which could synthesize vaccines within seconds of coming into contact with the saliva of an infected person, no matter the virus. You would probably choose to invest in developing this technology, or any other similarly protective technology, since invention would render any pandemic nearly harmless. If you think that such silver bullets are unlikely or prohibitively expensive, international cooperation on biosecurity would likely still be worthy of attention. I am unsure of the probability of invention of such “silver bullets” over the next decade.
Below I include what I view as key components of any future international agreements on AI / EP safety, specifically. Countries that cooperate on these points may recommend trade sanctions on non-cooperators to incentivize participation.
I believe international cooperation on existential risk would no longer be neglected, and would soon reach diminishing marginal returns if events on a scale similar to the below occurred:
A look into the history of GPGs is helpful in estimating the tractability of improving GPG provision. Roughly, GPG provision efforts have succeeded when (1) the benefits from cooperating clearly outweigh the costs, and (2) every country benefits from cooperation.
As defined on p. 3 in (Bostrom, 2001) ↩︎
p. 16 (Reisen, 2004) ↩︎
This example was borrowed from (Barrett, 2007) p.2 ↩︎
(Buchholz and Sandler, 2021) p.10 ↩︎
(Barrett, 2007) p.20 ↩︎
Other aggregator functions are discussed at length in (Buchholz and Sandler, 2021), p. 10-14. ↩︎
See Appendix 4 for further details on the eradication of smallpox. ↩︎
Barrett discusses the details of disease eradication in the context of polio in (Barrett, 2010). ↩︎
(Ord, 2020) p. 167 ↩︎
(Bostorm, 2014) p. 77 ↩︎
(Bostrom, 2014) p.100. This argument cuts both ways: if the first superintelligence is aligned to human values, it could also stop all progress on unaligned AI. ↩︎
(Bostrom, 2019) p. 458 ↩︎
Thanks to Michael Wulfsohn for raising this point. ↩︎
(Bostrom, 2019) p.455 ↩︎
(Ord 2020) p. 167 ↩︎
Ord estimates the risks of existential catastrophe from AI and engineered pandemics as approximately 10% and 3.3%, respectively, over the next century.(Ord, 2020) p. 167 ↩︎
(Ord, 2020) pp. 96-97 ↩︎
See the prior section: “Aggregator functions applied to two existential risks”, above, for why this path dependence exists. ↩︎
See Recommendations section for example criteria of a “strong level of cooperation” ↩︎
The amount of funding for Open AI, one research lab working on developing safe AI ↩︎
The approximate amount of viewers for the Super Bowl in 2021, typically the most-watched American television event each year. See footnote 22 for further details. ↩︎
$1B could have bought all of the Super Bowl commercials in 2021 (assuming $5M for a 30 second advertisement, and 50 minutes of ads), with $500M left over for production costs. Even more cost-effective advertising is probably achievable via Facebook or other targeted advertising. ↩︎
I have not listed organizations that advocate for international coordination on other risks, such as nuclear war or climate change, since the focus of this post is AI/EP risk. ↩︎
Despite its name, NTI also funds programs to reduce biosecurity risk. I excluded the Global Nuclear Policy Program and International Fuel Cycle Strategies line items, to arrive at an upper bound on pandemic response and prevention funding. Funding amounts can be found on p. 31 of (Nuclear Threat Initiative, 2019). ↩︎
I assumed that John Hopkins’ Bloomberg School of Public Health’s $598m budget (Johns Hopkins, 2021) was allocated to research centers on a pro rata basis based on the number of faculty in each research center. There are 12 faculty members ranking Senior Scholar or above at the Center for Health Security. There are 837 total faculty at the School of Public Health (12 / 837) * $598M yields a midpoint estimate of $9M. ↩︎
(Partnership on AI, 2019) p. 9 ↩︎
(Future of Life Institute, 2019) ↩︎
(Open Philanthropy, 2017). Used exchange rate of $1.39 per £1. ↩︎
I could not find an exact budget. I estimated these numbers based on FHI’s budget, a similar research center at Oxford. ↩︎
(Leverhulme, 2021) Calculated as 1/10 of the 10 million pound grant from the Leverhulme Trust. See footnote 28 for exchange rates. ↩︎
(Global Catastrophic Risk Institute, 2020), retrieved April 14, 2021 ↩︎
I used the upper bound of my estimates for the Johns Hopkins Center for Health Security and Centre for the Study of Existential Risk budgets. ↩︎
(Ord, 2020) p. 280 ↩︎
(Global Priorities Institute, 2020) p. 43 ↩︎
(Zheng, 2020) ↩︎
(Frischknecht, 2003), p. 2 ↩︎
See Appendix 4 for further details on the eradication of smallpox. ↩︎
(US State Department, 2019), p.47 ↩︎
In the Preventive Policing header in (Bostrom, 2019), Bostrom explores the tradeoff between enforcement efficacy and privacy. While a “high-tech panopticon” probably would not be necessary to reduce AI/EP risks to acceptable levels today, if the means to unleash existential catastrophes come into the hands of many, citizens may choose to trade civil liberties for increased safety. ↩︎
“Such as automated blurring of intimate body parts, and...the option to redact identity-revealing data such as faces and name tags”, (Bostrom, 2019). ↩︎
The Communist Party of China has set a goal to be the global leader in AI by 2030 (O’Meara, 2020). ↩︎
(Yudkowsky, 2008) p.333-338, Yudkowsky lists several reasons to invest in “local” efforts, which he defines as actions that require a “concentration of will, talent and funding to overcome a threshold”. Yudkowsky argues that “majoritarian” action (like international cooperation) may be possible but local action (like technical research) is probably faster and easier. My view is that both majoritarian and local actions should be undertaken to reduce AI risk, especially when there is not common knowledge of all actors’ potentially risky activities. ↩︎
(Ord, 2020) p. 201-202 ↩︎
“if a value-aligned, safety-conscious project comes close to building AGI before we do, we commit to stop competing with and start assisting this project.” ↩︎
(Bostrom, 2014) pp. 102-106 ↩︎
These final two measures are cited in (Nouri and Chyba, 2008) p. 463-464, ↩︎
(Wilson, 2013), pp. 351-364 discusses what such a treaty might include. ↩︎
(Barrett, 2007d) pp. 79-82 ↩︎
(Barrett, 2007) p. 52 ↩︎
(MasAskill, 2018) This statement is made at 3 minutes, 3 seconds. ↩︎
See p. 1252 of (Allwood, et al 2014), which defines Annex-1 countries ↩︎
(Dessai, 2001) p. 5 ↩︎
I think EAs focused on x-risks are typically pretty gung-ho about improving international cooperation and coordination, but it's hard to know what would actually be effective for reducing x-risk, rather than just e.g. writing more papers about how cooperation is desirable. There are a few ideas I'm exploring in the AI governance area, but I'm not sure how valuable and tractable they'll look upon further inspection. If you're curious, some concrete ideas in the AI space are laid out here and here.
Great points. I wonder if building awareness of x-risk in the general public (i.e. outside EAs) could help increase tractability and make research papers on cooperation more likely to get put into practice.
I'm curious which ideas you're exploring too. I saw your post on the topic from last year. Reading some of the research linked there has been super helpful!
Thanks for linking these resources too. Looking forward to reading them.