Discussion about this post

User's avatar
David Manheim's avatar

I strongly agree with your conclusion, but think it makes some good points, and what I will argue are two bad ones.

Yes, things are changing, and it's unclear that the barriers which have existed historically will stay in place in light of AI and cloud labs. (It's unclear, as in I think both this piece and Abi's overstated this.) It's also the case that biosecurity is an obviously good investment, and we're failing as a civilization by not bothering to wipe out more infectious diseases, and drastically under-invest in biosecurity.

The two points I strongly disagree with are about terrorist actors. First, it is simply untrue that "terrorists don’t have a great track record with following cost-benefit logic." There is extensive literature showing that terrorist groups do, in fact, make decisions that advance their goals in ways that limit their costs. It's contentious, and there are certainly debates within that, but modeling terrorist groups as rational actors is more predictive than any alternative. And you argued that they deal with trade-offs.

I will note that people love claiming terrorism is "strategically irrational," as you said, but what they mean is that their goals aren't ones that the "rational" westerners claim makes sense - which deeply confuses what rationality means. Goals cannot individually be irrational, only actions to reach those goals can be. (And if terrorists were actually irrational in their combination of goals, we should worry less about them, as they'd be trivially exploitable by those around them.) The question of whether bioweapons would be strategically rational must then grapple with their actual goals and what they want to achieve.

And this brings us to the second point, which is what you claim the goals are, namely, "indiscriminate civilizational-scale damage," It turns out that almost everyone has other goals they want to accomplish, and *almost* no-one is actually ominical. And, of course, the critical piece is the world almost. But terrorist groups very much are not among those that want to cause indiscriminant damage!

The best piece I know of about actually omnicidal actors is by (everyone's favorite person) Emile Torres: https://www.sciencedirect.com/science/article/abs/pii/S1359178917302859 which points out that there only needs to be one such actor to end the world. However, contra Torres, Aum Shinrikyo almost certainly wasn't among them. You quoted Kyle Olsen's piece, linking the words "trigger Armageddon" - but he says "the objective of the Tokyo subway attack was not irrational. The objective that day was to kill as many policemen as possible..." This is incredibly different than triggering a global catastrophe directly; they believed phowa legitimized killing people, but the claims that they were actually omnicidal are obviously false, as they wanted to survive to rule the world after the prophesied catastrophe. And as evidence for that fact, note that they used a non-infectious bioweapon, and then a chemical weapon, each designed to inflict low level harm.

The critical point Torres makes, of course, which you echoed, is that misaligned AGI would by default be omnicidal - but that is about superintelligence, not misuse of AI by malicious actors.

Lennart Justen's avatar

Nice! Much to agree with here but I feel obliged to push back on your claim that cloud labs are really democratizing capabilities to anyone with a credit card and internet connection. I wrote about my skepticism of the cloud lab threat model here: https://substack.com/home/post/p-192022274

14 more comments...

No posts

Ready for more?