Discussion about this post

User's avatar
David Spies's avatar

I'm having trouble imagining anything that could plausibly look like "missiles in Cuba" to set this off, at least not until after it's likely too late to do anything. Or rather, the sort of thing I'm imagining is "lab was training/evaluating an AI and it secretly broke out of the box, went rogue and started hacking things on the open internet" but that happened already a couple times

Do you have a story of what leads to getting serious?

Connor Williams's avatar

"2: Why not just pause now? [...]"

Putting aside the question of whether it would have been net-negative to pause earlier (I personally disagree in large part because the shorter the runway to dangerous capabilities, the more fragile the pause inherently is[1], it's much simpler to pause now than to "agree to pause in the future, and then do it at just the right moment".

[1] https://connorsscratchpad.substack.com/p/ai-breakout-times-a-mechanism-for

- Political capital is easier to build for the first, especially beyond narrow technical spheres, because it's a simpler, more intuitive, and more emotionally satisfying goal.

- Pausing doesn't happen in an instant. It will, in practice, likely take at least weeks if not months for an enforceable pause to be implemented once everyone agrees it is time. In that period, capabilities development will continue until the absolute last moment.

- Fewer points of failure. To agree to pause at some point in the future, you both have to get people to agree to that future goal, then when the time comes you have to get them to agree again. Each distinct time that everyone has to agree on something, there's an opportunity for failure/deception/defection.

- The concern I noted above about the length of the runway to truly dangerous capabilities applies just as much going forward as it does looking backward. In that sense, the sooner we pause the better.

[Note: I posted substantially the same comment in reply to your link-post of this on LessWrong]

6 more comments...

No posts

Ready for more?