Welcome to The Nonlinear Library, where we use Text-to-Speech software to convert the best writing from the Rationalist and EA communities into audio. This is: Ways I Expect AI Regulation To Increase Extinction Risk, published by 1a3orn on July 4, 2023 on LessWrong.
The following are some very short stories about some of the ways that I expect AI regulation to make x-risk worse.
I don't think that each of these stories will happen. Some of them are mutually exclusive, by design. But, conditional upon AI heavy regulations passing, I'd strongly expect the dynamics pointed at in more than one of them to happen.
Pointing out potential costs of different regulations is a necessary part of actual deliberation about regulation; like all policy debates, AI policy debates should not appear one-sided. If someone proposes an AI policy, and doesn't actively try to consider potential downsides -- or treats people who bring up these downsides as an opponent -- they are probably doing something closer to propaganda than actual policy work.
All these stories have excessive detail; these are archetypes meant to illustrate patterns more than they are predictions. Before mentioning one fix that would stop a problem, consider whether it would be subject to similar problems in the same class.
(a) Misdirected Regulations Reduce Effective Safety Effort; Regulations Will Almost Certainly Be Misdirected
"Alice, have you gotten the new AI model past the AMSA [AI and ML Safety Administration] yet?" said Alex, Alice's manager. "We need their approval of the pre-deployment risk-assessment to launch."
"I'm still in correspondence with them," said Alice, wincing. "The model is a little more faceblind than v4 -- particularly in the case of people with partially shaved heads, facial tattoos, or heavy piercings -- so that runs into the equity clauses and so on. They won't let us deploy without fixing this."
"Sheesh," said Alex. "Do we know why the model is like this?"
Alice shrugged.
"Well, getting this model past the AMSA is your number one priority," Alex said. "Check if you can do some model surgery to patch in the facial recognition from v4. It should still be compatible."
"Alex," said Alice. "I'm not really worried about the face-blindness. It's a regression, but it should have minimal impact."
"I mean, I agree," grinned Alex, "but don't put that in an email."
"But," Alice continued, hesitantly, "I'm really concerned about some of the capabilities of the model, which I don't think we've explored completely. I know that it passed the standard safety checks, and I know I've explored some of its capabilities, but I think this is much much more important than -- "
"Alice," interrupted Alex. "We've gone over this before. You're on the Alignment and Safety Team. That means part of your job is to make sure the model passes AMSA standards. That's what the Alignment and Safety team does."
"But the AMSA regulations went through Congress before Ilya's Abductor was even invented!" Alice said. "I know the regs are out of date; you know the regs out of date; everyone knows they're out of date, but they haven't even begun to change yet!--"
"Alice," interrupted Alex again. "Sorry for interrupting. And look, I'm really sorry about the situation in general. But you've spent two whole months on this already. I let you bring in Aaron on the problem, and he spent a whole month. Neither of you have found anything dangerous."
He continued: "I've shielded you from pressure from my boss for a while now. He was understanding the first three weeks, impatient the next three weeks, and has been increasingly irritated since. I wish we had more time for you, but we can't just ignore the AMSA -- it's the law. I need you to work on this and only this, or you're fired."
I expect actual instances of misdirected safety effort from safety laws to be far more universal, and only moderately more hidden, than is indicated in this dialogue.
If you think this is unlikely, consider IRBs, and co...