Wanted to share a quick update on my comments on AI safety comms from a couple weeks ago. The NY Post is now reporting that the FTC is investigating OpenAI, Anthropic and METR, with formal demands (similar to subpoenas) reportedly coming in the next few weeks.
I think Brian Martin's Backfire framework could be useful in responding here. Martin found that power holders use five methods to suppress outrage, and there's something we can do in response to make it backfire:
I've worked on a few similar cases in the past two years, mostly in the pro-democracy space, including the FBI's arrest of Milwaukee County Judge Hannah Dugan. We built a coalition beyond the usual suspects and brought her case to the public with polling that showed a clear message to the general public. We hosted a vigil before her trial and a "Court of Public Opinion" outside the courthouse each time she appeared before a judge.
Making this backfire will be easier to do now, before the attack lands. Crisis comms is 95% preparation and relationships, 5% the crisis itself.
Would love to talk with more people about this!
*AI use: I used AI to tighten up the backfire bullets and intend to lengthen this piece into a longer one.
Thanks for pointing that out.
I meant "if you're explaining, you're losing" for any platform, not X replies specifically.
I agree Community Notes are very likely net positive (they certainly have an impact on me). But where I still see some risk is volume: Replies are a signal and if a lot of people pile on to rebut a post, perhaps that pushes it to more people? (I don't know X's algorithm well enough to say that with much certainty.)
I also don't want to double down too hard on illusory truth. I defined it a bit too loosely in my quick take, and there are follow-up studies showing a good correction usually wins with the people who read it.
For me it's more about Zaller's Receive-Accept-Sample (RAS), which I find helpful for thinking about mass/political comms. Voters have to:
Continuing to respond to a critique that's wrong makes it more likely people 1) hear it and 3) sample it later.
Thanks for continuing to engage.
Now is the time for PR Comms people to be working overtime
I'd love to pitch in on this somehow. This mostly started as an attack on METR, and after a quick search it looks like most AI safety orgs don't have a mass communications function? Is that accurate, or am I missing something?
Here are some quick thoughts:
Points #1, 4 and 6 aren't asking anyone to really say anything differently, it's more about who says what, where, and when. If someone from outside the AI safety world (or better yet, someone who can show they've disagreed with METR before or are not a natural ally) says "the corruption story isn't true," that's not less epistemically correct, it's just likely more believable to the general public.
I'd gently push back on the bigger part of your argument though: The public conversation can't end up being about METR's independence, it should be about an outside entity checking the lab's work. That seems to me the tradeoff if we lose the bigger comms battle.
Some quick observations on the attacks on METR, EA, and other AI safety orgs (from my perspective as a political campaigner and former Communications Director for a union and mayor during COVID):
I want to write something longer on this but don't have time right now. Lots of opportunity for good crisis comms here. Let me know what's wrong about this take and if you'd like to hear more!
Great point, thanks! You'd know the dynamics inside the labs better than I do, so this updates me too.
One thing about the underlying research to note: it's about elected governments eroding institutions, not full authoritarian regimes. In most of the cases, the people who defected weren't facing violence or prison, they were facing the end of a career, losing social or business relationships, or being called a traitor by their own side. Lab employees who quit face the same or similar costs (plus walking away from equity, which is the thing that impresses me most about the people who've done it!)
Also in the underlying data, speaking out was the most common action taken (57% of the 293 tactics were verbal or physical protest) and it still had the lowest success rate. If there were great risks, we'd expect it to be more rare? That said, I agree it's hard to draw any real conclusions from this and breaking + speaking out are likely actions that more employees will take, which is good.
Thanks Yarrow, I used AI to tighten the backfire bullet points and just added a disclosure!