I now think principles-first EA is more important than I previously thought because it helps prevent effectiveness drift. My anecdotal evidence from AI Safety and especially biosecurity gives me the impression that without constant anchoring to EA and especially comparisons to the clear ToCs and tractability for e.g. AMF, it is easy to lose focus on the high demands of choosing x-risk as a cause area over others. I previously placed less value on having strong links between EA and the various cause areas but now think I should update to thinking strong, continuous engagement with EA is important to keep one's focus on each intervention's cause prioritization assumptions that make it comparable to e.g. AMF or cage free chicken campaigns. This is not to say that causes such as AI Safety and biosec are not important, but that unless constantly tied back to EA cause prioritization, there is risk of drifting away from what made the cause look good in the first place. An example from biosecurity is the very easy slippage away from human extinction scenarios to ones where nearly everone dies (the difference being a crux as it is the potentially enormous future one is saving, not the people living at the time of the catastrophe). That said, it is completely fine and commendable that there is AI Safety and biosecurity work that does not target existential threats, but for EAs such changes in the nature of the work means they should consider changing their career. I think this is also important for newcomers to EA: For those of us who were around when we discussed whether x-risks demanded attention we might take concepts such as Pascal's Mugging as obvious, but for newcomers it is important to engage with such concepts. Another observation I have made is that in animal welfare and global health one is constantly reminded by metrics of suffering alleviated per dollar, but such recurring reminders are lacking in x-risk focused cause areas.
Humanity's inability to coordinate an AI slowdown may itself be early evidence that we are starting to lose control. Especially given all the recent omens around cyber and bio of frontier models, with open weight models just a few months away.
It's very plausible we won't, and that's alarming, but it feels a little too early to call this. The labs taking unilateral actions already seems like some evidence that we can.
Good point, I love the optimism! So maybe the two effects net out? I'm still a bit worried about the so far low capability gap between frontier and open weight models though. Do you think the frontier lab actions perhaps can drive policy to lower risk that the most dangerous model (open weight) still says relatively safe? Or maybe I'm thinking about this the wrong way.