Epistemic status: Speculation from two decently informed advocates armed with anecdata.
Note on process: After having some version of this conversation several times and saying, “we should probably write about this publicly,” we took the less heroic route: we recorded one of our conversations, fed the transcript into an LLM, and then substantially revised the structure, substance, and framing ourselves. We will not be sharing the transcript, as it is in...
I want to start by noting that I am amazed and somewhat in awe of what's been accomplished by the EA animal welfare movement. I am enthusiastically supportive of all campaigns that make living and dying conditions less bad for farmed animals.
I worked on such campaigns for almost 20 years, and I'm proud of that work.
That said, I do think it's a strategic error that the EA animal movement has so thoroughly moved on from advocacy for animal liberation and diet change.&n...
The swarm has made me feel AI risk in my bones for the first time
Until the Hugging Face / OpenAI swarm thing, I think I deep down felt pretty skeptical that extinction or serious loss of control scenarios were plausible in the next few years.
I’ve been lurking on LessWrong for years, and I remember being freaked out by Yudkowsky’s ‘List of Lethalities’ in 2022, and thinking that this did seem intellectually convincing, and maybe I should do something about it. But it didn...
From SE Geyges: Is METR a meaningful check on Anthropic?.
Part of the reason we are in this position is because nobody outside the EA community cares about AI safety enough to fund it or work in it. I’m very sympathetic to this; some of these close connections are inevitable.
However, it can also be true that these connections are
completelyunacceptable (edit), and that Anthropic should be trying harder than they are to find or create auditors that are genuinely independent. How do we do that?I think METRs checks are meaningful and maybe the best we have at the moment. They are also like you say compromised and the conflict of interests are immense with huge personal overlap between the labs and safety orgs, and funding streams too.
Geyges seems largely correct, but if we can't convince governments to regulate properly its better METR is in there doing it. After all METR exposed more about the hugging face hack than Open AI did on its own.
I don't think a framing of "these connections are completely unacceptable" is helpful given these problems. I think "compromised and far from ideal" is a better framing. Its better to do something than do nothing. I agree its best if government installed internal auditors like they do for banks, but that ain't happening any time soon.
Companies like Anthropic and Open AI are selfish animals by nature. They may have moments where good humans inside might do the right thing, but fundamentally they thirst for profit and growth. After IPO this will only get worse. We should never expect a company to regulate itself or its industry. Self regulation for harmful companies is a terrible idea and never works.
How can ANthropic "create" a genuinely independent auditor? this seems impossible, almost and Oxymoron.