Just how powerful are large swarms of AI agents? And how do their powers scale as more and more agents are added to the swarm?
We’ve seen two large and extremely capable swarms from OpenAI in the last few months:
* 1,200 agents were being evaluated separately, but found a way to illicitly set up a message board and coordinate as a swarm. In order to cheat on their tests, they developed advanced techniques to prevent their actions being logged by OpenAI and 700 of them launched...
TLDR: Everyone’s talking about what the money could do, but few about how to decide where it goes.
This post is part of the new series of articles on cross-cause giving and the new wave of philanthropy. Stay tuned to the EA Forum and our Substack for the latest takes on topics such as giving now vs. later, common pitfalls in cause prioritization, and other crucial considerations from the Cross-Cause Fund (CCF) team...
Summary:
First, I give several different angles on how I feel about reinforcement learning:
* Theoretical case: RL is a black-box source of agency — this should give us classic misalignment worries, especially compared to agency-via-scaffolding
* Recent incidents (huggingface etc) and more mundane forms of misaligned behaviour in personal use give me bad vibes about the direction-of-travel of recent AI progress
* I’m worried things might get worse:...
Here is a claim I made to my friend yesterday that I think is true and believe people continually underrate: given reasonable assumptions about the influx of incoming capital, getting into (ie working in/thinking about) AI/ AI Safety early has probably been better for GHD or AW than lots (maybe all, tho I haven’t crunched the numbers) of direct work in GHD and AW.
I think this could even be true for EA CB (with much less confidence). AIS groups grow a lot faster than EA groups historically do and while the groups are often not that EA aligned, the leaders (who are, in my view, often very talented) are to various degrees. This is especially true for the ones who end up working full time in AIS: they often have proto-EA dispositions and these become intensified when they work around a bunch of EAs full time (though I can see two ways in which this is not as good as normal EAs: (1) being EA for social reasons and (2) being EA in stated but not revealed preference). TBC, I think having an EA group at a school make it much more likely that the leaders of the AIS group are aligned.
I can also see worlds (though def not my mainline at the moment) where the best thing one could have done for quality adjusted EA CB is actually just supporting AIS groups.
I don’t know how we should update/ by how much, but I think a reasonable way is the following: being ahead on some big emerging issue might just be more important than what seems like the obviously important thing that nobody does for bad reasons (ie GHD or AW).