Headline finding: I audited 17 AI Safety Talent programmes. Zero of 17 have published any comparison group, rejected-applicant follow-up, matched control or randomisation. Not one. Every programme that mentions a counterfactual does it by asking participants to self-report.
Background
At least $70 million...
Author: Grace Ryba (she/her), Executive Director
TL;DR:
* BluePerch is a new animal welfare grantmaker.
* Grant applications aren't open yet - we're aiming to open applications in roughly mid-2027.
* We welcome early expressions of interest from potential grantseekers so we can notify you when applications open - express your interest here or see further details below.
* We're also seeking boa...
In celebration of still being alive and fighting, we are giving away 1,000 Amazon e-books of “If Anyone Builds It, Everyone Dies”. Feel free to send a copy to yourself, a loved one, or a friend—we need all hands on deck.
Today marks exactly one year since If Anyone Builds It, Everyone Dies: Why Superhuman AI Would Kill Us All, by Eliezer Yudkowsky and Nate Soare...
I fear the weird hugbox EAs do towards their critics in order to signal good faith means over time a lot of critics just end up not being sharpened in their arguments.
I feel pretty strongly against "weird hugboxing" but I think the main negative effect is an erosion of our own epistemic standards and a reduction in the degree to which we can epistemically defer to one another. I want the EA community to consist of people whose pronouncements I can fully trust, rather than have to wonder if they are saying something because it reflects their considered judgment on that topic or instead because they are signaling good faith, "steelmanning", etc.
What's the comparative?
I think an inviting form of decoupling norms where it's fractured in chains. I don't think decoupling norms work when both parties don't opt-in and so people should switch to the dominant norm of the sphere. An illustrative example is as follows:
Some EAs would see this as being a motte-and-bailey instead of getting to the crux but cruxes can be asymmetric in that different critics combine claims together (e.g. the "woke" combining with more centrist sensibilities deontologists). But I think explanations which are done well are persuasive because they reframe truth-seeking ideas within accessible language that dissolve cruxes to seek agreement and cooperation.
Another illustration on the macro-level of the comparative:
To be clear, there are harms with trying to be persuasive (e.g. sophistry, lying, motivated reasoning etc.). But sometimes being persuasive is about speaking the argumentative language of another side.
This is a great comment and I think made me get much more of what you're driving at than the (much terser) top-level comment.
Yeah I should have written more but I try to keep my short form casual to make the barrier of entry lower and to allow for expansions based on different reader's issues.
What do you mean by "resource" here?
Examples of resources that come to mind:
Again this is predicated on good faith critics.
I think something like 30% hugboxing is good. I think that the cases where you see it maybe it could happen less, but a lot of the time I think we are too brutal to non-rationalist critics.
It's really tiring to criticise and I think it's nice to have someone listen and engage at least a bit. If I move straight to "here is how I disagree" I think I lose out on useful criticism in the long run.
But that's conditional on people not interpreting the hugboxing as a tactic/weird norm. E.g. mormon missionaries being nice to people doesn't elicit the same response as a person off the street because they adjust their set point.
Can you give examples of hugboxing you don't like?
Because my internal response is "people think we are too aggressive/dismissive" rather than "people think we listen to them but in a weird/patronising way" and if you mean internally you don't like it, then I am confused as to why you read it.
On the forum I agree hugboxing is worse.
Poll!