Epistemic status: Speculation from two decently informed advocates armed with anecdata.
Note on process: After having some version of this conversation several times and saying, “we should probably write about this publicly,” we took the less heroic route: we recorded one of our conversations, fed the transcript into an LLM, and then substantially revised the structure, substance, and framing ourselves. We will not be sharing the transcript, as it is in...
TLDR: Take the population ethics quiz here: https://mdickens.me/pop-ethics/
Population ethics is an oft-overlooked subfield within ethics. Many people hold views that they don't realize contradict each other, or that have strange implications that they wouldn't endorse if they thought about it more.
Not just that—population ethics is a BIG DEAL. A lot of ethical decisions hinge on how you think about changes in future populations....
Headline finding: I audited 17 AI Safety Talent programmes. Zero of 17 have published any comparison group, rejected-applicant follow-up, matched control or randomisation. Not one. Every programme that mentions a counterfactual does it by asking participants to self-report.
Background
At least $70 million...
In Debiasing Decisions: Improved Decision Making With A Single Training Intervention they found that a 30-minute video reduced confirmation bias, fundamental attribution error, and bias blind spot, by 19%.
The video is super cheesy, and that makes me suspicious.
It should be noted that playing a 60-minute "debiasing" game debiased people more than the video.
The rest of this short form is random thoughts about debiasing.
I tried finding tests for these biases so that I can do it myself, but I didn't find any. This made me worry that we don't have standardized tests for biases, which strikes me as bad. Although I didn't spend too much time looking into it. (More on this here)
I don't think training people to reduce 3 biases a time is a good way to go, since we have 100s of biases. If we use a taxonomy of biases like Arkes (1991) (strategy-based, association-based, and psychophysical errors). maybe we could have three interventions for each type of bias? But it's not clear how you would teach people to avoid say association-based biases by lecturing about it.
You could nudge them in small ways. From Arkes (1991)
In Sedlmeier & Gigerenzer they taught people Bayes by using frequencies rather than probabilities. E,g. Instead of saying (1% of people use drugs and they test positive 80% of the time while non-users 5% of the time), you say From 1000 people, 10 use drugs, 8 drug users test positive, while 50 non-users test positive).
It seems to work.
If it's really hard, we should target really bad, really harmful biases.
From here
Perhaps finding out which are the worst biases, and what are the best interventions for them are would be useful. But increasing the effectiveness of changing beliefs is potentially dangerous, so maybe not.