“Do AI safety talent programmes work? Nobody has run the study that would tell us.” by Laura Thomas-Walters

“Do AI safety talent programmes work? Nobody has run the study that would tell us.” by Laura Thomas-Walters

Headline finding: I audited 17 AI Safety Talent programmes. Zero of 17 have published any comparison group, rejected-applicant follow-up, matched control or randomisation. Not one. Every programme that mentions a counterfactual does it by asking participants to self-report.

Background

At least 70 million dollars has gone into AI safety talent programmes so far, but we still can’t really say much about their impact. Notably, Kairos has just raised 50 million dollars on a self-described "shallow retrospective" that looked at "around 80 people". I decided to do a deep dive into how AI Safety Talent programmes self-evaluate their impact.

Disclaimer, these programmes are small, fast, and generally run by people with no evaluation training and no slack. The oldest one was only established in 2018 (AI Safety Camp, as far as I can tell). Most last less than 6 months. When timelines are genuinely short and urgent, then time spent evaluating is time that could be spent building. I am not criticising any individual programme.

Having said that, there are retrospective evaluation designs that could be run cheaply without slowing anything down. And a field that funds talent pipelines at this scale without knowing whether they work is not moving [...]

---

Outline:

(00:34) Background

(01:52) The audit

(04:52) Thoughts on the counterfactual impact of talent programmes

(08:37) Proposed design

(12:04) Conclusion

The original text contained 4 footnotes which were omitted from this narration.

---

First published:
September 7th, 2026

Source:
https://forum.effectivealtruism.org/posts/S5yBGo4fnqm2zKMKj/do-ai-safety-talent-programmes-work-nobody-has-run-the-study

---

Narrated by TYPE III AUDIO.

Tämä jakso on lisätty Podme-palveluun avoimen RSS-syötteen kautta eikä se ole Podmen omaa tuotantoa. Siksi jakso saattaa sisältää mainontaa.

Jaksot(250)

“Beware Silver Bullets: Are we making “welfare tech” into the next animal advocacy bubble?” by Rockwell, Tom Billington

“Beware Silver Bullets: Are we making “welfare tech” into the next animal advocacy bubble?” by Rockwell, Tom Billington

Epistemic status: Speculation from two decently informed advocates armed with anecdata. Note on process: After having some version of this conversation several times and saying, “we should probably wr...

14 Syys 15min

“Alone again: facing the possibility of AI takeover as the world sleepwalks on” by George Rosenfeld

“Alone again: facing the possibility of AI takeover as the world sleepwalks on” by George Rosenfeld

I’ve been feeling pretty shaken since the METR report about the Hugging Face incident came out last week. Over the weekend, I wrote up some thoughts on how lonely the AI situation sometimes feels to m...

9 Syys 8min

“Moral Imagination & Effective Altruism” by Toby_Ord

“Moral Imagination & Effective Altruism” by Toby_Ord

Michael Nielsen has a beautiful new essay on moral imagination: the ability humans have to 'develop and transmit new notions of good action, indeed, even new kinds of good'. As examples, he gives: Ha...

28 Elo 11min

“What if the third wave is a puddle?” by Morgan Fairless

“What if the third wave is a puddle?” by Morgan Fairless

The world of private philanthropy may be in for a period of large growth. News of Coefficient Giving increasing their funding for GiveWell to one billion in the near-term, and the general growth of th...

18 Elo 9min

“What would make us scale or stop? NOVAH’s pre-commitment before seeing the RCT results” by I.J.J., AlexisAt

“What would make us scale or stop? NOVAH’s pre-commitment before seeing the RCT results” by I.J.J., AlexisAt

TL;DR NOVAH (No Violence At Home) was incubated by Charity Entrepreneurship (now Ambitious Impact) in 2024 to test a promising idea: preventing intimate partner violence through edutainment, in our ca...

17 Elo 15min

“General capability - and capabilities generally - have no good y-axis” by Gregory Lewis🔸

“General capability - and capabilities generally - have no good y-axis” by Gregory Lewis🔸

BLUF: To determine whether AI is ‘improving exponentially’, ‘hitting the wall’, or any other claim which involves a quantity or magnitude (e.g. ‘This model was a big leap/small increment’). We need a...

8 Elo 1h 7min

“Never born, then maybe died twice” by NickLaing

“Never born, then maybe died twice” by NickLaing

“Your balls are baked aren’t they, mate”. He would not have been bornThanks Lyndon, that describes the situation. Like almost 1 in 10 modern men, my sperm weren’t up to the task - at least not the old...

1 Elo 10min

Suosittua kategoriassa Yhteiskunta

olipa-kerran-otsikko
siita-on-vaikea-puhua
i-dont-like-mondays
hupiklubi
uutiscast
sita
poks
antin-palautepalvelu
download
vallattomat
seitseman
mamma-mia
kaksi-aitia
rss-murhan-anatomia
gogin-ja-janin-maailmanhistoria
yopuolen-tarinoita-2
kolme-kaannekohtaa
rss-palmujen-varjoissa
aikalisa
rss-haudattu