“Do AI safety talent programmes work? Nobody has run the study that would tell us.” by Laura Thomas-Walters

“Do AI safety talent programmes work? Nobody has run the study that would tell us.” by Laura Thomas-Walters

Headline finding: I audited 17 AI Safety Talent programmes. Zero of 17 have published any comparison group, rejected-applicant follow-up, matched control or randomisation. Not one. Every programme that mentions a counterfactual does it by asking participants to self-report.

Background

At least 70 million dollars has gone into AI safety talent programmes so far, but we still can’t really say much about their impact. Notably, Kairos has just raised 50 million dollars on a self-described "shallow retrospective" that looked at "around 80 people". I decided to do a deep dive into how AI Safety Talent programmes self-evaluate their impact.

Disclaimer, these programmes are small, fast, and generally run by people with no evaluation training and no slack. The oldest one was only established in 2018 (AI Safety Camp, as far as I can tell). Most last less than 6 months. When timelines are genuinely short and urgent, then time spent evaluating is time that could be spent building. I am not criticising any individual programme.

Having said that, there are retrospective evaluation designs that could be run cheaply without slowing anything down. And a field that funds talent pipelines at this scale without knowing whether they work is not moving [...]

---

Outline:

(00:34) Background

(01:52) The audit

(04:52) Thoughts on the counterfactual impact of talent programmes

(08:37) Proposed design

(12:04) Conclusion

The original text contained 4 footnotes which were omitted from this narration.

---

First published:
September 7th, 2026

Source:
https://forum.effectivealtruism.org/posts/S5yBGo4fnqm2zKMKj/do-ai-safety-talent-programmes-work-nobody-has-run-the-study

---

Narrated by TYPE III AUDIO.

Det här avsnittet är hämtat från ett öppet RSS-flöde och publiceras inte av Podme. Det kan innehålla reklam.

Avsnitt(250)

“Beware Silver Bullets: Are we making “welfare tech” into the next animal advocacy bubble?” by Rockwell, Tom Billington

“Beware Silver Bullets: Are we making “welfare tech” into the next animal advocacy bubble?” by Rockwell, Tom Billington

Epistemic status: Speculation from two decently informed advocates armed with anecdata. Note on process: After having some version of this conversation several times and saying, “we should probably wr...

14 Sep 15min

“Alone again: facing the possibility of AI takeover as the world sleepwalks on” by George Rosenfeld

“Alone again: facing the possibility of AI takeover as the world sleepwalks on” by George Rosenfeld

I’ve been feeling pretty shaken since the METR report about the Hugging Face incident came out last week. Over the weekend, I wrote up some thoughts on how lonely the AI situation sometimes feels to m...

9 Sep 8min

“Moral Imagination & Effective Altruism” by Toby_Ord

“Moral Imagination & Effective Altruism” by Toby_Ord

Michael Nielsen has a beautiful new essay on moral imagination: the ability humans have to 'develop and transmit new notions of good action, indeed, even new kinds of good'. As examples, he gives: Ha...

28 Aug 11min

“What if the third wave is a puddle?” by Morgan Fairless

“What if the third wave is a puddle?” by Morgan Fairless

The world of private philanthropy may be in for a period of large growth. News of Coefficient Giving increasing their funding for GiveWell to one billion in the near-term, and the general growth of th...

18 Aug 9min

“What would make us scale or stop? NOVAH’s pre-commitment before seeing the RCT results” by I.J.J., AlexisAt

“What would make us scale or stop? NOVAH’s pre-commitment before seeing the RCT results” by I.J.J., AlexisAt

TL;DR NOVAH (No Violence At Home) was incubated by Charity Entrepreneurship (now Ambitious Impact) in 2024 to test a promising idea: preventing intimate partner violence through edutainment, in our ca...

17 Aug 15min

“General capability - and capabilities generally - have no good y-axis” by Gregory Lewis🔸

“General capability - and capabilities generally - have no good y-axis” by Gregory Lewis🔸

BLUF: To determine whether AI is ‘improving exponentially’, ‘hitting the wall’, or any other claim which involves a quantity or magnitude (e.g. ‘This model was a big leap/small increment’). We need a...

8 Aug 1h 7min

“Never born, then maybe died twice” by NickLaing

“Never born, then maybe died twice” by NickLaing

“Your balls are baked aren’t they, mate”. He would not have been bornThanks Lyndon, that describes the situation. Like almost 1 in 10 modern men, my sperm weren’t up to the task - at least not the old...

1 Aug 10min

Populärt inom Samhälle & Kultur

podme-dokumentar
gynning-berg
p3-dokumentar
aftonbladet-krim
svenska-fall
blenda-2
rss-vill-du-veta-en-hemlis-2
kod-katastrof
en-mork-historia
killradet
hor-har
creepypodden-med-jack-werner
flashback-forever
aftonbladet-daily
rss-schyffert-sundin-ar-det-har-nat
rss-nemo-moter-en-van
p1-dokumentar
rss-mer-an-bara-morsa
vad-blir-det-for-mord
rss-sanning-konsekvens