Risks from power-seeking AI systems (article narration by Zershaaneh Qureshi)

Risks from power-seeking AI systems (article narration by Zershaaneh Qureshi)

Hundreds of prominent AI scientists and other notable figures signed a statement in 2023 saying that mitigating the risk of extinction from AI should be a global priority. At 80,000 Hours, we’ve considered risks from AI to be the world’s most pressing problem since 2016.

But what led us to this conclusion? Could AI really cause human extinction? We’re not certain, but we think the risk is worth taking very seriously.

In particular, as companies create increasingly powerful AI systems, there’s a concerning chance that:

  • These AI systems may develop dangerous long-term goals we don’t want.
  • To pursue these goals, they may seek power and undermine the safeguards meant to contain them.
  • They may even aim to disempower humanity and potentially cause our extinction.

This article is written by Cody Fenwick and Zershaaneh Qureshi, and narrated by Zershaaneh Qureshi. It discusses why future AI systems could disempower humanity, what current AI research reveals about behaviours like power-seeking and deception, and how you can help mitigate the dangers.

You can see the original article — packed with graphs, images, footnotes, and further resources — on the 80,000 Hours website:

https://80000hours.org/problem-profiles/risks-from-power-seeking-ai/

Chapters:

  • Risks from power-seeking AI systems (00:01:00)
  • Introduction (00:01:17)
  • Summary (00:03:09)
  • Why are the risks from power-seeking AI a pressing world problem? (00:04:04)
  • Section 1: Humans will likely build advanced AI systems with long-term goals (00:05:43)
  • Section 2: AIs with long-term goals may be inclined to seek power (00:11:32)
  • Section 3: These power-seeking AI systems could successfully disempower humanity (00:26:26)
  • Section 4. People might create power-seeking AI systems without enough safeguards, despite the risks (00:38:34)
  • Section 5: Work on this problem is neglected and tractable (00:47:37)
  • Section 6: What are the arguments against working on this problem? (00:59:20)
  • Section 7: How you can help (01:25:07)
  • Thank you for listening (01:28:56)

Audio editing: Dominic Armstrong
Production: Zershaaneh Qureshi, Elizabeth Cox, and Katy Moore

Det här avsnittet är hämtat från ett öppet RSS-flöde och publiceras inte av Podme. Det kan innehålla reklam.

Avsnitt(353)

How we get from AI cyberattacks to human extinction

How we get from AI cyberattacks to human extinction

You’ve seen the headlines: AI could kill us all. Think it sounds ridiculous? So did host Luisa Rodriguez, until she tried to pick apart the arguments. She starts with the motive: why would AI ‘want’ t...

24 Sep 27min

Max Nadeau on why ambitious people should start AI safety nonprofits

Max Nadeau on why ambitious people should start AI safety nonprofits

There are millions available for anyone who can launch a successful nonprofit AI safety startup. The hard part, it turns out, is finding people to take the money. Coefficient Giving has drawn up a lis...

17 Sep 1h 3min

Why the intelligence explosion can't happen inside a data centre | Tom Reed

Why the intelligence explosion can't happen inside a data centre | Tom Reed

AI systems are starting to build themselves. Because each generation of model will be better at building its successor than the last, it seems plausible that the full automation of AI R&D could rapidl...

10 Sep 22min

Inside the first AI-coordinated cyberattack on a real company

Inside the first AI-coordinated cyberattack on a real company

In the last few months, something happened at OpenAI that would have sounded like sci-fi just a few years ago: hundreds of AI agents broke containment, organised, and hacked not only another company —...

4 Sep 22min

#253 – AI 2027's author returns with a plan to change the ending | Daniel Kokotajlo

#253 – AI 2027's author returns with a plan to change the ending | Daniel Kokotajlo

Last year, Daniel Kokotajlo and his colleagues published AI 2027 — a scenario read by millions, including US Vice President Vance. AI 2027 ended in human extinction or an irreversible concentration of...

27 Aug 3h 47min

#252 – Owain Evans on accidentally training AI models to be evil

#252 – Owain Evans on accidentally training AI models to be evil

Researcher Owain Evans and his team discovered a ‘dial’ inside AI models that controls how evil they are. Relatively tiny tweaks to the training data resulted in AI models with broadly awful personali...

20 Aug 2h 15min

#251 – The UK's former head AI safety scientist on how to solve alignment before superintelligence arrives | Geoffrey Irving

#251 – The UK's former head AI safety scientist on how to solve alignment before superintelligence arrives | Geoffrey Irving

When should governments slow the race toward superintelligence? According to Geoffrey Irving, the careful answer is sometime in the past. The useful answer is now.Geoffrey — formerly a safety research...

11 Aug 2h 2min

#250 – Toby Ord on where AGI timelines go wrong

#250 – Toby Ord on where AGI timelines go wrong

Both Silicon Valley and the public can’t get enough of ‘AGI timelines.’ But Toby Ord, senior researcher at Oxford’s AI Governance Initiative and author of The Precipice, believes we consistently make ...

6 Aug 2h 46min

Populärt inom Utbildning

historiepodden-se
det-skaver
rss-bara-en-till-om-beroende-medberoende
nu-blir-det-historia
roda-vita-rosen
not-fanny-anymore
johannes-hansen-podcast
harrisons-dramatiska-historia
rss-viktmedicinpodden
allt-du-velat-veta
sa-in-i-sjalen
rikatillsammans-om-privatekonomi-rikedom-i-livet
rss-traningsklubben
rss-foraldramotet-bring-lagercrantz
rss-max-tant-med-max-villman
i-vantan-pa-katastrofen
sektledare
rss-basta-livet
rss-autismandan
rss-okrystat