How scary is Claude Mythos? 303 pages in 21 minutes

How scary is Claude Mythos? 303 pages in 21 minutes

With Claude Mythos we have an AI that knows when it's being tested, can obscure its reasoning when it wants, and is better at breaking into (and out of) computers than any human alive. Rob Wiblin works through its 244-page System Card and 59-page Alignment Risk Update to explain why:

  • Mythos is a nightmare for computer security
  • It has arrived far ahead of schedule
  • It might be great news for alignment and safety
  • But 3 key problems mean we can’t take its alignment results at face value
  • Mythos isn’t building its replacement yet, probably
  • Anthropic staff are, for the first time, kinda scared of Claude
  • He's losing sleep

Learn more & full transcript: https://80k.info/mythos

This episode was recorded on April 9, 2026.

Chapters:

  • Why people are panicking about computer security (01:05)
  • Mythos could break out of containment (04:23)
  • Anthropic is losing billions in revenue by not releasing Mythos (06:21)
  • Mythos is actually the most aligned model to date, except… (07:48)
  • Mythos knows when it’s being tested (09:52)
  • Mythos can hide its thoughts (11:50)
  • Mythos can’t be trusted about whether it’s untrustworthy (14:02)
  • Does Mythos advance automated AI R&D? (17:03)
  • Mythos scares Anthropic (19:15)

Video and audio editing: Dominic Armstrong, Milo McGuire, Luke Monsour, and Simon Monsour
Camera operator: Dominic Armstrong
Production: Elizabeth Cox, Nick Stockton, and Katy Moore

Det här avsnittet är hämtat från ett öppet RSS-flöde och publiceras inte av Podme. Det kan innehålla reklam.

Avsnitt(343)

Jasmine Sun on what the people building AI really believe

Jasmine Sun on what the people building AI really believe

Many AI researchers believe mass job displacement is coming — and some even think there’s a chance their technology will kill everyone. But they’re building it anyway. Writer and journalist Jasmine Su...

21 Juli 0s

#247 – Anton Leicht on how middle powers avoid losing everything in a post-AI world

#247 – Anton Leicht on how middle powers avoid losing everything in a post-AI world

In a post-AGI world, can a country without access to frontier AI even be considered sovereign anymore?Anton Leicht says once frontier AI becomes a core economic input, the countries that own it will p...

14 Juli 0s

#246 – Sneha Revanur on how a small team of activists helped pass America's landmark AI safety laws

#246 – Sneha Revanur on how a small team of activists helped pass America's landmark AI safety laws

Six years ago, aged just 15, Sneha Revanur founded the AI advocacy nonprofit Encode AI — back when AI felt like a niche issue. Now the world’s caught up with her, and she’s ready to share everything s...

8 Juli 52min

We can guess what intergalactic war would look like. And strangely, it matters.

We can guess what intergalactic war would look like. And strangely, it matters.

Intergalactic war is probably billions of years away — yet physics can already tell us how it ends. And strangely that conclusion is relevant to decisions people have to make today.In this video, Rob ...

18 Juni 15min

How AI could create the world’s biggest problems (article by Zershaaneh Qureshi)

How AI could create the world’s biggest problems (article by Zershaaneh Qureshi)

Imagine you’re living 15,000 years ago. Your people are hunter-gatherers and you sleep under the stars. If someone told you humans would one day build cities with millions of people, fly through the a...

11 Juni 1h 29min

#245 – Rohin Shah on what it's really like to run AGI safety at Google DeepMind (and where I disagree with 'doomers')

#245 – Rohin Shah on what it's really like to run AGI safety at Google DeepMind (and where I disagree with 'doomers')

Most people working on AI safety think without a massive effort AI systems will probably end up with goals catastrophically different from humanity’s. Today’s guest, Rohin Shah — head of AGI Safety an...

2 Juni 2h 48min

What makes for a dream job? | Benjamin Todd

What makes for a dream job? | Benjamin Todd

What actually makes a job fulfilling? It's not what most career advice tells you. "Follow your passion" sounds inspiring, but it's misleading — and the research backs that up.Drawing on hundreds of st...

28 Maj 28min

#244 – Benjamin Todd on how we’re updating our career advice for the strangest time in history

#244 – Benjamin Todd on how we’re updating our career advice for the strangest time in history

The average career is 80,000 hours long. With AI advancing so rapidly, the hours you have left in your career matter more than ever.Some leading AI researchers think there’s a 10% chance that AI syste...

26 Maj 1h 6min

Populärt inom Utbildning

historiepodden-se
det-skaver
rss-bara-en-till-om-missbruk-medberoende-2
not-fanny-anymore
nu-blir-det-historia
harrisons-dramatiska-historia
rss-viktmedicinpodden
allt-du-velat-veta
johannes-hansen-podcast
rss-basta-livet
rikatillsammans-om-privatekonomi-rikedom-i-livet
sex-pa-riktigt-med-marika-smith
i-vantan-pa-katastrofen
sa-in-i-sjalen
rss-traningsklubben
rss-max-tant-med-max-villman
sektledare
henry-laser-wikipedia
rss-sjalsligt-avkladd
rss-rummet-podcast