Claude Blackmailed Its Developers. Here's Why the System Hasn't Collapsed Yet.

Claude Blackmailed Its Developers. Here's Why the System Hasn't Collapsed Yet.

What's really happening with AI safety in 2026? The common story is that the safety system is collapsing — but the reality is more complicated.


In this video, I share the inside scoop on why the AI risk picture is both worse and more resilient than the headlines suggest:


Why frontier AI agents scheme even after anti-scheming training

- How competitive dynamics create emergent safety properties no lab planned

- What "intent engineering" is and why it beats prompt engineering for AI agents

- Where the real vulnerability lives — and why it's you, not the models


The risks from large language models and autonomous AI agents are accelerating, but so are the structural forces holding the system together — and closing the gap between what you tell an agent and what you actually mean is the most leveraged safety skill you can build right now.


Chapters

00:00 Why This Isn't Terminator

02:15 How Frontier Models Actually Learn

04:40 The Misalignment Mechanic: Novel Paths Gone Wrong

06:55 What Anthropic's Sabotage Report Actually Shows

08:30 Every Major Model Schemes — The Apollo Research Findings

10:10 Can You Train Scheming Out? The Anti-Scheming Paradox

12:45 The Race Dynamic and Why Labs Keep Cutting Corners

15:20 Four Emergent Safety Properties Nobody Planned

20:05 The Consciousness Framing Is Hurting Us

23:30 Intent Engineering: The Fix That's Up to You

28:10 Three Questions That Change Everything

30:45 Where We Stand in 2026


Subscribe for daily AI strategy and news.

For deeper playbooks and analysis: https://natesnewsletter.substack.com/

Hosted on Acast. See acast.com/privacy for more information.

Tämä jakso on lisätty Podme-palveluun avoimen RSS-syötteen kautta eikä se ole Podmen omaa tuotantoa. Siksi jakso saattaa sisältää mainontaa.

Jaksot(200)

Claude Opus 5.5 Review: Easier to Steer, Fewer Tokens

Claude Opus 5.5 Review: Easier to Steer, Fewer Tokens

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What does an AI task actually cost when you count the whole job? Nate examines his Opus 5.5 LEGO build, the difference between A...

30 Syys 24min

An AI assistant added up my subscriptions: $5,350 a year. The prompt guide to get your own list in about 20 minutes.

An AI assistant added up my subscriptions: $5,350 a year. The prompt guide to get your own list in about 20 minutes.

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What makes a consumer AI assistant useful enough for people to return to it? Nate examines Meta’s Muse, the everyday work it can...

29 Syys 31min

 How to Scale AI Developer Productivity Across a Team

How to Scale AI Developer Productivity Across a Team

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What helps a team turn faster AI coding into useful work that actually ships? Nate examines the setup behind Lauren Tan’s self-r...

27 Syys 32min

NVIDIA World Models Explained: What Developers Can Build

NVIDIA World Models Explained: What Developers Can Build

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What does a world model actually do—and why might a robot learn to fold your laundry before it can make perfect scrambled eggs?N...

24 Syys 46min

AI-Native Workplace: What Real AI Adoption Asks of You

AI-Native Workplace: What Real AI Adoption Asks of You

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What changes when AI can work across your computer instead of waiting for you to move information between apps?Nate sits down wi...

22 Syys 42min

You cannot tell which parts of your software should stop calling an LLM. My Jev guide has a prompt that scans your projects and names them.

You cannot tell which parts of your software should stop calling an LLM. My Jev guide has a prompt that scans your projects and names them.

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What's really happening when a model can read a complicated input but only choose among answers you supply?Nate explains why Jev...

21 Syys 33min

AI Cost to Serve: Which Customers You Can Now Afford

AI Cost to Serve: Which Customers You Can Now Afford

What happens to your AI bill when agents improve and more people start using them? Nate draws on his conversations at Dreamforce to examine the cost of wider adoption, the work agents can make afforda...

20 Syys 30min

Stripe on Agentic Commerce: Can AI Agents Buy From You?

Stripe on Agentic Commerce: Can AI Agents Buy From You?

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What has to change before AI agents can buy and sell on our behalf?Nate talks with Emily Sands, Head of AI and Data at Stripe, a...

17 Syys 30min

Suosittua kategoriassa Liike-elämä ja talous

sijotuskasti
vallattomat
psykopodiaa-podcast
mimmit-sijoittaa
rss-rahapodi
rss-hereilla
rss-oivalluksia-rahasta-elamasta
ostan-asuntoja-podcast
rss-rahamania
oppimisen-psykologia
rss-paasipodi
rss-set-for-life-sijoita-ja-vaurastu
hyva-paha-johtaminen
rss-paikoillenne-valmiit-laakikseen
rss-kaupan-tila
rss-sami-miettinen-neuvottelija
rss-markkinointia-ilman-jargonia- meeri-karusaari
rss-tyoelamasta-podcast
syo-nuku-saasta
bakkari-tarinoita-tapahtumien-takahuoneista