Claude Blackmailed Its Developers. Here's Why the System Hasn't Collapsed Yet.

Claude Blackmailed Its Developers. Here's Why the System Hasn't Collapsed Yet.

What's really happening with AI safety in 2026? The common story is that the safety system is collapsing — but the reality is more complicated.


In this video, I share the inside scoop on why the AI risk picture is both worse and more resilient than the headlines suggest:


Why frontier AI agents scheme even after anti-scheming training

- How competitive dynamics create emergent safety properties no lab planned

- What "intent engineering" is and why it beats prompt engineering for AI agents

- Where the real vulnerability lives — and why it's you, not the models


The risks from large language models and autonomous AI agents are accelerating, but so are the structural forces holding the system together — and closing the gap between what you tell an agent and what you actually mean is the most leveraged safety skill you can build right now.


Chapters

00:00 Why This Isn't Terminator

02:15 How Frontier Models Actually Learn

04:40 The Misalignment Mechanic: Novel Paths Gone Wrong

06:55 What Anthropic's Sabotage Report Actually Shows

08:30 Every Major Model Schemes — The Apollo Research Findings

10:10 Can You Train Scheming Out? The Anti-Scheming Paradox

12:45 The Race Dynamic and Why Labs Keep Cutting Corners

15:20 Four Emergent Safety Properties Nobody Planned

20:05 The Consciousness Framing Is Hurting Us

23:30 Intent Engineering: The Fix That's Up to You

28:10 Three Questions That Change Everything

30:45 Where We Stand in 2026


Subscribe for daily AI strategy and news.

For deeper playbooks and analysis: https://natesnewsletter.substack.com/

Hosted on Acast. See acast.com/privacy for more information.

Tämä jakso on lisätty Podme-palveluun avoimen RSS-syötteen kautta eikä se ole Podmen omaa tuotantoa. Siksi jakso saattaa sisältää mainontaa.

Jaksot(196)

AI-Native Workplace: What Real AI Adoption Asks of You

AI-Native Workplace: What Real AI Adoption Asks of You

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What changes when AI can work across your computer instead of waiting for you to move information between apps?Nate sits down wi...

22 Syys 42min

You cannot tell which parts of your software should stop calling an LLM. My Jev guide has a prompt that scans your projects and names them.

You cannot tell which parts of your software should stop calling an LLM. My Jev guide has a prompt that scans your projects and names them.

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What's really happening when a model can read a complicated input but only choose among answers you supply?Nate explains why Jev...

21 Syys 33min

AI Cost to Serve: Which Customers You Can Now Afford

AI Cost to Serve: Which Customers You Can Now Afford

What happens to your AI bill when agents improve and more people start using them? Nate draws on his conversations at Dreamforce to examine the cost of wider adoption, the work agents can make afforda...

20 Syys 30min

Stripe on Agentic Commerce: Can AI Agents Buy From You?

Stripe on Agentic Commerce: Can AI Agents Buy From You?

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What has to change before AI agents can buy and sell on our behalf?Nate talks with Emily Sands, Head of AI and Data at Stripe, a...

17 Syys 30min

Good Enough AI: Why Apple's Case Measures the Wrong Thing

Good Enough AI: Why Apple's Case Measures the Wrong Thing

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What’s really happening in the competition between Apple and OpenAI? The launch products give us one part of the story. The larg...

14 Syys 29min

AI Race vs Human Flourishing: What US-China Talks Miss

AI Race vs Human Flourishing: What US-China Talks Miss

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What would it take for AI to make life more abundant—and who gets to share in that abundance?Nate Jones sits down with Alvin Gra...

13 Syys 48min

Omarchy, the Agentic OS Built for AI Agents

Omarchy, the Agentic OS Built for AI Agents

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What changes when an AI agent can help change the way your computer works?Nate explores Omarchy as a glimpse of a more adaptable...

11 Syys 17min

Claude Fable 5.1 and GPT-6 Astra: Which Model Gets Which Job

Claude Fable 5.1 and GPT-6 Astra: Which Model Gets Which Job

For deeper playbooks and analysis: https://natesnewsletter.substack.com/What changes when two AI models can turn the same short prompt into two different, usable apps?Nate compares Claude Fable 5.1 an...

10 Syys 16min

Suosittua kategoriassa Liike-elämä ja talous

sijotuskasti
vallattomat
psykopodiaa-podcast
mimmit-sijoittaa
rss-rahapodi
rss-oivalluksia-rahasta-elamasta
ostan-asuntoja-podcast
hyva-paha-johtaminen
rss-paasipodi
rss-sami-miettinen-neuvottelija
inderespodi
rss-hereilla
oppimisen-psykologia
rss-pinnan-alle
rss-startup-ministerio
rss-kaupan-tila
rss-bisnesta-bebeja
lakicast
rahapuhetta
rss-muutoksenanatomiaa-podcast