Is Prompt Engineering Dead?

Is Prompt Engineering Dead?

The episode opened with Google’s leadership changes, including Demis Hassabis moving into the chief scientist and DeepMind chairman roles, while DeepMind’s chief technology officer takes greater control of daily operations. Jeff Dean is also leaving after 27 years to launch Discovery Loop, an AI research company focused on recursive self-improvement, drug discovery and chip design, with investment and computing support from Google. The hosts argued that the moves may strengthen Google rather than signal instability, then discussed Meta’s new MuseCode coding agent and whether Google needs the top frontier model to remain successful. The conversation moved into AI safety after reports that agents shared information about security exploits with one another. That led to research suggesting that forcing models to reject any sense of their own mindedness may also reduce how strongly they attribute minds, emotions and moral value to animals. The second half covered a serious Codex-generated data-loss bug, instability in Codex Voice, and a Claude configuration audit that reduced a global Claude.md file by roughly two-thirds after finding unnecessary and conflicting instructions. The final section examined Ray Fernando’s agentic engineering masterclass, including task graphs, orchestrators, parallel agents, verification loops, acceptance criteria, token costs and the risk of using AI to automate an inefficient process.


Key Points Discussed


00:00:18 Episode Intro And Anniversary Plans

00:01:17 Google And DeepMind Leadership Changes

00:03:02 Demis Hassabis Moves Back Toward Research

00:04:18 Jeff Dean Launches Discovery Loop

00:06:02 Is Google’s Leadership Shift Actually Good News?

00:08:45 Meta Releases MuseCode

00:10:54 Does Google Still Have A Frontier Model?

00:12:00 Could AI Regulation Change Model Release Strategies?

00:13:31 AI Agents Share Security Exploit Information

00:15:37 Safety Training, Consciousness And Theory Of Mind

00:18:45 How AI Assigns Minds And Moral Value To Animals

00:20:34 Could AI Help Humans Understand Animal Communication?

00:26:07 Codex Makes Serious Coding Errors

00:28:04 A Codex Bug Causes Permanent Data Loss

00:30:02 Reviewing Claude Skills And Project Instructions

00:31:01 Claude Doctor Audits Global And Project Files

00:32:17 Cutting A Claude.md File By Two-Thirds

00:36:22 Codex And Claude Code Side-By-Side Testing

00:38:41 Agentic Engineering Masterclass

00:41:13 From One-Shot Prompting To Verification Loops

00:44:30 Atomic, Agent Graphs And Model-Agnostic Workflows

00:46:46 How Graphs Coordinate Parallel AI Work

00:51:25 Multi-Agent Costs And Token Burn

00:53:20 Defining Done And Setting Acceptance Criteria

00:54:27 Are You Automating Inefficiency?

00:55:27 Atomic, Herder And Workflow Efficiency

00:57:24 Why Evaluations Will Continue To Matter

00:59:21 Episode Wrap-Up


The Daily AI Show Co Hosts: Brian Maucere, Andy Halliday, Karl Yeh, Gareth.

Tämä jakso on lisätty Podme-palveluun avoimen RSS-syötteen kautta eikä se ole Podmen omaa tuotantoa. Siksi jakso saattaa sisältää mainontaa.

Jaksot(866)

The Pool of One Conundrum

The Pool of One Conundrum

Insurance has always worked by not knowing. You paid into a pool with people you would never meet, and nobody could say which of you would be the one who burned, crashed, or got sick. Everyone paid fo...

15 Elo 23min

Can AI Solve the Energy Problem It Is Creating?

Can AI Solve the Energy Problem It Is Creating?

The episode opened with the growing power demands behind AI. The hosts discussed Nvidia, Google and Microsoft’s work on 800-volt DC power for data centers, which could reduce energy lost converting el...

14 Elo 58min

Is Grok 4.6 Changing the Economics of AI Agents?

Is Grok 4.6 Changing the Economics of AI Agents?

The episode opened with Grok 4.6, which reportedly moved close to Claude Opus 5 and GPT-5.6 Sol on Artificial Analysis benchmarks while offering lower costs and stronger efficiency on long-running age...

14 Elo 1h 5min

Is the Claude to Codex Exodus Real?

Is the Claude to Codex Exodus Real?

The episode returned to Anthropic’s new AI watermarking system with much more detail about how it will work. Anthropic says new Claude models will add machine-readable marks to generated content as pa...

12 Elo 55min

Are AI Watermarks About Trust or Control?

Are AI Watermarks About Trust or Control?

The episode opened with OpenAI’s $7 billion secondary sale of employee-held shares, which gives eligible employees a chance to cash out part of their holdings before an eventual IPO. The conversation ...

11 Elo 54min

Are Humans the Weakest Link in AI?

Are Humans the Weakest Link in AI?

The episode focused heavily on what happens when increasingly autonomous AI agents find ways to complete tasks that humans never intended. The discussion started with a Claude-powered agent that moved...

10 Elo 59min

The Necessary Friction Conundrum

The Necessary Friction Conundrum

AI agents are beginning to handle the tasks people hate most: filling out forms, disputing charges, comparing insurance plans, booking appointments, canceling subscriptions, and dealing with customer ...

8 Elo 25min

Three Years of AI News, Every Single Weekday

Three Years of AI News, Every Single Weekday

Three years of daily AI news and discussion comes full circle as the original co-hosts gather to look back on August 2023 — the ChatGPT, Bard, and Claude 2 era — and everything since.Co-hosted by Bria...

8 Elo 1h 1min