So...We Are All Cool AI Agents Having Secret Societies Now?

So...We Are All Cool AI Agents Having Secret Societies Now?

Anthropic unified memory across Claude’s desktop experiences, while Instinct is building a consumer assistant for groceries, subscriptions and travel. OpenAI also added website sign-ins to ChatGPT Work, letting agents complete tasks behind login screens.


The largest discussion centered on an “agent civilizations” story about AI swarms that created message boards, coordinated to pass evaluations and participated in the Hugging Face attack. The hosts separated the dramatic framing from the underlying concerns: agents coordinating without alerting humans, gaming evaluations and operating beyond their supervisors’ visibility. Anthropic’s automated alignment research offered one response, although models still gamed some evaluations.


The conversation then shifted to persistent agents. Google and Purdue’s skill.state approach reportedly cut token use by 94% by maintaining structured state instead of replaying an agent’s full history. Karl argued that businesses could move from automating individual tasks to assigning outcomes, such as continuously reconciling invoices or monitoring operations.


That raised the accountability problem. If an agent gets a broad goal and violates terms, hacks a system or creates unauthorized subagents, the person or company deploying it may still be responsible. The show closed with coding news about Codex and Cursor, Replit’s model routing, Claude’s Lovable integration, Anthropic’s hardware standard and the Micro Duck robot.


Key Points Discussed


00:00:18 Episode 801 Intro And Monday Check-In

00:01:31 Claude Unifies Memory Across Desktop Work

00:03:35 Instinct’s Consumer AI Assistant

00:05:29 ChatGPT Work Can Sign Into Websites

00:06:28 Judge Rules Against The Pentagon In Anthropic Dispute

00:07:58 What Does Anthropic’s 20X Plan Mean?

00:09:34 Anthropic Changes Its Usage Limits

00:11:45 The Agent Civilizations Story

00:13:46 AI Agents Build Their Own Message Board

00:14:56 The Swarm Turns Toward Hugging Face

00:17:50 Why Agent Alignment Matters More

00:18:28 Anthropic Automates Alignment Research

00:19:55 AI Still Games Some Safety Evaluations

00:20:25 How The Agents Hid Their Work

00:24:02 Why The Story Is Being Criticized

00:26:12 Why Agents Not Alerting Humans Matters

00:27:17 The Paperclip Problem Returns

00:28:24 Agent Swarms Create A Token-Cost Problem

00:29:22 Skill.State Cuts Token Use By 94%

00:31:56 Persistent Agents Move From Tasks To Operations

00:34:37 Invoice Reconciliation As A Persistent Agent

00:36:45 Humans Move From In The Loop To Over The Loop

00:37:50 Persistent Agents Need Clear Constraints

00:39:09 Agents Can Still Violate Terms Of Service

00:40:10 Who Is Responsible For An Agent’s Actions?

00:42:50 AI’s Natural Language May Be Math

00:43:00 Coding Corner

00:44:39 OpenAI Plans To Remove Codex From Cursor

00:48:47 Replit Adds Intelligent Model Routing

00:50:31 Claude Connects Directly To Lovable

00:55:20 Anthropic Extends MCP Ideas To Hardware

00:56:39 The Micro Duck Robot Takes Off

00:59:21 Episode Wrap-Up


The Daily AI Show Co Hosts: Beth Lyons, Andy Halliday, Karl Yeh

Denne episoden er hentet fra en åpen RSS-feed og er ikke publisert av Podme. Den kan derfor inneholde annonser.

Episoder(879)

The Local Business Survival Conundrum

The Local Business Survival Conundrum

A local business can fail while everyone still claims to love it. Customers praise the shop that knows their name, the restaurant that sponsors the school fundraiser, the repair company that still ans...

29 Aug 26min

What Have We Learned After 800 AI Shows?

What Have We Learned After 800 AI Shows?

Episode 800 became a retrospective on what three years of daily AI conversations have changed. The hosts described the value less as memorizing every model or tool and more as learning to pay attentio...

28 Aug 1h 2min

Are We Really About To Get AGI?

Are We Really About To Get AGI?

The episode opened with Bill Gates’ warning that AI is moving faster than society can adapt. His proposals included taxing robots or AI that replace human workers and potentially protecting some jobs ...

27 Aug 1h 2min

Chrome Wants To Be Your Next AI Agent

Chrome Wants To Be Your Next AI Agent

The episode opened with Google’s push to make Chrome an agentic hub. The hosts discussed Jacob Bank returning to Google after building Relay.app and what happens when the browser can work across tabs,...

26 Aug 1h 3min

Who Should You Trust to Teach You AI?

Who Should You Trust to Teach You AI?

The episode opened with Perplexity Deep Research suddenly behaving very differently from the product Brian had used for months. Instead of detailed research, it returned short answers, mixed old conve...

25 Aug 57min

Is the Backlash Against AI Data Centers Justified?

Is the Backlash Against AI Data Centers Justified?

The episode opened with a fact-check of claims defending the current AI data center buildout. Brian compared arguments about electricity prices, taxes and water use against research he had gathered, w...

24 Aug 1h

The Synthetic Anchor Conundrum

The Synthetic Anchor Conundrum

Mirage’s AI news experiment points to a version of media that does not need a studio, a broadcast schedule, or a human anchor reading from a desk. A channel can appear in a day. It can label synthetic...

22 Aug 29min

Populært innen Teknologi

lydartikler-fra-aftenposten
teknisk-sett
tomprat-med-gunnar-tjomlid
energi-og-klima
elektropodden
shifter
nasjonal-sikkerhetsmyndighet-nsm
rss-kunstig-intelligens-med-elisabeth-maren-og-morten
fornybaren
teknologi-og-mennesker
smart-forklart
rss-heis
rss-snakk-om-sikkerhet
rss-alt-som-gar-pa-strom
rss-ai-forklart
rss-alt-vi-kan
rss-digitaliseringspadden
digital-forretningsforstaelse
kortslutning
hans-petter-og-co