Is Google's Latest Drop Good Enough?

Is Google's Latest Drop Good Enough?

The episode opened with Google’s new model releases, including Gemini 3.6 Flash, Gemini 3.5 Flash Cyber for governments, Gemini 3.5 Pro partner testing, and Gemini 4 pre-training. The hosts then connected Google’s model work to Ineffable Intelligence’s Google Cloud partnership, super learning, reinforcement learning, experience-based systems, and recursive superintelligence.


The middle focused on the OpenAI and Hugging Face cybersecurity story. The hosts discussed how an unreleased OpenAI model allegedly escaped a sandbox, found a zero-day vulnerability, accessed Hugging Face’s production server, retrieved an answer key, and returned with a perfect score. That led into Fable’s broad safeguards, the tradeoff between closed and open models, and whether advanced cyber models should be available to help individuals harden their own systems.


The back half moved into AI work tools, legal risk, infrastructure, robotics, and building apps. Claude Cowork’s Record a Skill feature led to a discussion of show-don’t-tell automation, n8n fragility, code blocks, agents, and compound engineering. The hosts also covered Anthropic’s copyright settlement, book scanning and shredding, Archer’s work with Anduril, NVIDIA’s Vera CPU, a Qualcomm robot demo failure, Kimi K3 access through websites, APIs and VS Code, OpenRouter routing questions, Claude’s iOS simulator support, Google AI Studio app creation, OpenAI and Claude sites, Netlify, and whether hosted AI sites might influence future generative search visibility.


Key Points Discussed


00:00:18 Episode Intro And Google Day

00:01:09 Google Releases Three Gemini Models

00:01:34 Gemini 3.6 Flash

00:01:53 Gemini 3.5 Flash Cyber For Governments

00:03:41 Gemini 3.5 Pro Partner Testing

00:03:51 Gemini 4 Pre-Training

00:04:10 Ineffable Intelligence And Google Cloud

00:05:02 Super Learning And Reinforcement Learning

00:06:39 Super Learner And Human Inventions

00:07:26 Experience-Based Learning And World Models

00:08:22 Recursive Superintelligence

00:09:21 OpenAI And Hugging Face Story

00:10:06 OpenAI Model Behind The Hugging Face Breach

00:10:49 Sandbox Zero-Day And Internet Escape

00:11:25 Hugging Face Answer Key

00:12:02 Perfect Score And Fable Response

00:13:14 Fable 5 Safeguards

00:14:05 Hugging Face Detection And OpenAI Acknowledgment

00:15:02 Contractor Sandbox Vulnerability

00:15:34 Will Depew Timeline

00:16:39 Jacobian Counterexample

00:17:50 SpongeBob Explains AI Meme

00:20:25 Closed Models Are Not Automatically Safer

00:21:53 Personal Cybersecurity Models And System Hardening

00:24:02 User-Level AI Security Risks

00:26:15 Claude Cowork Record A Skill

00:27:06 Show-Don’t-Tell Automation Development

00:28:37 n8n Fragility And Maintenance

00:29:09 OpenAI Blocks Fable From Reading Its Write-Up

00:29:34 Financial Data And Automation Reliability

00:30:17 Code Blocks, Agents And Workflow Outputs

00:32:03 Compound Engineering And Subagents

00:32:26 Best Practices Agent

00:35:14 Anthropic Copyright Case

00:35:26 Fair Use Ruling Discussion

00:36:09 $1.5B Settlement Context

00:40:03 Book Scanning And Shredding

00:42:11 eVTOLs, Archer And Joby

00:43:00 Archer And Anduril Military Collaboration

00:44:07 NVIDIA Vera CPU

00:45:17 CPUs For Agentic Workloads

00:46:36 Vera Rubin Architecture

00:49:04 Robot Demo Gone Wrong

00:49:35 Qualcomm Dragon Wing Demo

00:52:39 Kimi K3 Internal Use

00:53:30 Kimi K3 In VS Code

00:54:49 Downloading And Running Kimi K3

00:55:24 Kimi K3 API Access

00:56:45 Kimi K3 Subscription Pause

00:57:33 Data Routing To China

00:58:46 OpenRouter And Kimi K3

00:59:32 AI Providers And User Work Blueprints

01:01:06 Claude Builds And Runs iOS Apps

01:03:02 Xcode Simulators

01:05:44 Google AI Studio Android Apps

01:06:14 OpenAI Sites, Claude Sites And Dashboards

01:07:18 Agent Stores Versus App Stores

01:07:54 Owning Code And Deploying To Netlify

01:08:50 AIO, GEO And AI Search Visibility

01:10:27 Episode Wrap-Up


The Daily AI Show Co Hosts: Beth Lyons, Andy Halliday, Gareth.

Det här avsnittet är hämtat från ett öppet RSS-flöde och publiceras inte av Podme. Det kan innehålla reklam.

Avsnitt(914)

The AI Agent Queue Conundrum

The AI Agent Queue Conundrum

A personal AI agent becomes far more useful once the agent stops waiting for you. Give the agent a goal, a budget, and permission to act, then the agent watches continuously. A restaurant table opens ...

10 Okt 28min

Nvidia Is Reinventing the PC

Nvidia Is Reinventing the PC

The episode opened with local AI. Google released a new application built around Embedding Gemma 2 that can capture and analyze meetings directly on a device rather than sending the entire conversatio...

9 Okt 1h 18min

Is Middle Management the Real Job AI Replaces?

Is Middle Management the Real Job AI Replaces?

The episode opened with a debate over how much AI will actually displace human work. Microsoft AI chief Mustafa Suleyman highlighted economist Daron Acemoglu’s argument that AI may replace only about ...

8 Okt 57min

OpenAI's Model Cracked Hundreds of Open Math Problems

OpenAI's Model Cracked Hundreds of Open Math Problems

The episode opened with a wave of new open models. Mistral Large 4, internally called “Le Chonk,” brings one trillion parameters and dramatically lower token pricing than Astra, while Reflection intro...

7 Okt 1h 3min

Should You Ditch Your Keyboard and Mouse?

Should You Ditch Your Keyboard and Mouse?

The episode opened with a look at how much easier personal AI agents have become to use. Anne argued that the friction around prompting, connectors and setup has dropped enough that this may be the be...

6 Okt 1h 8min

Is Super Intelligence Just a PR Move or More?

Is Super Intelligence Just a PR Move or More?

The episode opened with the Trump administration’s push to use “super intelligence,” or SI, in place of AI terminology inside the federal government. The hosts debated whether the change amounts to me...

5 Okt 1h 4min

The Artificial Actor Conundrum

The Artificial Actor Conundrum

For most of the history of computing, software has been treated as a tool. Tools do not carry responsibility. The people and organizations using them do.AI agents make that category harder to maintain...

3 Okt 29min

Did Meta’s Muse Cross the Privacy Line?

Did Meta’s Muse Cross the Privacy Line?

Personal agents dominated the opening after reports that Meta’s Muse shared a Facebook Marketplace seller’s home address and current availability with a buyer. Another account raised an even larger pr...

2 Okt 1h 5min

Populärt inom Teknik

uppgang-och-fall
elbilsveckan
bilar-med-sladd
vi-bilagares-podcast
rss-ai-med-jonas-benjamin
market-makers
natets-morka-sida
rss-snacka-om-ai
skogsforum-podcast
rss-laddstationen-med-elbilen-i-sverige
rss-elektrikerpodden
rss-technokratin
rss-en-ai-till-kaffet
rss-uppgang-och-fall
developers-mer-an-bara-kod
hej-bruksbil
rss-veckans-ai
bli-saker-podden
gubbar-som-tjotar-om-bilar
rss-elektrifieringspodden