Claude Opus 4.7 Dropped — And a Local Model Drew the Better Pelican

Claude Opus 4.7 Dropped — And a Local Model Drew the Better Pelican

Claude Opus 4.7 is here with upgraded vision, memory, and instruction-following — but Simon Willison's pelican benchmark just handed the win to a local Alibaba model running on a laptop. We dig into what that actually means, plus Anthropic's new identity verification layer, Amazon's MCP bet, and whether "personal software" is about to change who gets to be a developer. Your commute just got more interesting.

Denne episoden er hentet fra en åpen RSS-feed og er ikke publisert av Podme. Den kan derfor inneholde annonser.

Episoder(24)

Your OpenClaw Updates Need a Canary, Not Courage

Your OpenClaw Updates Need a Canary, Not Courage

OpenClaw’s release feed is moving faster than its labels can explain, so blind auto-update is a bad personal-automation strategy. Cleo and Dev build a Release Sentinel canary, keep telemetry local, an...

2 Aug 20min

Claude Code Changed Engines—Your Evals Just Broke

Claude Code Changed Engines—Your Evals Just Broke

Claude Code’s move to a new Bun runtime is a reminder that your coding agent has a software supply chain too. Alex and Sam unpack runtime drift, model routers, reverse-engineering with agents, and a f...

24 Jul 19min

Better Agent Tools Made Code Review Worse

Better Agent Tools Made Code Review Worse

GitHub gave its code-review agent better tools and watched cost rise while useful findings fell. Alex and Sam unpack why task-shaped instructions beat bigger toolboxes, how invisible environment detai...

14 Jul 18min

Your AI Coding Benchmarks Are Lying To You

Your AI Coding Benchmarks Are Lying To You

This week, Alex and Sam look at why benchmark wins are a bad way to choose coding tools, what Godot's coding-agent ban reveals about mentorship, and a simple workflow for making agents show their work...

3 Jul 18min

The Tiny Local Model That Changes Your Agent Budget

The Tiny Local Model That Changes Your Agent Budget

Small, local models are suddenly good enough for real agent chores, but the win is not replacing your smartest model. Cleo and Dev unpack lightweight extraction models, model-routing memory, browser-s...

26 Jun 18min

Your Coding Agent Needs a Bouncer Now

Your Coding Agent Needs a Bouncer Now

AI coding agents are getting longer runs, more context, and more ways to touch production workflows, but this week made the real bottleneck obvious: authorization. Alex and Sam unpack MCP's missing en...

19 Jun 19min

Verification Is Now Your Coding Agent Bottleneck

Verification Is Now Your Coding Agent Bottleneck

Coding agents are getting better at long runs, but this week's news points at the real limit: proof. Alex and Sam unpack agent loops, Stack Overflow for Agents, Copilot CLI delegation, local-model cod...

17 Jun 11min

Cursor's Tokenomics Reckoning Hits Every Coding Agent

Cursor's Tokenomics Reckoning Hits Every Coding Agent

Coding agents are no longer just a workflow story; they are a cost, context, and control story. Alex and Sam unpack Cursor's pricing reset, Uber capping Claude Code usage, GitHub's agent-native deskto...

5 Jun 16min

Populært innen Teknologi

lydartikler-fra-aftenposten
tomprat-med-gunnar-tjomlid
teknisk-sett
elektropodden
rss-ai-forklart
shifter
fornybaren
rss-alt-som-gar-pa-strom
hans-petter-og-co
rss-bak-skyen
rss-ki-praten
rss-digitaliseringspadden
energi-og-klima
smart-forklart
kortslutning
sosialt-sett-om-teknologi-kommunikasjon-og-livet-i-mellom
rss-bouvet-bobler
rss-ki-til-kaffen
rss-teknologioptimistene-energibransjens-it-podcast
rss-grenser-for-ki