CoT Evolved 3 New Chains for the Reasoning AI Era (Ep. 460)

CoT Evolved 3 New Chains for the Reasoning AI Era (Ep. 460)

Want to keep the conversation going?

Join our Slack community at thedailyaishowcommunity.com


What started as a simple “let’s think step by step” trick has grown into a rich landscape of reasoning models that simulate logic, branch and revise in real time, and now even collaborate with the user. The episode explores three specific advancements: speculative chain of thought, collaborative chain of thought, and retrieval-augmented chain of thought (CoT-RAG).


Key Points Discussed

Chain of thought prompting began in 2022 as a method for improving reasoning by asking models to slow down and show their steps.


By 2023, tree-of-thought prompting and more branching logic began emerging.


In 2024, tools like DeepSeek and O3 showed dynamic reasoning with visible steps, sparking renewed interest in more transparent models.


Andy explains that while chain of thought looks like sequential reasoning, it’s really token-by-token prediction with each output influencing the next.


The illusion of “thinking” is shaped by the model’s training on step-by-step human logic and clever UI elements like “thinking…” animations.


Speculative chain of thought uses a smaller model to generate multiple candidate reasoning paths, which a larger model then evaluates and improves.


Collaborative chain of thought lets the user review and guide reasoning steps as they unfold, encouraging transparency and human oversight.


Chain of Thought RAG combines structured reasoning with retrieval, using pseudocode-like planning and knowledge graphs to boost accuracy.


Jyunmi highlighted how collaborative CoT mirrors his ideal creative workflow by giving humans checkpoints to guide AI thinking.


Beth noted that these patterns often mirror familiar software roles, like sous chef and head chef, or project management tools like Gantt charts.


The team discussed limits to context windows, attention, and how reasoning starts to break down with large inputs or long tasks.


Several ideas were pitched for improving memory, including token overlays, modular context management, and step weighting.


The conversation wrapped with a reflection on how each CoT model addresses different needs: speed, accuracy, or collaboration.


Timestamps & Topics

00:00:00 🧠 What is Chain of Thought evolved?


00:02:49 📜 Timeline of CoT progress (2022 to 2025)


00:04:57 🔄 How models simulate reasoning


00:09:36 🤖 Agents vs LLMs in CoT


00:14:28 📚 Research behind the three CoT variants


00:23:18 ✍️ Overview of Speculative, Collaborative, and RAG CoT


00:25:02 🧑‍🤝‍🧑 Why collaborative CoT fits real-world workflows


00:29:23 📌 Brian highlights human-in-the-loop value


00:32:20 ⚙️ CoT-RAG and pseudo-code style logic


00:34:35 📋 Pretraining and structured self-ask methods


00:41:11 🧵 Importance of short-term memory and chat history


00:46:32 🗃️ Ideas for modular memory and reg-based workflows


00:50:17 🧩 Visualizing reasoning: Gantt charts and context overlays


00:52:32 ⏱️ Tradeoffs: speed vs accuracy vs transparency


00:54:22 📬 Wrap-up and show announcements


Hashtags

#ChainOfThought #ReasoningAI #AIprompting #DailyAIShow #SpeculativeAI #CollaborativeAI #RetrievalAugmentedGeneration #LLMs #AIthinking #FutureOfAI


The Daily AI Show Co-Hosts: Andy Halliday, Beth Lyons, Brian Maucere, Eran Malloch, Jyunmi Hatcher, and Karl Yeh

Denne episoden er hentet fra en åpen RSS-feed og er ikke publisert av Podme. Den kan derfor inneholde annonser.

Episoder(888)

The Economics of Work In An Age of AI

The Economics of Work In An Age of AI

The episode centered on what happens to the economics of work as AI becomes capable of doing more of it. Anthropic’s new Economic Scenarios Explorer provided the starting point, allowing users to mode...

10 Sep 1h 6min

10,000 AI Agents Attack One Problem

10,000 AI Agents Attack One Problem

The episode opened with the dispute surrounding OpenAI’s newly announced mathematical result and what may be the more important story behind it. Tristan Buckmaster of NYU and Anthropic researcher Leve...

9 Sep 1h 3min

Our Real Atlas Builds and Use Cases

Our Real Atlas Builds and Use Cases

The episode moved quickly from theory to practical experience with GPT-6 Astra. After revisiting OpenAI’s “Alien Mind” paper and the conundrum of using more powerful AI to monitor frontier systems, th...

8 Sep 1h 5min

Can We Truly Control The Alien Mind?

Can We Truly Control The Alien Mind?

The episode focused heavily on GPT-6 Astra and a new essay from OpenAI chief scientist Jakub Pachocki describing advanced AI systems as increasingly alien forms of intelligence that humans grow throug...

7 Sep 54min

The Democratic Bandwidth Conundrum

The Democratic Bandwidth Conundrum

Public participation has always contained a hidden constraint: time.Writing a serious response to a tax rule, zoning plan, environmental permit, school policy, or agency proposal takes hours. Filing r...

5 Sep 28min

Is GPT-6 Astra the Biggest AI Leap Yet?

Is GPT-6 Astra the Biggest AI Leap Yet?

OpenAI’s GPT-6 Astra dominated the episode after its unusual rollout. The hosts discussed access, OpenAI’s plan to bring Astra to paid users, and why some cybersecurity users may receive capabilities ...

4 Sep 1h 1min

Will Stores Use AI to Charge You More?

Will Stores Use AI to Charge You More?

The episode opened with the downside of increasingly capable AI harnesses. OpenClaw 2.0 made setup easier, but some self-hosted users reported broken gateways, failed migrations and unusable systems a...

3 Sep 1h 3min

Is Fable 5.1 Good Enough to Make You Leave Codex?

Is Fable 5.1 Good Enough to Make You Leave Codex?

Anthropic’s Fable 5.1 dominated the first half of the episode. Beth and Andy compared its higher output costs with improved caching, stronger benchmark performance and better agentic task results. The...

2 Sep 58min

Populært innen Teknologi

lydartikler-fra-aftenposten
tomprat-med-gunnar-tjomlid
teknisk-sett
energi-og-klima
rss-ai-forklart
nasjonal-sikkerhetsmyndighet-nsm
shifter
rss-heis
rss-kunstig-intelligens-med-elisabeth-maren-og-morten
elektropodden
rss-alt-som-gar-pa-strom
rss-bouvet-bobler
fornybaren
teknologi-og-mennesker
rss-alt-vi-kan
rss-teknologioptimistene-en-podkast-om-teknologi-og-mennesker
rss-bak-skyen
i-loopen
rss-larervarelset
rss-forvarelset