A Definition of AGI
GenAI Level UP23 Okt 2025

A Definition of AGI

For decades, Artificial General Intelligence has been a moving target, a nebulous concept that shifts every time a new AI masters a complex task. This ambiguity fuels unproductive debates and obscures the real gap between today's specialized models and true human-level cognition.

This episode changes everything.

We unpack a groundbreaking, quantifiable framework that finally stops the goalposts from moving. Grounded in the most empirically validated model of human intelligence (CHC theory), this approach introduces a standardized "AGI Score"—a single number from 0 to 100% that measures an AI against the cognitive versatility of a well-educated adult.

The scores are in, and they are astonishing. While GPT-4 scores 27%, the next generation leaps to 58%, revealing dizzying progress. But the total score isn't the real story. The true revelation is the "jagged profile" of AI's capabilities—a shocking disparity between superhuman brilliance and profound cognitive deficits.

This is your guide to understanding the true state of AI, moving beyond the hype to see the critical bottlenecks and the real path forward.

In this episode, you will discover:

    • (00:59) The AGI Scorecard: How a new framework, based on 10 core cognitive domains, provides a concrete, measurable definition of AGI for the first time.

    • (02:56) The Shocking Results: Unpacking the AGI scores for GPT-4 (27%) and the next-gen GPT-5 (58%), revealing both massive leaps and a substantial remaining gap.

    • (08:37) The Jagged Frontier & The 0% Problem: The most critical insight—why today's AI scores perfectly in math and reading yet gets a 0% in Long-Term Memory Storage, the system's most significant bottleneck.

    • (13:12) "Capability Contortions": The non-obvious ways AI masks its fundamental flaws, using enormous context windows and RAG to create a brittle illusion of general intelligence.

    • (16:21) AGI vs. Replacement AI: The provocative final question—can an AI become economically disruptive long before it ever achieves a perfect 100% AGI score?

Denne episoden er hentet fra en åpen RSS-feed og er ikke publisert av Podme. Den kan derfor inneholde annonser.

Episoder(45)

Recursive Self Improvement

Recursive Self Improvement

Imagine holding a wrench on an assembly line. Suddenly, it leaps from your hand, sprouts its own mechanical arms, and begins forging a faster, lighter wrench without you. You are no longer the creator...

7 Jun 1h

Master the New Physics of AI with Context Graphs & GraphRAG

Master the New Physics of AI with Context Graphs & GraphRAG

Stop trying to find the "magic words" to hack your LLM. The era of the Prompt Engineer—tweaking adjectives and hoping for the best—is officially over. We are entering the age of the Context Engineer, ...

1 Feb 17min

Context Graph

Context Graph

Stop feeding your AI static facts in a dynamic world.Most RAG systems and Knowledge Graphs rely on a fundamental unit called the "Triple" (Subject, Verb, Object). It’s efficient, but it’s brittle. It ...

25 Jan 19min

Nested Learning: The Illusion of Deep Learning Architectures

Nested Learning: The Illusion of Deep Learning Architectures

Why do today's most powerful Large Language Models feel... frozen in time? Despite their vast knowledge, they suffer from a fundamental flaw: a form of digital amnesia that prevents them from truly le...

14 Nov 202513min

Memento: Fine-tuning LLM Agents without Fine-tuning LLMs

Memento: Fine-tuning LLM Agents without Fine-tuning LLMs

What if you could build AI agents that get smarter with every task, learning from successes and failures in real-time—without the astronomical cost and complexity of constant fine-tuning? This isn't a...

1 Nov 202518min

MemGPT: Towards LLMs as Operating Systems

MemGPT: Towards LLMs as Operating Systems

Have you ever felt the frustration of an LLM losing the plot mid-conversation, its brilliant insights vanishing like a dream? This "goldfish memory"—the limited context window—is the Achilles' heel of...

1 Nov 202518min

DeepSeek-OCR: Contexts Optical Compression

DeepSeek-OCR: Contexts Optical Compression

The single biggest bottleneck for Large Language Models isn't intelligence—it's cost. The quadratic scaling of self-attention makes processing truly long documents prohibitively expensive, a fundament...

24 Okt 202513min

Populært innen Teknologi

lydartikler-fra-aftenposten
teknisk-sett
tomprat-med-gunnar-tjomlid
shifter
rss-ki-praten
elektropodden
rss-ai-forklart
fornybaren
nasjonal-sikkerhetsmyndighet-nsm
rss-bouvet-bobler
hans-petter-og-co
pedagogisk-intelligens
rss-polypod
rss-alt-som-gar-pa-strom
rss-fish-ships
smart-forklart
rss-kunstig-intelligens-med-elisabeth-maren-og-morten
rss-fisketimen
rss-nkom-innsikt
energi-og-klima