#325 Phelim Brady: Why AI's Future Depends on Human Judgement

#325 Phelim Brady: Why AI's Future Depends on Human Judgement

AI often looks fully automated. But behind the scenes, a huge amount of human judgment is shaping how these systems actually work.

In this episode, Craig Smith speaks with Phelim Bradley, co-founder and CEO of Prolific, a platform that connects millions of real people with researchers and AI labs to evaluate and improve AI systems.

They explore the hidden human layer behind modern AI, why traditional benchmarks are becoming less reliable, and why AI companies increasingly rely on real human feedback to measure model performance in the real world.

Phelim also explains how demographic differences influence how models are evaluated, why human judgment remains critical even as AI improves, and how the collaboration between humans and AI will shape the next phase of development.

This conversation reveals the human backbone behind today's AI systems.


Stay Updated:

Craig Smith on X: https://x.com/craigss

Eye on A.I. on X: https://x.com/EyeOn_AI

(00:00) Preview and Intro

(02:45) Founding Prolific And Early Pain Points

(06:30) From Mechanical Turk To Representativeness

(09:55) Academic Research And AI Use Cases Split

(13:40) Vetting Real Participants And Fighting Fraud

(17:45) Scale, Community Growth, And Talent Mix

(22:00) High-Complexity Projects Over Commoditised Labeling

(26:40) Measuring Model Persuasion With Live Conversations

(30:20) Demographic-Aware Model Preference Benchmarks

(34:10) The Rise Of Human Evaluation Over Benchmarks

(38:00) Enterprise Model Choice And Continuous Evaluation

(42:00) Why Humans Won't Disappear From The Loop





Denne episoden er hentet fra en åpen RSS-feed og er ikke publisert av Podme. Den kan derfor inneholde annonser.

Episoder(386)

Chat GPT's Creator Explains The Shocking Truth About How AI Understands the World | Ilya Sutskever

Chat GPT's Creator Explains The Shocking Truth About How AI Understands the World | Ilya Sutskever

Craig Smith sits down with Ilya Sutskever - Co-Founder and Chief Scientist at Safe Superintelligence Inc. & former chief scientist at OpenAI and one of the primary minds behind GPT-3, GPT-4, and the d...

9 Okt 41min

The Coordination Tax: Why AI Is Burning Out Your Best People | Dan O'Connell, Front

The Coordination Tax: Why AI Is Burning Out Your Best People | Dan O'Connell, Front

Everyone assumes AI in customer operations means fewer people and lower costs. The data from 700 customer operations leaders is more complicated, and more honest. Dan O'Connell, CEO of Front, joins Cr...

6 Okt 51min

Ukraine's Secret Weapon: The Points System Winning the Drone War | Andrii Hrytseniuk

Ukraine's Secret Weapon: The Points System Winning the Drone War | Andrii Hrytseniuk

Ukraine tracks every confirmed drone kill with video evidence, deduplicates the data to prevent double-counting, converts verified strikes into e-points, and delivers newly ordered weapons to frontlin...

2 Okt 33min

Why Current AI Cannot Be Conscious | Dr. Christof Koch

Why Current AI Cannot Be Conscious | Dr. Christof Koch

One in four patients currently being considered for life support withdrawal may actually be fully conscious, they just can't signal it. That single finding, from a landmark New England Journal of Medi...

30 Sep 1h 1min

The Technology for Fully Autonomous Attack Is Already Here | Alex Liannyi, NORDA Dynamics

The Technology for Fully Autonomous Attack Is Already Here | Alex Liannyi, NORDA Dynamics

The AI systems guiding Ukrainian combat drones aren't running on expensive Nvidia chips. They're running on a Raspberry Pi Zero - a $15 hobby computer - and that single detail tells you more about how...

24 Sep 19min

Inside Ukraine's Drone War: Maj. "Phoenix" of Lasar's Group

Inside Ukraine's Drone War: Maj. "Phoenix" of Lasar's Group

The first armed drone Ukraine ever fielded wasn't built in a factory or procured from a defense contractor. It was built in four months by a network engineer using a Starlink terminal and a large agri...

21 Sep 1h 5min

The Hidden Algorithm That Decides Which Software AI Will Recommend | Tim Sanders, G2

The Hidden Algorithm That Decides Which Software AI Will Recommend | Tim Sanders, G2

Most companies investing in AI visibility are optimizing for the wrong thing. Being cited by an AI response and being recommended by an AI response are completely different outcomes, with click-throug...

14 Sep 58min

The Reason 30 Years of Cybersecurity Has Failed - and What Actually Fixes It | Trent Telford, Qanapi

The Reason 30 Years of Cybersecurity Has Failed - and What Actually Fixes It | Trent Telford, Qanapi

Every major data breach in the last 30 years shares the same root cause: the data inside the wall was never protected, only the wall. And AI frontier models are now making that wall easier to breach t...

10 Sep 55min

Populært innen Teknologi

energi-og-klima
teknisk-sett
lydartikler-fra-aftenposten
nasjonal-sikkerhetsmyndighet-nsm
smart-forklart
elektropodden
shifter
rss-ai-forklart
tomprat-med-gunnar-tjomlid
rss-alt-vi-kan
rss-bouvet-bobler
kortslutning
fornybaren
rss-bak-skyen
rss-teknologioptimistene-en-podkast-om-teknologi-og-mennesker
pedagogisk-intelligens
rss-larervarelset
rss-kunstig-intelligens-med-elisabeth-maren-og-morten
rss-ki-praten
rss-digitaliseringspadden