We Measure What AI Can Do. We Should Measure What It Does to Us.

We Measure What AI Can Do. We Should Measure What It Does to Us.

In AI, what gets measured gets optimized. Right now, we're spending all our efforts to measure how capable and powerful models are, narrowly optimizing for those metrics while ignoring downstream consequences.

What if we could flip this dynamic on its head? What if, instead of what AI can do, we start to measure what AI does to us? What if, instead of races to the bottom on capabilities and engagement, we could incentivize races to the top on safety, or better yet, on making us more resilient and developed human beings?

That’s the mission of CHT’s Humane Evals program: we're bringing together researchers, psychologists, engineers, and technologists from across the entire AI ecosystem and beyond to build out the expertise and infrastructure we need to measure AI's impact on humans.

Today on the show, Aza Raskin explores the Humane Evals project with Imran Khan, a researcher and strategist who's been leading CHT's efforts in this area, and Jared Moore, a computer scientist and researcher who's been at the forefront of measuring AI's psychological impact on users.

If this sounds like something you're interested in working on, you can email us at evals@humanetech.com.

RECOMMENDED MEDIA

You can read more about Jared’s work at his website.

Related pieces by Imran on the CHT Substack:

The website for the UC Berkeley Center for the Science of Psychedelics

KoraBench, the child safety AI benchmark that Imran referenced

RECOMMENDED YUA EPISODES

The AI Dilemma

Attachment Hacking and the Rise of AI Psychosis

How OpenAI's ChatGPT Guided a Teen to His Death

Corrections:

Aza gave the wrong year for Sewell Setzer’s death. It was in 2024, not 2025.

Aza referred to Joseph Henrich as an evolutionary psychologist; he was actually a biological anthropologist.


Hosted by Simplecast, an AdsWizz company. See pcm.adswizz.com for information about our collection and use of personal data for advertising.

Denne episoden er hentet fra en åpen RSS-feed og er ikke publisert av Podme. Den kan derfor inneholde annonser.

Episoder(167)

Flock is Just the Beginning: Inside the Era of AI-Powered Policing

Flock is Just the Beginning: Inside the Era of AI-Powered Policing

In a 2018 TED Talk, Yuval Harari argued that democracy prevailed over fascism and communism in the 20th century not because of any inherent advantage, but because the technology of that age favored op...

10 Sep 59min

Enough Debate about the AI Jobpocalypse. We Need To Plan for the Messy Middle.

Enough Debate about the AI Jobpocalypse. We Need To Plan for the Messy Middle.

It feels like we’re stuck in an endless debate about what AI is going to mean for jobs and the economy. The prediction you hear from the people closest to the technology — both its critics and its boo...

13 Aug 57min

Can AI Be Built in Service of Life? A Conversation with Krista Tippett

Can AI Be Built in Service of Life? A Conversation with Krista Tippett

This week, we're bringing you a conversation that Tristan Harris had with Krista Tippett. Krista is the Peabody Award-winning host of the On Being podcast, where she explores spiritual inquiry, scienc...

16 Jul 53min

“Magnifica Humanitas:” Pope Leo’s Clarion Call on AI

“Magnifica Humanitas:” Pope Leo’s Clarion Call on AI

Since stepping into the Papacy, Pope Leo XIV has been a forceful voice pushing back against the anti-human path we’re on with AI. In May, he released “Magnifica Humanitas,” a sprawling encyclical warn...

2 Jul 30min

We Need AI Treaties. This is How We Get Them

We Need AI Treaties. This is How We Get Them

In the middle of the twentieth century, the existential threat posed by nuclear weapons seemed inevitable. The number of countries with nukes was climbing rapidly, and the idea of stopping the nuclear...

18 Jun 51min

What Do We Mean by Humane Tech?

What Do We Mean by Humane Tech?

We often think of the challenges created by technology as separate and disconnected, so trying to solve them feels like playing the world's hardest game of Whac-A-Mole.  What if, instead, we tackled t...

4 Jun 52min

Anthropic’s Mythos Has Changed Cybersecurity Forever. What Now?

Anthropic’s Mythos Has Changed Cybersecurity Forever. What Now?

A generation ago, the world's critical infrastructure was physical. Today, it’s largely digital. Your bank vault is a database, your filing cabinet is a server, your car is a robot on wheels. And in a...

14 Mai 46min

Populært innen Samfunn

rss-spartsklubben
giver-og-gjengen-vg
aftenpodden
aftenpodden-usa
konspirasjonspodden
rss-nesten-hele-uka-med-lepperod
popradet
alt-fortalt
rss-henlagt-andy-larsgaard
min-barneoppdragelse
wolfgang-wee-uncut
rss-espen-lee-usensurert
grenselos
synnve-og-vanessa
rss-dette-ma-aldri-skje-igjen
fladseth
frokostshowet-pa-p5
lydartikler-fra-aftenposten
krisemoter
rss-siktet