AI's Unsung Hero: Data Labeling and Expert Evals
AI + a16z27 Juni 2025

AI's Unsung Hero: Data Labeling and Expert Evals

Labelbox CEO Manu Sharma joins a16z Infra partner Matt Bornstein to explore the evolution of data labeling and evaluation in AI — from early supervised learning to today’s sophisticated reinforcement learning loops.

Manu recounts Labelbox’s origins in computer vision, and then how the shift to foundation models and generative AI changed the game. The value moved from pre-training to post-training and, today, models are trained not just to answer questions, but to assess the quality of their own responses. Labelbox has responded by building a global network of “aligners” — top professionals from fields like coding, healthcare, and customer service, who label and evaluate data used to fine-tune AI systems.

The conversation also touches on Meta’s acquisition of Scale AI, underscoring how critical data and talent have become in the AGI race.

Here's a sample of Manu explaining how Labelbox was able to transition from one era of AI to another:

It took us some time to really understand like that the world is shifting from building AI models to renting AI intelligence. A vast number of enterprises around the world are no longer building their own models; they're actually renting base intelligence and adding on top of it to make that work for their company. And that was a very big shift.

But then the even bigger opportunity was the hyperscalers and the AI labs that are spending billions of dollars of capital developing these models and data sets. We really ought to go and figure out and innovate for them. For us, it was a big shift from the DNA perspective because Labelbox was built with a hardcore software-tools mindset. Our go-to market, engineering, and product and design teams operated like software companies.

But I think the hardest part for many of us, at that time, was to just make the decision that we're going just go try it and do it. And nothing is better than that: "Let's just go build an MVP and see what happens."

Follow everyone on X:

Manu Sharma

Matt Bornstein

Check out everything a16z is doing with artificial intelligence here, including articles, projects, and more podcasts.

Please note that the content here is for informational purposes only; should NOT be taken as legal, business, tax, or investment advice or be used to evaluate any investment or security; and is not directed at any investors or potential investors in any a16z fund. a16z and its affiliates may maintain investments in the companies discussed. For more details please see a16z.com/disclosures.


Hosted by Simplecast, an AdsWizz company. See pcm.adswizz.com for information about our collection and use of personal data for advertising.

Avsnitt(81)

Why This Isn't the Dot-Com Bubble | Martin Casado on WSJ's BOLD NAMES

Why This Isn't the Dot-Com Bubble | Martin Casado on WSJ's BOLD NAMES

Christopher Mims and Tim Higgins of the Wall Street Journal sit down with a16z General Partner Martin Casado on WSJ’s Bold Names to ask whether the AI spending boom is a bubble waiting to burst. Marti...

3 Feb 29min

Martin Casado on the Demand Forces Behind AI

Martin Casado on the Demand Forces Behind AI

In this feed drop from The Six Five Pod, a16z General Partner Martin Casado discusses how AI is changing infrastructure, software, and enterprise purchasing. He explains why current constraints are dr...

27 Jan 27min

How Mintlify Is Rebuilding Documentation for Coding Agents

How Mintlify Is Rebuilding Documentation for Coding Agents

Mintlify is a documentation platform built by cofounders Han Wang and Hahnbee Lee to help teams create and maintain developer docs. In this episode, Andreessen Horowitz general partners Jennifer Li an...

23 Jan 44min

Inferact: Building the Infrastructure That Runs Modern AI

Inferact: Building the Infrastructure That Runs Modern AI

Inferact is a new AI infrastructure company founded by the creators and core maintainers of vLLM. Its mission is to build a universal, open-source inference layer that makes large AI models faster, ch...

22 Jan 43min

How Should AI Be Regulated? Use vs. Development

How Should AI Be Regulated? Use vs. Development

To Regulate AI Effectively, Focus on How It’s UsedA conversation with Martin Casado on learning from past computing platform shifts, understanding marginal risk in AI, and why open source matters for ...

20 Jan 46min

Michael Truell: How Cursor Builds at the Speed of AI

Michael Truell: How Cursor Builds at the Speed of AI

When four MIT grads decided to build a code editor while everyone else was building AI agents, they created the fastest-growing developer tool ever built. Cursor CEO Michael Truell joins a16z’s Martin...

13 Jan 27min

Dylan Patel on the AI Chip Race - NVIDIA, Intel & the US Government

Dylan Patel on the AI Chip Race - NVIDIA, Intel & the US Government

Nvidia’s $5 billion investment in Intel is one of the biggest surprises in semiconductors in years. Two longtime rivals are now teaming up, and the ripple effects could reshape AI, cloud, and the glob...

6 Jan 1h 40min

Feed Drop from The Generalist: Why a16z's Martin Casado believes the AI boom still has years to run

Feed Drop from The Generalist: Why a16z's Martin Casado believes the AI boom still has years to run

This episode is a special replay from The Generalist Podcast, featuring a conversation with a16z General Partner Martin Casado. Martin has lived through multiple tech waves as a founder, researcher, a...

30 Dec 20251h 21min

Populärt inom Business & ekonomi

badfluence
framgangspodden
rss-jossan-nina
varvet
rss-borsens-finest
uppgang-och-fall
avanzapodden
svd-tech-brief
fill-or-kill
bathina-en-podcast
lastbilspodden
borsmorgon
rss-inga-dumma-fragor-om-pengar
rss-kort-lang-analyspodden-fran-di
kapitalet-en-podd-om-ekonomi
rss-dagen-med-di
rss-den-nya-ekonomin
affarsvarlden
rss-borslunch
rikatillsammans-om-privatekonomi-rikedom-i-livet