Beyond Transcripts:  Language Nuances and Audio Signals with Carter Huffman of Modulate

Beyond Transcripts: Language Nuances and Audio Signals with Carter Huffman of Modulate

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub talk with Carter Huffman, CTO and co-founder of Modulate AI, about his path from machine learning work at NASA's Jet Propulsion Lab to building voice AI that understands conversations. Carter explains why moderation in gaming is hard because you don't want to ban players unfairly, and contrasts big foundation models with orchestrated ensembles of many tiny models that require high-quality, globally vetted data labeling. They discuss the nuance of classifying hate speech, expansion into detecting fraud and manipulation in delivery and call-center contexts, and monitoring misbehaving AI voice agents. The conversation covers why conversation is more than transcripts, possible therapeutic/telehealth uses of Modulate, analyzing data at a massive scale, and ambitions for audio generation using hierarchical edge-and-cloud approaches. The episode ends with a humorous anecdote about two factor authenticaiton failure. 00:00 Podcast Cold Open 00:48 Meet Carter Huffman 02:06 JPL Spacecraft Autonomy 04:18 From JPL to Audio AI 06:18 Why Audio Is Hard 07:44 Voice AI Use Cases 12:49 Tiny Models Orchestration 15:56 Data Labeling at Scale 17:17 Defining Toxic Behavior 18:58 Nuanced Language Moderation 20:04 Scaling Ensemble Models 21:39 GPU Crunch During Launch 22:29 Beyond Gaming Use Cases 26:03 AI Agents Gone Wrong 28:45 Telehealth and Diagnostics 30:26 Ambient Audio and Privacy 32:26 Edge Ensembles Everywhere 33:25 Audio Synthesis Ambitions 35:24 Latency Hierarchies Explained 38:10 Two Factor Key Fob Fiasco 39:14 Wrap Up and Credits

Resources:

#TechPodcast #EngineeringPodcast #DevTalks #PodcastForDevs #HowManyCTOs #Podcast #CTOs #CTOPodcast #ChiefTechnologyOfficer #Technology #Engineering #SoftwareDevelopment #SoftwareEngineering #TechLeadership #EngineeringLeadership #EngineeringCulture #TechDebates #AI #VoiceTech #MachineLearning #MachineLearningModels #GamingIndustry #AIinnovation #Entrepreneurship #AIConversation #VoiceAssistant #LanguageModeration #GPU #LLMs #LargeLanguageModels

Denne episoden er hentet fra en åpen RSS-feed og er ikke publisert av Podme. Den kan derfor inneholde annonser.

Episoder(85)

The Hardware/Software Blend: Visionary Innovations in Health Tech with Eyebot's Jack Moldave

The Hardware/Software Blend: Visionary Innovations in Health Tech with Eyebot's Jack Moldave

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub interview Jack Moldave of Eyebot, a company he co-founded nearly six years ago to make getting glasses a...

1 Sep 52min

Fast, On Time, or Fully Utilized? Choosing What to Optimize in Engineering Delivery

Fast, On Time, or Fully Utilized? Choosing What to Optimize in Engineering Delivery

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub discuss a leadership conversation where three senior stakeholders wanted different outcomes from enginee...

25 Aug 16min

The Pace of Innovation: Systems Thinking for Faster Engineering

The Pace of Innovation: Systems Thinking for Faster Engineering

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub open with travel and sports talk, including World Cup rules, then pivot to Scott's change-management tra...

18 Aug 56min

Method to the Madness: A CTO Framework for Systemic AI Adoption

Method to the Madness: A CTO Framework for Systemic AI Adoption

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub discuss how CTOs must frame narratives that help teams organize around ideas, inspired by how Scott expl...

11 Aug 45min

There's A Pattern To Follow: An Interview with Robert "Uncle Bob" Martin

There's A Pattern To Follow: An Interview with Robert "Uncle Bob" Martin

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub are honored to interview Robert "Uncle Bob" Martin, who recounts his 60+ year programming journey from a...

4 Aug 1h 1min

Is Uncle Bob Right? Reading Code vs. Architecting the Future

Is Uncle Bob Right? Reading Code vs. Architecting the Future

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub discuss how Robert "Uncle Bob" Martin—author of Clean Code and Agile Manifesto signer—sparked online con...

28 Jul 43min

Token Maxxing and Local Inference: Navigating the Shift in AI Usage for Developers

Token Maxxing and Local Inference: Navigating the Shift in AI Usage for Developers

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub discuss token spending and the "token maxing" backlash, sparked by a GitHub Copilot budget issue and bro...

21 Jul 1h

This Has Always Been A Problem: Slot Machine Devs and Growing Junior Engineers in the Age of AI

This Has Always Been A Problem: Slot Machine Devs and Growing Junior Engineers in the Age of AI

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub discuss how managing armies of coding agents can be physically draining, likening it to context switchin...

14 Jul 56min

Populært innen Business og økonomi

stopp-verden
dine-penger-pengeradet
e24-podden
rss-penger-polser-og-politikk
lydartikler-fra-aftenposten
rss-borsmorgen-okonominyhetene
rss-skravla-gar
finansredaksjonen
rss-pa-konto
pengepodden-2
tid-er-penger-en-podcast-med-peter-warren
pengesnakk
lederpodden
livet-pa-veien-med-jan-erik-larssen
rss-orjasater
stormkast-med-valebrokk-stordalen
morgenkaffen-med-finansavisen
utbytte
liberal-halvtime
rss-markedspuls-2