Beyond Transcripts:  Language Nuances and Audio Signals with Carter Huffman of Modulate

Beyond Transcripts: Language Nuances and Audio Signals with Carter Huffman of Modulate

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub talk with Carter Huffman, CTO and co-founder of Modulate AI, about his path from machine learning work at NASA's Jet Propulsion Lab to building voice AI that understands conversations. Carter explains why moderation in gaming is hard because you don't want to ban players unfairly, and contrasts big foundation models with orchestrated ensembles of many tiny models that require high-quality, globally vetted data labeling. They discuss the nuance of classifying hate speech, expansion into detecting fraud and manipulation in delivery and call-center contexts, and monitoring misbehaving AI voice agents. The conversation covers why conversation is more than transcripts, possible therapeutic/telehealth uses of Modulate, analyzing data at a massive scale, and ambitions for audio generation using hierarchical edge-and-cloud approaches. The episode ends with a humorous anecdote about two factor authenticaiton failure. 00:00 Podcast Cold Open 00:48 Meet Carter Huffman 02:06 JPL Spacecraft Autonomy 04:18 From JPL to Audio AI 06:18 Why Audio Is Hard 07:44 Voice AI Use Cases 12:49 Tiny Models Orchestration 15:56 Data Labeling at Scale 17:17 Defining Toxic Behavior 18:58 Nuanced Language Moderation 20:04 Scaling Ensemble Models 21:39 GPU Crunch During Launch 22:29 Beyond Gaming Use Cases 26:03 AI Agents Gone Wrong 28:45 Telehealth and Diagnostics 30:26 Ambient Audio and Privacy 32:26 Edge Ensembles Everywhere 33:25 Audio Synthesis Ambitions 35:24 Latency Hierarchies Explained 38:10 Two Factor Key Fob Fiasco 39:14 Wrap Up and Credits

Resources:

#TechPodcast #EngineeringPodcast #DevTalks #PodcastForDevs #HowManyCTOs #Podcast #CTOs #CTOPodcast #ChiefTechnologyOfficer #Technology #Engineering #SoftwareDevelopment #SoftwareEngineering #TechLeadership #EngineeringLeadership #EngineeringCulture #TechDebates #AI #VoiceTech #MachineLearning #MachineLearningModels #GamingIndustry #AIinnovation #Entrepreneurship #AIConversation #VoiceAssistant #LanguageModeration #GPU #LLMs #LargeLanguageModels

Det här avsnittet är hämtat från ett öppet RSS-flöde och publiceras inte av Podme. Det kan innehålla reklam.

Avsnitt(84)

Fast, On Time, or Fully Utilized? Choosing What to Optimize in Engineering Delivery

Fast, On Time, or Fully Utilized? Choosing What to Optimize in Engineering Delivery

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub discuss a leadership conversation where three senior stakeholders wanted different outcomes from enginee...

25 Aug 16min

The Pace of Innovation: Systems Thinking for Faster Engineering

The Pace of Innovation: Systems Thinking for Faster Engineering

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub open with travel and sports talk, including World Cup rules, then pivot to Scott's change-management tra...

18 Aug 56min

Method to the Madness: A CTO Framework for Systemic AI Adoption

Method to the Madness: A CTO Framework for Systemic AI Adoption

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub discuss how CTOs must frame narratives that help teams organize around ideas, inspired by how Scott expl...

11 Aug 45min

There's A Pattern To Follow: An Interview with Robert "Uncle Bob" Martin

There's A Pattern To Follow: An Interview with Robert "Uncle Bob" Martin

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub are honored to interview Robert "Uncle Bob" Martin, who recounts his 60+ year programming journey from a...

4 Aug 1h 1min

Is Uncle Bob Right? Reading Code vs. Architecting the Future

Is Uncle Bob Right? Reading Code vs. Architecting the Future

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub discuss how Robert "Uncle Bob" Martin—author of Clean Code and Agile Manifesto signer—sparked online con...

28 Juli 43min

Token Maxxing and Local Inference: Navigating the Shift in AI Usage for Developers

Token Maxxing and Local Inference: Navigating the Shift in AI Usage for Developers

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub discuss token spending and the "token maxing" backlash, sparked by a GitHub Copilot budget issue and bro...

21 Juli 1h

This Has Always Been A Problem: Slot Machine Devs and Growing Junior Engineers in the Age of AI

This Has Always Been A Problem: Slot Machine Devs and Growing Junior Engineers in the Age of AI

In this episode of "How Many CTOs Does It Take?" podcast, hosts Scott Porad and Brad Hefta-Gaub discuss how managing armies of coding agents can be physically draining, likening it to context switchin...

14 Juli 56min

From Chaos to Control: Bringing Specification Driven Design to Vibe Coding with David Bauer

From Chaos to Control: Bringing Specification Driven Design to Vibe Coding with David Bauer

In this episode of "How Many CTOs Does It Take?" podcast, host Brad Hefta-Gaub is joined by Dr. David Bauer of Axonis AI, who shares how his background in distributed computing, startups, and the U.S....

7 Juli 51min

Populärt inom Business & ekonomi

framgangspodden
rss-jossan-nina
varvet
badfluence
rss-borsens-finest
avanzapodden
svd-tech-brief
uppgang-och-fall
rss-inga-dumma-fragor-om-pengar
lastbilspodden
fill-or-kill
tabberaset
24fragor
borsmorgon
rikatillsammans-om-privatekonomi-rikedom-i-livet
bathina-en-podcast
rss-dagen-med-di
skaraborgspodden
rss-hos-psykologen
rss-kort-lang-analyspodden-fran-di