#324 Sharon Zhou: Inside AMD's Plan to Build Self-Improving AI

#324 Sharon Zhou: Inside AMD's Plan to Build Self-Improving AI

AI is not just getting smarter. It is getting faster by learning how to optimize the hardware it runs on.

In this episode, Sharon Zhou, VP of AI at AMD and former Stanford AI researcher, explains how language models are beginning to write and optimize their own GPU kernel code. We explore what self improving AI actually means, how reinforcement learning is used in post training, and why kernel optimization could be one of the most overlooked scaling levers in modern AI.

Sharon breaks down how GPU efficiency impacts the cost of training and inference, why catastrophic forgetting remains a challenge in continual learning, and how verifiable rewards from hardware profiling can help models improve themselves. The conversation also dives into compute economics, synthetic data, RLHF, and why infrastructure may define the next phase of AI progress.

If you want to understand where AI scaling is really happening beyond bigger models and more data, this episode goes under the hood.


Stay Updated:

Craig Smith on X: https://x.com/craigss

Eye on A.I. on X: https://x.com/EyeOn_AI


(00:00) Preview and Intro

(00:25) Sharon Zhou's Background and Transition to AMD

(02:00) What Is Self-Improving AI?

(04:16) What Is a GPU Kernel and Why It Matters

(07:01) Using AI Agents and Evolutionary Strategies to Write Kernels

(11:31) Just-In-Time Optimization and Continual Learning

(13:59) Self-Improving AI at the Infrastructure Layer

(16:15) Synthetic Data and Models Generating Their Own Training Data

(20:48) AMD's AI Strategy: Research Meets Product

(23:22) Inside the NeurIPS Tutorial on AI-Generated Kernels

(30:59) Reinforcement Learning Beyond RLHF

(39:09) 10x Faster Kernels vs 10x More Compute

(41:50) Will Efficiency Reduce Chip Demand?

(42:18) Beyond Language Models: Diffusion, JEPA, and Robotics

(45:34) Educating the Next Generation of AI Builders

Det här avsnittet är hämtat från ett öppet RSS-flöde och publiceras inte av Podme. Det kan innehålla reklam.

Avsnitt(374)

95% of AI Agent Projects Fail to Reach Production. Here's Why | Manoj Saxena, TrustWise

95% of AI Agent Projects Fail to Reach Production. Here's Why | Manoj Saxena, TrustWise

It takes a few hours to build an AI agent. It takes six to seven months to get it into production. Manoj Saxena - the executive who commercialized IBM Watson - built TrustWise around one thesis: intel...

24 Aug 1h 1min

From Zero to 150 Robots in Just 20 Months | Mike LeBlanc, Foundation Future Industries

From Zero to 150 Robots in Just 20 Months | Mike LeBlanc, Foundation Future Industries

Most humanoid robot companies are still running curated demos in replica environments. Foundation Future Industries is running 150 robots on real automotive production lines in Georgia, and heading to...

19 Aug 1h 6min

Why People Are Paying 10x More for AI - and What That Means for the Chip Market | Sid Sheth, d-Matrix

Why People Are Paying 10x More for AI - and What That Means for the Chip Market | Sid Sheth, d-Matrix

The AI chip market looks monolithic from the outside - NVIDIA dominates, and everyone else is fighting for scraps. But d-Matrix's CEO Sid Sheth argues that the market is quietly splitting into two dis...

17 Aug 50min

American Companies Have 36 Months to Go AI-Native or Get Left Behind | Drew Cukor, TWG AI

American Companies Have 36 Months to Go AI-Native or Get Left Behind | Drew Cukor, TWG AI

The same tools that slowed the U.S. military down in Afghanistan (PowerPoint, Excel, email, and Word) are now slowing American businesses down in the AI race. Drew Cukor spent 30 years as a Marine int...

13 Aug 58min

In 5 Years, 90% of What You Use AI For Will Run on Your Smartphone | Paolo Ardoino, Tether

In 5 Years, 90% of What You Use AI For Will Run on Your Smartphone | Paolo Ardoino, Tether

Hundreds of billions of dollars are flowing into AI data centers right now, and Paolo Ardoino, CEO of Tether - the company behind the world's most widely used stablecoin with 573 million users - think...

10 Aug 58min

AI Agents Fixing Your IT Before You Even Know Something Broke | Erhan Giral & Ryan Manning, BMC Helix

AI Agents Fixing Your IT Before You Even Know Something Broke | Erhan Giral & Ryan Manning, BMC Helix

Most enterprise IT teams spend the majority of their time fighting the same fires repeatedly. BMC Helix is building the AI system that handles those fires automatically, detecting anomalies, tracing r...

3 Aug 59min

Real AI Transformation Costs HALF of Everyone's Salary for 2 Years | Chris Blackburn, Liatrio

Real AI Transformation Costs HALF of Everyone's Salary for 2 Years | Chris Blackburn, Liatrio

Most companies think they're transforming with AI. They're not, and the gap between what they believe and what's actually happening on the ground is costing them far more than they realize. In this e...

30 Juli 1h 5min

"According to NASA's Definition of Life, I'm Not Alive" - Why Nobody Can Define Life | Dr. Kate Adamala

"According to NASA's Definition of Life, I'm Not Alive" - Why Nobody Can Define Life | Dr. Kate Adamala

Nobody has ever built a cell from scratch - assembled entirely from purified molecules on a shelf - that can feed itself, grow, and split into daughter cells through its own genetic activity. Until no...

29 Juli 46min

Populärt inom Teknik

uppgang-och-fall
market-makers
rss-elektrikerpodden
rss-laddstationen-med-elbilen-i-sverige
 och-bilen-gar-bra
skogsforum-podcast
natets-morka-sida
bli-saker-podden
rss-technokratin
rss-uppgang-och-fall
bilar-med-sladd
rss-veckans-ai
rss-fabriken-2
klocksnack-tillsammans-med-nymans-ur-1851
hej-bruksbil
rss-en-ai-till-kaffet
elbilsveckan
bosse-bildoktorn-och-hasse-p
ai-sweden-podcast
developers-mer-an-bara-kod