Practical AI

Optimizing for efficiency with IBM’s Granite

Mar 14, 2025 · 43 min · Episode 306 · 42.0 MB
0:00-43:38

Streams straight from the publisher. podnod never proxies or re-hosts episode audio.

We often judge AI models by leaderboard scores, but what if efficiency matters more? Kate Soule from IBM joins us to discuss how Granite AI is rethinking AI at the edge—breaking tasks into smaller, efficient components and co-designing models with hardware. She also shares why AI should prioritize efficiency frontiers over incremental benchmark gains and how seamless model routing can optimize performance. 

Featuring:

Links: