EPISODE · Aug 18, 2026 · 6 MIN
Inside the 20x AI Cost Trap and Model Routing
from The AI Engineering Podcast · host Jellypod
This episode explores how enterprise AI teams are slashing runaway inference costs with intelligent model routing, prompt harnessing, and open-weight alternatives. It also dives into the use of agentic search, shadow evals, and real-world feedback loops to decide when smaller models can outperform expensive frontier systems.
Embed this episode
Ready to play
Inside the 20x AI Cost Trap and Model Routing
No transcript for this episode yet
Similar Episodes
No similar episodes found.
Similar Podcasts
No similar podcasts found.