EPISODE · Aug 27, 2026 · 6 MIN
AWS and NVIDIA's 2M GPU Push, OpenAI's Jalapeño Chip, and Google Gemini Legal Agents
from AI Convo Cast · host AI Convo Cast
In this episode, we cover AWS and NVIDIA's plan to deploy two million more GPUs, OpenAI's first benchmarks for its custom Jalapeño inference chip built with Broadcom, and Google Cloud's new Gemini Enterprise agents for legal and financial work. We also break down a report that China's Moonshot AI is negotiating to bring its massive Kimi K3 model onto Azure, AWS, and Google Cloud. From NVIDIA's Vera CPUs and Nemotron open models to OpenAI's inference efficiency gains and agentic AI in regulated industries, we explore the strategic tensions shaping AI infrastructure, custom chips, and enterprise adoption.https://www.aiconvocast.comHelp support the podcast by using our affiliate links:Eleven Labs: https://try.elevenlabs.io/ibl30sgkibkvDisclaimer:This podcast is an independent production and is not affiliated with, endorsed by, or sponsored by AWS, NVIDIA, OpenAI, Broadcom, Google, Microsoft, Moonshot AI, or any other entities mentioned unless explicitly mentioned. The content provided is for educational and entertainment purposes only and does not constitute professional, financial, or legal advice. Affiliate links may earn the podcast a commission at no additional cost to you. All trademarks, logos, and copyrights mentioned are the property of their respective owners.
Embed this episode
NOW PLAYING
AWS and NVIDIA's 2M GPU Push, OpenAI's Jalapeño Chip, and Google Gemini Legal Agents
No transcript for this episode yet
Similar Episodes
No similar episodes found.
Similar Podcasts
No similar podcasts found.