Nvidia's Record Revenue 📈 // OpenAI's News Corp Deal 📰 // Your Transformer is Secretly Linear 🔍 episode artwork

EPISODE · May 23, 2024 · 15 MIN

Nvidia's Record Revenue 📈 // OpenAI's News Corp Deal 📰 // Your Transformer is Secretly Linear 🔍

from GPT Reviews · host Earkind

Nvidia's Q1 revenue up 262% to $26.0B, beating estimates. OpenAI's News Corp deal licenses content from WSJ, New York Post and more. PyramidInfer compresses KV cache to save memory during inference for Large Language Models. Your Transformer is Secretly Linear challenges our existing understanding of transformer architectures. Contact:  [email protected] Timestamps: 00:34 Introduction 01:55 Nvidia's Q1 revenue up 262% to $26.0B, beating estimates 03:23 OpenAI’s News Corp deal licenses content from WSJ, New York Post, and more 04:57 Systematically Improving Your RAG 06:18 Fake sponsor 08:17 PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference 09:49 Reducing Transformer Key-Value Cache Size with Cross-Layer Attention 11:48 Your Transformer is Secretly Linear 13:26 Outro

Episode metadata supplied by the publisher feed · Published May 23, 2024

Embed this episode

Ready to play

Nvidia's Record Revenue 📈 // OpenAI's News Corp Deal 📰 // Your Transformer is Secretly Linear 🔍

0:00 15:05

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of GPT Reviews?

This episode is 15 minutes long.

When was this GPT Reviews episode published?

This episode was published on May 23, 2024.

Can I download this GPT Reviews episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!