Retrieval Transformer episode artwork

EPISODE · Jan 14, 2025 · 14 MIN

Retrieval Transformer

from Large Language Model (LLM) Talk · host AI-Talk

The sources describe RETRO (Retrieval-Enhanced Transformer), a language model that enhances its performance by retrieving information from a large database. RETRO uses a key-value store where keys are BERT embeddings of text chunks and values are the text chunks themselves. When processing input, it retrieves similar text chunks from the database to augment the input, allowing it to perform comparably to much larger models. By incorporating this retrieved information through a chunked cross-attention mechanism, RETRO reduces the need to memorize facts and improves its performance on knowledge-intensive tasks. The database contains trillions of tokens.

Episode metadata supplied by the publisher feed · Published Jan 14, 2025

Embed this episode

NOW PLAYING

Retrieval Transformer

0:00 14:12

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Large Language Model (LLM) Talk?

This episode is 14 minutes long.

When was this Large Language Model (LLM) Talk episode published?

This episode was published on January 14, 2025.

Can I download this Large Language Model (LLM) Talk episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!