RAFT: Adapting Language Model to Domain Specific RAG episode artwork

EPISODE · Jun 28, 2024 · 44 MIN

RAFT: Adapting Language Model to Domain Specific RAG

from Deep Papers

Where adapting LLMs to specialized domains is essential (e.g., recent news, enterprise private documents), we discuss a paper that asks how we adapt pre-trained LLMs for RAG in specialized domains. SallyAnn DeLucia is joined by Sai Kolasani, researcher at UC Berkeley’s RISE Lab (and Arize AI Intern), to talk about his work on RAFT: Adapting Language Model to Domain Specific RAG. RAFT (Retrieval-Augmented FineTuning) is a training recipe that improves an LLM’s ability to answer questions in a “open-book” in-domain settings. Given a question, and a set of retrieved documents, the model is trained to ignore documents that don’t help in answering the question (aka distractor documents). This coupled with RAFT’s chain-of-thought-style response, helps improve the model’s ability to reason. In domain-specific RAG, RAFT consistently improves the model’s performance across PubMed, HotpotQA, and Gorilla datasets, presenting a post-training recipe to improve pre-trained LLMs to in-domain RAG. Read it on the blog: https://arize.com/blog/raft-adapting-language-model-to-domain-specific-rag/Learn more about AI observability and evaluation, join the Arize AI Slack community or get the latest on LinkedIn and X.

Episode metadata supplied by the publisher feed · Published Jun 28, 2024

Embed this episode

Where adapting LLMs to specialized domains is essential (e.g., recent news, enterprise private documents), we discuss a paper that asks how we adapt pre-trained LLMs for RAG in specialized domains. SallyAnn DeLucia is joined by Sai Kolasani, researcher at UC Berkeley’s RISE Lab (and Arize AI Intern), to talk about his work on RAFT: Adapting Language Model to Domain Specific RAG. RAFT (Retrieval-Augmented FineTuning) is a training recipe that improves an LLM’s ability to answer questions...

Distinct summary based on available episode metadata or transcript content.

NOW PLAYING

RAFT: Adapting Language Model to Domain Specific RAG

0:00 44:01

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Deep Papers?

This episode is 44 minutes long.

When was this Deep Papers episode published?

This episode was published on June 28, 2024.

Can I download this Deep Papers episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!