PodParley PodParley
LLMs Still Can't Plan; Can LRMs?

EPISODE · Oct 18, 2024 · 8 MIN

LLMs Still Can't Plan; Can LRMs?

from LlamaCast · host Shahriar Shariati

📈 LLMs Still Can't Plan; Can LRMs?The paper "LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench" investigates the ability of large language models (LLMs) to plan, using a benchmark called PlanBench. The authors find that while OpenAI's new "Large Reasoning Model" (LRM) o1 shows significant improvement in planning abilities, it still falls short of fully achieving the task. This research highlights the need for further investigation into the accuracy, efficiency, and guarantees associated with these advanced models.📎 Link to paper

NOW PLAYING

LLMs Still Can't Plan; Can LRMs?

0:00 8:19

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

URL copied to clipboard!