When and Why LLMs Fail to Reason Globally episode artwork

EPISODE · May 31, 2025 · 18 MIN

When and Why LLMs Fail to Reason Globally

from Best AI papers explained · host Enoch H. Kang

This research explores why Large Language Models (LLMs) struggle with tasks requiring global reasoning over long inputs. The authors propose that these limitations stem from constraints on information flow within LLMs, formalizing this with the Bounded Attention Prefix Oracle (BAPO) model. They classify problems as BAPO-easy or BAPO-hard, predicting that LLMs will fail on the latter. Empirical results with models like GPT-4o, Claude, and Gemini support this prediction, showing poor performance on BAPO-hard tasks even for relatively small inputs. Crucially, the paper demonstrates theoretically and empirically that using Chain of Thought (CoT) reasoning can transform BAPO-hard problems into BAPO-easy ones, significantly improving performance despite potentially high token usage.

Episode metadata supplied by the publisher feed · Published May 31, 2025

Embed this episode

NOW PLAYING

When and Why LLMs Fail to Reason Globally

0:00 18:19

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of Best AI papers explained?

This episode is 18 minutes long.

When was this Best AI papers explained episode published?

This episode was published on May 31, 2025.

Can I download this Best AI papers explained episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!