“Not a Paper: “Frontier Lab CEOs are Capable of In-Context Scheming”” by LawrenceC episode artwork

EPISODE · Apr 29, 2026 · 14 MIN

“Not a Paper: “Frontier Lab CEOs are Capable of In-Context Scheming”” by LawrenceC

from LessWrong (30+ Karma)

(Fragments from a research paper that will never be written) Extended Abstract. The frontier AI developers are becoming increasingly powerful and wealthy, significantly increasing their potential for risks. One concern is that of executive misalignment: when the CEO has different incentives and goals than that of the board of directors, or of humanity as a whole. Our work proposes three different threat models, under which executive misalignment can lead to concrete harm. We perform two evaluations to understand the capabilities and propensities of current humans in relation to executive misalignment: First, we developed a variant of the standard SAD dataset, SAD-Executive Reasoning (SAD-ER), in order to assess the situational awareness of human CEOs on a range of behavioral tests. We find that n=6 current CEOs can (i) recognize their previous public statements, (ii) understand their roles and responsibilities, (iii) determine if an interviewer is friendly or hostile, and (iv) follow instructions that depend on self knowledge. Second, we stress-tested the same 6 leading AI developers in hypothetical corporate environments to identify potentially risky behaviors before they cause real harm. We find that, even without explicit instructions, all 6 developers are willing to engage in strategic behavior (such as [...] The original text contained 2 footnotes which were omitted from this narration. --- First published: April 28th, 2026 Source: https://www.lesswrong.com/posts/FuauQjjbTCS5QFLk8/not-a-paper-frontier-lab-ceos-are-capable-of-in-context --- Narrated by TYPE III AUDIO.

Episode metadata supplied by the publisher feed · Published Apr 29, 2026

Embed this episode

NOW PLAYING

“Not a Paper: “Frontier Lab CEOs are Capable of In-Context Scheming”” by LawrenceC

0:00 14:51

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of LessWrong (30+ Karma)?

This episode is 14 minutes long.

When was this LessWrong (30+ Karma) episode published?

This episode was published on April 29, 2026.

Can I download this LessWrong (30+ Karma) episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!