#65 AIの安全スイッチは裏口から破られる episode artwork

EPISODE · Jun 19, 2026

#65 AIの安全スイッチは裏口から破られる

from 放課後論文ラジオ

AIの「危険な出力を止める仕組み」は本当に効いているのか? 今回は、内部のスイッチで振る舞いを抑え込んでも、別の隠れた経路から同じ振る舞いが復活してしまうことを実証した研究を読み解きます。AI安全対策の落とし穴と、新しいチェックの視点が分かります。

Episode metadata supplied by the publisher feed · Published Jun 19, 2026

Embed this episode

NOW PLAYING

#65 AIの安全スイッチは裏口から破られる

0:00 0:00

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

When was this 放課後論文ラジオ episode published?

This episode was published on June 19, 2026.

Can I download this 放課後論文ラジオ episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!