Active Learning for Moral Preference Elicitation: Challenges and Nuances episode artwork

EPISODE · Apr 11, 2025 · 21 MIN

Active Learning for Moral Preference Elicitation: Challenges and Nuances

from Best AI papers explained · host Enoch H. Kang

We explore the efficacy of active learning for understanding moral preferences, which are people's views on right actions when harm is involved. While active learning efficiently learns preferences in some areas, the authors argue it relies on assumptions like stable preferences, accurate models, and limited response noise, which may not hold for moral judgments. Through simulations testing these assumptions, the study finds that active learning's performance can be similar to or worse than random questioning when moral preferences are unstable, models are misspecified, or responses are very noisy, highlighting the need for caution when applying active learning to elicit moral preferences.

Episode metadata supplied by the publisher feed · Published Apr 11, 2025

Embed this episode

NOW PLAYING

Active Learning for Moral Preference Elicitation: Challenges and Nuances

0:00 21:58

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of Best AI papers explained?

This episode is 21 minutes long.

When was this Best AI papers explained episode published?

This episode was published on April 11, 2025.

Can I download this Best AI papers explained episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!