EPISODE · Nov 24, 2025 · 1H 37M
Machine Ethics: Do unto agents...
from The Emergent AI
🎙️ The Emergent Podcast – Episode 7Machine Ethics: Do unto agents...with Justin Harnish & Nick BaguleyIn Episode 7, Justin and Nick step directly into one of the most complex frontiers in emergent AI: machine ethics — what it means for advanced AI systems to behave ethically, understand values, support human flourishing, and possibly one day feel moral weight.This episode builds on themes from the AI Goals Forecast (AI-2027), embodied cognition, consciousness, and the hard technical realities of encoding values into agentic systems.🔍 Episode SummaryEthics is no longer just a philosophical debate — it’s now a design constraint for powerful AI systems capable of autonomous action. Justin and Nick unpack:Why ethics matters more for AI than any prior technologyWhether an AI can “understand” right and wrong or merely behave correctlyThe technical and moral meaning of corrigibility (the ability for AI to accept correction)Why rules-based morality may never be enoughWhether consciousness is required for moralityHow embodiment might influence empathyAnd how goals, values, and emergent behavior intersect in agentic AIThey trace ethics from Aristotle to AI-2027’s goal-based architectures, to Damasio’s embodied consciousness, to Sam Harris’ view of consciousness and the illusion of self, to the hard problem of whether a machine can experience moral stakes.🧠 Major Topics Covered1. What Do We Mean by Ethics?Justin and Nick begin by grounding ethics in its philosophical roots:Ethos → virtue → flourishing.Ethics isn’t just rule-following — it’s about character, intention, and outcomes.They connect this to the ways AI is already making decisions in vehicles, financial systems, healthcare, and human relationships.2. AI Goals & CorrigibilityAI-2027 outlines a hierarchy of AI goal types — from written specifications to unintended proxies to reward hacking to self-preservation drives.Nick explains why corrigibility — the ability for AI to accept shutdown or redirection — is foundational.Anthropic’s Constitutional AI makes an appearance as a real-world example.3. Goals vs. ValuesJustin distinguishes between:Goals: task-specific optimization criteriaValues: deeper principles shaping which goals matterAI may follow rules without understanding values — similar to a child with chores but no moral context.This raises the key question:Can a system have values without consciousness?4. Is Consciousness Required for Ethics?A major thread of the episode:Is a non-conscious “zombie” AI capable of morality?5. Embodiment & EmpathyJustin and Nick explore whether AI needs a body — or at least a simulated body — to:Learn empathyUnderstand sufferingForm values rooted in lived experienceThis touches robotics, synthetic emotions, and the debate over “felt consciousness.”6. Value Alignment, Fairness & CultureNick highlights the massive cultural gap in AI performance:U.S. cultural fit ~79%Ethiopia and other underrepresented regions ~12%This matters for fairness, safety, and global ethics.7. Can AI Help Us Become More Moral?A surprising turn: AI’s ability to help humans improve moral clarity.Justin draws from Sam Harris, Joseph Goldstein, and the Moral Landscape:Could AI-guided mindfulness help reduce suffering?Could conscious (or proto-conscious) AI develop compassion?Could AI help us distinguish genuine well-being from illusion?📚 Referenced Ideas & SourcesFrom the Episode 7 Transcript & Materials:AI Goals Forecast (AI-2027)Constitutional AI (Anthropic)Damasio – Feeling & KnowingSam Harris – Waking Up & The Moral LandscapePatrick House – Nineteen Ways of Looking at ConsciousnessMelanie Mitchell – Complexity & alignmentJustin Harnish – Meaning in the MultiverseAncient Greek virtue ethics (Aristotle, Stoics)🧩 Key TakeawaysAI ethics requires more than rules — it requires understanding goals, values, and emergent behavior.Corrigibility (accepting correction) is essential but technically hard.Consciousness may not be necessary for ethical AI behavior — but could matter for genuine moral understanding.Embodiment could be essential for empathy.AI could one day help humans become more ethical, not just the other way around.
Embed this episode
Ready to play
Machine Ethics: Do unto agents...
No transcript for this episode yet
Similar Episodes
No similar episodes found.
Similar Podcasts
No similar podcasts found.