EPISODE · Jun 17, 2024 · 16 MIN
AI Safety: Constitutional AI vs Human Feedback
from Super Prompt: Generative AI · host Tony Wan
With great power comes great responsibility. How do leading AI companies implement safety and ethics as language models scale? OpenAI uses Model Spec combined with RLHF (Reinforcement Learning from Human Feedback). Anthropic uses Constitutional AI. The technical approaches to maximizing usefulness while minimizing harm. Solo episode on AI alignment.REFERENCEOpenAI Model Spechttps://cdn.openai.com/spec/model-spec-2024-05-08.html#overviewAnthropic Constitutional AIhttps://www.anthropic.com/news/claudes-constitutionTo stay in touch, sign up for our newsletter at https://superprompt.substack.com
Embed this episode
Ready to play
AI Safety: Constitutional AI vs Human Feedback
No transcript for this episode yet
Similar Episodes
No similar episodes found.