Attacking LLMs for fun and profit (Ep. 239)
Episode 241 of the Data Science at Home podcast, hosted by Francesco Gadaleta, titled "Attacking LLMs for fun and profit (Ep. 239)" was published on September 18, 2023 and runs 22 minutes.
September 18, 2023 ·22m · Data Science at Home
Summary
As a continuation of Episode 238, I explain some effective and fun attacks to conduct against LLMs. Such attacks are even more effective on models served locally, that are hardly controlled by human feedback. Have great fun and learn them responsibly. References https://www.jailbreakchat.com/ https://www.reddit.com/r/ChatGPT/comments/10tevu1/new_jailbreak_proudly_unveiling_the_tried_and/ https://arxiv.org/abs/2305.13860
Episode Description
As a continuation of Episode 238, I explain some effective and fun attacks to conduct against LLMs. Such attacks are even more effective on models served locally, that are hardly controlled by human feedback.
Have great fun and learn them responsibly.
References
https://www.jailbreakchat.com/
https://www.reddit.com/r/ChatGPT/comments/10tevu1/new_jailbreak_proudly_unveiling_the_tried_and/
https://arxiv.org/abs/2305.13860
Similar Episodes
Apr 13, 2026 ·4m
Apr 12, 2026 ·5m
Apr 11, 2026 ·5m
Apr 10, 2026 ·4m
Apr 9, 2026 ·3m
Apr 8, 2026 ·3m