Attacking LLMs for fun and profit (Ep. 239) - Data Science at Home Podcast

September 20, 2023

Attacking LLMs for fun and profit

As a continuation of Episode 238, I explain some effective and fun attacks to conduct against LLMs. Such attacks are even more effective on models served locally, that are hardly controlled by human feedback.

Have great fun and learn them responsibly.

References

https://www.jailbreakchat.com/

New jailbreak! Proudly unveiling the tried and tested DAN 5.0 – it actually works – Returning to DAN, and assessing its limitations and capabilities.
byu/SessionGloomy inChatGPT

https://arxiv.org/abs/2305.13860

Image Credit

June 23, 2026

AI is the Concorde of our time (Ep. 309)

May 19, 2026

Recommend and manipulate: the dangers of the attention economy

May 19, 2026

Attacking LLMs for fun and profit (Ep. 239) - Data Science at Home Podcast

Related posts

AI is the Concorde of our time (Ep. 309)

Recommend and manipulate: the dangers of the attention economy

Social media is an ant mill (Internet is a disaster) (Ep. 303)