Attacking LLMs for fun and profit (Ep. 239)

Data Science at Home - Een podcast door Francesco Gadaleta

Probeer Podimo de eerste 60! dagen gratis

Luister 30 dagen gratis naar exclusieve podcasts en duizenden luisterboeken

Categorieën:

As a continuation of Episode 238, I explain some effective and fun attacks to conduct against LLMs. Such attacks are even more effective on models served locally, that are hardly controlled by human feedback. Have great fun and learn them responsibly. References https://www.jailbreakchat.com/ https://www.reddit.com/r/ChatGPT/comments/10tevu1/new_jailbreak_proudly_unveiling_the_tried_and/ https://arxiv.org/abs/2305.13860

Visit the podcast's native language site