Artwork

Indhold leveret af Security Weekly Productions and Security Weekly. Alt podcastindhold inklusive episoder, grafik og podcastbeskrivelser uploades og leveres direkte af Security Weekly Productions and Security Weekly eller deres podcastplatformspartner. Hvis du mener, at nogen bruger dit ophavsretligt beskyttede værk uden din tilladelse, kan du følge processen beskrevet her https://da.player.fm/legal.
Player FM - Podcast-app
Gå offline med appen Player FM !

AI Red Teaming and AI Safety - Amanda Minnich - ESW #371

41:17
 
Del
 

Manage episode 433346603 series 72776
Indhold leveret af Security Weekly Productions and Security Weekly. Alt podcastindhold inklusive episoder, grafik og podcastbeskrivelser uploades og leveres direkte af Security Weekly Productions and Security Weekly eller deres podcastplatformspartner. Hvis du mener, at nogen bruger dit ophavsretligt beskyttede værk uden din tilladelse, kan du følge processen beskrevet her https://da.player.fm/legal.

In this interview we explore the new and sometimes strange world of redteaming AI. I have SO many questions, like what is AI safety?

We'll discuss her presence at Black Hat, where she delivered two days of training and participated on an AI safety panel.

We'll also discuss the process of pentesting an AI. Will pentesters just have giant cheatsheets or text files full of adversarial prompts? How can we automate this? Will an AI generate adversarial prompts you can use against another AI? And finally, what do we do with the results?

Resources:

Show Notes: https://securityweekly.com/esw-371

  continue reading

4275 episoder

Artwork
iconDel
 
Manage episode 433346603 series 72776
Indhold leveret af Security Weekly Productions and Security Weekly. Alt podcastindhold inklusive episoder, grafik og podcastbeskrivelser uploades og leveres direkte af Security Weekly Productions and Security Weekly eller deres podcastplatformspartner. Hvis du mener, at nogen bruger dit ophavsretligt beskyttede værk uden din tilladelse, kan du følge processen beskrevet her https://da.player.fm/legal.

In this interview we explore the new and sometimes strange world of redteaming AI. I have SO many questions, like what is AI safety?

We'll discuss her presence at Black Hat, where she delivered two days of training and participated on an AI safety panel.

We'll also discuss the process of pentesting an AI. Will pentesters just have giant cheatsheets or text files full of adversarial prompts? How can we automate this? Will an AI generate adversarial prompts you can use against another AI? And finally, what do we do with the results?

Resources:

Show Notes: https://securityweekly.com/esw-371

  continue reading

4275 episoder

Alle episoder

×
 
Loading …

Velkommen til Player FM!

Player FM is scanning the web for high-quality podcasts for you to enjoy right now. It's the best podcast app and works on Android, iPhone, and the web. Signup to sync subscriptions across devices.

 

Hurtig referencevejledning