Localizing And Editing Knowledge In LLMs With Peter Hase - #679 The TWIML AI Podcast (formerly This Week In Machine Learning & Artificial Intelligence) podcast

Artwork

Artificial Intelligence Tech News Artificialintelligence Machinelearning Samcharrington Technology Thisweekinmachinelearning Sam Charrington Thetwimlaipocast Twimlaipodcast Tech News China TWIML Datascience Science

Indhold leveret af TWIML and Sam Charrington. Alt podcastindhold inklusive episoder, grafik og podcastbeskrivelser uploades og leveres direkte af TWIML and Sam Charrington eller deres podcastplatformspartner. Hvis du mener, at nogen bruger dit ophavsretligt beskyttede værk uden din tilladelse, kan du følge processen beskrevet her https://da.player.fm/legal.

The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence) « »
Localizing and Editing Knowledge in LLMs with Peter Hase - #679

3M ago 49:46

Del

MP3•Episode hjem

Indhold leveret af TWIML and Sam Charrington. Alt podcastindhold inklusive episoder, grafik og podcastbeskrivelser uploades og leveres direkte af TWIML and Sam Charrington eller deres podcastplatformspartner. Hvis du mener, at nogen bruger dit ophavsretligt beskyttede værk uden din tilladelse, kan du følge processen beskrevet her https://da.player.fm/legal.

Today we're joined by Peter Hase, a fifth-year PhD student at the University of North Carolina NLP lab. We discuss "scalable oversight", and the importance of developing a deeper understanding of how large neural networks make decisions. We learn how matrices are probed by interpretability researchers, and explore the two schools of thought regarding how LLMs store knowledge. Finally, we discuss the importance of deleting sensitive information from model weights, and how "easy-to-hard generalization" could increase the risk of releasing open-source foundation models.

The complete show notes for this episode can be found at twimlai.com/go/679.

… continue reading

710 episoder

#Artificial Intelligence #Tech News #Artificialintelligence #Machinelearning #Samcharrington #Technology #Thisweekinmachinelearning #Sam Charrington #Thetwimlaipocast #Twimlaipodcast #Tech #News #China #TWIML #Datascience #Science

Artwork

Localizing and Editing Knowledge in LLMs with Peter Hase - #679

The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

1,696 subscribers

published 3M ago

Del

MP3•Episode hjem

Indhold leveret af TWIML and Sam Charrington. Alt podcastindhold inklusive episoder, grafik og podcastbeskrivelser uploades og leveres direkte af TWIML and Sam Charrington eller deres podcastplatformspartner. Hvis du mener, at nogen bruger dit ophavsretligt beskyttede værk uden din tilladelse, kan du følge processen beskrevet her https://da.player.fm/legal.

Today we're joined by Peter Hase, a fifth-year PhD student at the University of North Carolina NLP lab. We discuss "scalable oversight", and the importance of developing a deeper understanding of how large neural networks make decisions. We learn how matrices are probed by interpretability researchers, and explore the two schools of thought regarding how LLMs store knowledge. Finally, we discuss the importance of deleting sensitive information from model weights, and how "easy-to-hard generalization" could increase the risk of releasing open-source foundation models.

The complete show notes for this episode can be found at twimlai.com/go/679.

… continue reading

710 episoder

#Artificial Intelligence #Tech News #Artificialintelligence #Machinelearning #Samcharrington #Technology #Thisweekinmachinelearning #Sam Charrington #Thetwimlaipocast #Twimlaipodcast #Tech #News #China #TWIML #Datascience #Science

所有剧集

×

Velkommen til Player FM!

Player FM is scanning the web for high-quality podcasts for you to enjoy right now. It's the best podcast app and works on Android, iPhone, and the web. Signup to sync subscriptions across devices.

Lyt til 500+ emner