Localizing and Editing Knowledge in LLMs with Peter Hase - Listen

The TWIML AI Podcast (formerly This Week in...

Localizing and Editing Knowledge in LLMs with Peter Hase

Listen now

Description

Today we're joined by Peter Hase, a fifth-year PhD student at the University of North Carolina NLP lab. We discuss "scalable oversight", and the importance of developing a deeper understanding of how large neural networks make decisions. We learn how matrices are probed by interpretability researchers, and explore the two schools of thought regarding how LLMs store knowledge. Finally, we discuss the importance of deleting sensitive information from model weights, and how "easy-to-hard generalization" could increase the risk of releasing open-source foundation models. The complete show notes for this episode can be found at twimlai.com/go/679.

More Episodes

See all »

Controlling Fusion Reactor Instability with Deep Reinforcement Learning with Aza Jalalvand

Today we're joined by Azarakhsh (Aza) Jalalvand, a research scholar at Princeton University, to discuss his work using deep reinforcement learning to control plasma instabilities in nuclear fusion reactors. Aza explains his team developed a model to detect and avoid a fatal plasma instability...

Published 04/29/24

GraphRAG: Knowledge Graphs for AI Applications with Kirk Marple

Today we're joined by Kirk Marple, CEO and founder of Graphlit, to explore the emerging paradigm of "GraphRAG," or Graph Retrieval Augmented Generation. In our conversation, Kirk digs into the GraphRAG architecture and how Graphlit uses it to offer a multi-stage workflow for ingesting,...

Published 04/22/24

The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

Published 04/22/24