Assistant Professor, Department of Computational and Data Sciences, Indian Institute of Science (IISc), Bangalore
- Date & time
- Monday, April 17, 2023 · 1:00 PM
- Location
- Donald Bren Hall 4011
Abstract
While large deep learning models have become increasingly accurate, concerns about their (lack of) interpretability have taken center stage. In response, a growing subfield on interpretability and analysis of these models has emerged. While hundreds of techniques have been proposed to explain predictions of models, what aims these explanations serve and how they ought to be evaluated are often unstated. In this talk, I will present a framework to quantify the value of explanations, along with specific applications in a variety of contexts. I would end with some of my thoughts on evaluating large language models and the rationales they generate.
About the speaker
Danish Pruthi is an assistant professor at the Indian Institute of Science (IISc), Bangalore. He received his Ph.D. from the School of Computer Science at Carnegie Mellon University, advised by Graham Neubig and Zachary Lipton. He is broadly interested in natural language processing and deep learning, with a focus on model interpretability, and is a recipient of the Siebel Scholarship and the CMU Presidential Fellowship.