Amnesic Probing: Behavioral Explanation With Amnesic Counterfactuals

Yanai Elazar; Shauli Ravfogel; Alon Jacovi; Yoav Goldberg

Vol. 9 (2021)

TACL approved

Amnesic Probing: Behavioral Explanation With Amnesic Counterfactuals

Published 2022-01-04

Yanai Elazar
Shauli Ravfogel
Alon Jacovi
Yoav Goldberg

Yanai Elazar
Bar-Ilan University AI2

Shauli Ravfogel
Bar-Ilan University AI2

Alon Jacovi
Bar-Ilan University

Yoav Goldberg
Bar-Ilan University AI2

Abstract

A growing body of work makes use of probing in order to investigate the working of neural models, often considered black boxes. Recently, an ongoing debate emerged surrounding the limitations of the probing paradigm. In this work, we point out the inability to infer behavioral conclusions from probing results, and offer an alternative method which focuses on how the information is being used, rather than on what information is encoded.

Our method, Amnesic Probing, follows the intuition that the utility of a property for a given task can be assessed by measuring the influence of a causal intervention which removes it from the representation. Equipped with this new analysis tool, we can ask questions that were not possible before, e.g. is part-of-speech information important for word prediction? We perform a series of analyses on BERT to answer these types of questions. Our findings demonstrate that conventional probing performance is not correlated to task importance, and we call for increased scrutiny of claims that draw behavioral or causal conclusions from probing results.

Article at MIT Press Presented at ACL 2021