RESEARCH

NLP

Information-Theoretic Probing for Linguistic Structure

June 26, 2020

Abstract

The success of neural networks on a diverse set of NLP tasks has led researchers to question how much these networks actually "know" about natural language. Probes are a natural way of assessing this. When probing, a researcher chooses a linguistic task and trains a supervised model to predict annotations in that linguistic task from the network's learned representations. If the probe does well, the researcher may conclude that the representations encode knowledge related to the task. A commonly held belief is that using simpler models as probes is better; the logic is that simpler models will identify linguistic structure, but not learn the task itself. We propose an information-theoretic operationalization of probing as estimating mutual information that contradicts this received wisdom: one should always select the highest performing probe one can, even if it is more complex, since it will result in a tighter estimate, and thus reveal more of the linguistic information inherent in the representation. The experimental portion of our paper focuses on empirically estimating the mutual information between a linguistic property and BERT, comparing these estimates to several baselines. We evaluate on a set of ten typologically diverse languages often underrepresented in NLP research—plus English—totaling eleven languages.

Download the Paper

AUTHORS

Written by

Adina Williams

Joseph Valvoda

Ran Zmigrod

Rowan Hall Maudsley

Ryan Cotterell

Tiago Pimentel

Publisher

ACL

Related Publications

October 02, 2026

RESEARCH

Tightness of the Cycle-Based Relaxation for Completed Length-Three Alpha-Cycles

Aykut Arslan

October 02, 2026

October 02, 2026

RESEARCH

On Solvable Evolution Algebras and a Conjecture by García-Martínez and Pérez-Rodríguez

Andres Barei Bueno

October 02, 2026

October 02, 2026

RESEARCH

String Two-Point Function = Height Function on a Curve

Anindya Dey, Gabriel Herczeg, An Huang, Nicolas Jaramillo Torres, Jacob H. Swenberg

October 02, 2026

October 02, 2026

RESEARCH

Semiabelian Groups Need Not Be Monomial

Joseph Phillip Brennan, Milana Golich

October 02, 2026

Help Us Pioneer The Future of AI

We share our open source frameworks, tools, libraries, and models for everything from research exploration to large-scale production deployment.