Bird's Eye: Probing for Linguistic Graph Structures with a Simple Information-Theoretic Approach

Hou, Yifan; Sachan, Mrinmaya

Computer Science > Computation and Language

arXiv:2105.02629 (cs)

[Submitted on 6 May 2021 (v1), last revised 25 May 2021 (this version, v4)]

Title:Bird's Eye: Probing for Linguistic Graph Structures with a Simple Information-Theoretic Approach

Authors:Yifan Hou, Mrinmaya Sachan

View PDF

Abstract:NLP has a rich history of representing our prior understanding of language in the form of graphs. Recent work on analyzing contextualized text representations has focused on hand-designed probe models to understand how and to what extent do these representations encode a particular linguistic phenomenon. However, due to the inter-dependence of various phenomena and randomness of training probe models, detecting how these representations encode the rich information in these linguistic graphs remains a challenging problem. In this paper, we propose a new information-theoretic probe, Bird's Eye, which is a fairly simple probe method for detecting if and how these representations encode the information in these linguistic graphs. Instead of using classifier performance, our probe takes an information-theoretic view of probing and estimates the mutual information between the linguistic graph embedded in a continuous space and the contextualized word representations. Furthermore, we also propose an approach to use our probe to investigate localized linguistic information in the linguistic graphs using perturbation analysis. We call this probing setup Worm's Eye. Using these probes, we analyze BERT models on their ability to encode a syntactic and a semantic graph structure, and find that these models encode to some degree both syntactic as well as semantic information; albeit syntactic information to a greater extent.

Comments:	Accepted for publication at ACL 2021. This is the camera ready version. Our implementation is available in this https URL
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2105.02629 [cs.CL]
	(or arXiv:2105.02629v4 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2105.02629

Submission history

From: Yifan Hou [view email]
[v1] Thu, 6 May 2021 13:01:57 UTC (127 KB)
[v2] Sun, 16 May 2021 13:25:16 UTC (155 KB)
[v3] Sat, 22 May 2021 08:13:05 UTC (160 KB)
[v4] Tue, 25 May 2021 21:44:25 UTC (160 KB)

Computer Science > Computation and Language

Title:Bird's Eye: Probing for Linguistic Graph Structures with a Simple Information-Theoretic Approach

Submission history

Access Paper:

Current browse context:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Bird's Eye: Probing for Linguistic Graph Structures with a Simple Information-Theoretic Approach

Submission history

Access Paper:

Current browse context:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators