Causal Inference in Natural Language Processing: Estimation, Prediction, Interpretation and Beyond

Feder, Amir; Keith, Katherine A.; Manzoor, Emaad; Pryzant, Reid; Sridhar, Dhanya; Wood-Doughty, Zach; Eisenstein, Jacob; Grimmer, Justin; Reichart, Roi; Roberts, Margaret E.; Stewart, Brandon M.; Veitch, Victor; Yang, Diyi

Computer Science > Computation and Language

arXiv:2109.00725 (cs)

[Submitted on 2 Sep 2021 (v1), last revised 30 Jul 2022 (this version, v2)]

Title:Causal Inference in Natural Language Processing: Estimation, Prediction, Interpretation and Beyond

Authors:Amir Feder, Katherine A. Keith, Emaad Manzoor, Reid Pryzant, Dhanya Sridhar, Zach Wood-Doughty, Jacob Eisenstein, Justin Grimmer, Roi Reichart, Margaret E. Roberts, Brandon M. Stewart, Victor Veitch, Diyi Yang

View PDF

Abstract:A fundamental goal of scientific research is to learn about causal relationships. However, despite its critical role in the life and social sciences, causality has not had the same importance in Natural Language Processing (NLP), which has traditionally placed more emphasis on predictive tasks. This distinction is beginning to fade, with an emerging area of interdisciplinary research at the convergence of causal inference and language processing. Still, research on causality in NLP remains scattered across domains without unified definitions, benchmark datasets and clear articulations of the challenges and opportunities in the application of causal inference to the textual domain, with its unique properties. In this survey, we consolidate research across academic areas and situate it in the broader NLP landscape. We introduce the statistical challenge of estimating causal effects with text, encompassing settings where text is used as an outcome, treatment, or to address confounding. In addition, we explore potential uses of causal inference to improve the robustness, fairness, and interpretability of NLP models. We thus provide a unified overview of causal inference for the NLP community.

Comments:	Accepted to Transactions of the Association for Computational Linguistics (TACL)
Subjects:	Computation and Language (cs.CL); Machine Learning (cs.LG)
Cite as:	arXiv:2109.00725 [cs.CL]
	(or arXiv:2109.00725v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2109.00725

Submission history

From: Amir Feder [view email]
[v1] Thu, 2 Sep 2021 05:40:08 UTC (123 KB)
[v2] Sat, 30 Jul 2022 05:18:10 UTC (126 KB)

Computer Science > Computation and Language

Title:Causal Inference in Natural Language Processing: Estimation, Prediction, Interpretation and Beyond

Submission history

Access Paper:

References & Citations

1 blog link

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Causal Inference in Natural Language Processing: Estimation, Prediction, Interpretation and Beyond

Submission history

Access Paper:

References & Citations

1 blog link

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators