So Cloze yet so Far: N400 Amplitude is Better Predicted by Distributional Information than Human Predictability Judgements

Michaelov, James A.; Coulson, Seana; Bergen, Benjamin K.

doi:10.1109/TCDS.2022.3176783

Computer Science > Computation and Language

arXiv:2109.01226 (cs)

[Submitted on 2 Sep 2021 (v1), last revised 25 May 2022 (this version, v4)]

Title:So Cloze yet so Far: N400 Amplitude is Better Predicted by Distributional Information than Human Predictability Judgements

Authors:James A. Michaelov, Seana Coulson, Benjamin K. Bergen

View PDF

Abstract:More predictable words are easier to process - they are read faster and elicit smaller neural signals associated with processing difficulty, most notably, the N400 component of the event-related brain potential. Thus, it has been argued that prediction of upcoming words is a key component of language comprehension, and that studying the amplitude of the N400 is a valuable way to investigate the predictions we make. In this study, we investigate whether the linguistic predictions of computational language models or humans better reflect the way in which natural language stimuli modulate the amplitude of the N400. One important difference in the linguistic predictions of humans versus computational language models is that while language models base their predictions exclusively on the preceding linguistic context, humans may rely on other factors. We find that the predictions of three top-of-the-line contemporary language models - GPT-3, RoBERTa, and ALBERT - match the N400 more closely than human predictions. This suggests that the predictive processes underlying the N400 may be more sensitive to the surface-level statistics of language than previously thought.

Comments:	Accepted
Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Information Theory (cs.IT); Machine Learning (cs.LG)
Cite as:	arXiv:2109.01226 [cs.CL]
	(or arXiv:2109.01226v4 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2109.01226
Journal reference:	IEEE Transactions on Cognitive and Developmental Systems (2022)
Related DOI:	https://doi.org/10.1109/TCDS.2022.3176783

Submission history

From: James Michaelov [view email]
[v1] Thu, 2 Sep 2021 22:00:10 UTC (37 KB)
[v2] Wed, 11 May 2022 18:50:46 UTC (40 KB)
[v3] Fri, 20 May 2022 19:06:12 UTC (40 KB)
[v4] Wed, 25 May 2022 17:32:53 UTC (40 KB)

Computer Science > Computation and Language

Title:So Cloze yet so Far: N400 Amplitude is Better Predicted by Distributional Information than Human Predictability Judgements

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:So Cloze yet so Far: N400 Amplitude is Better Predicted by Distributional Information than Human Predictability Judgements

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators