Brains and algorithms partially converge in natural language processing-Reference-Cited by-同舟云学术

Brains and algorithms partially converge in natural language processing

Published:2022-02-16 Issue:1 Volume:5 Page:
ISSN:2399-3642
Container-title:Communications Biology
language:en
Short-container-title:Commun Biol

Author:

Caucheteux Charlotte,King Jean-Rémi^ORCID

Abstract

AbstractDeep learning algorithms trained to predict masked words from large amount of text have recently been shown to generate activations similar to those of the human brain. However, what drives this similarity remains currently unknown. Here, we systematically compare a variety of deep language models to identify the computational principles that lead them to generate brain-like representations of sentences. Specifically, we analyze the brain responses to 400 isolated sentences in a large cohort of 102 subjects, each recorded for two hours with functional magnetic resonance imaging (fMRI) and magnetoencephalography (MEG). We then test where and when each of these algorithms maps onto the brain responses. Finally, we estimate how the architecture, training, and performance of these models independently account for the generation of brain-like representations. Our analyses reveal two main findings. First, the similarity between the algorithms and the brain primarily depends on their ability to predict words from context. Second, this similarity reveals the rise and maintenance of perceptual, lexical, and compositional representations within each cortical region. Overall, this study shows that modern language algorithms partially converge towards brain-like solutions, and thus delineates a promising path to unravel the foundations of natural language processing.

Publisher

Springer Science and Business Media LLC

Subject

General Agricultural and Biological Sciences,General Biochemistry, Genetics and Molecular Biology,Medicine (miscellaneous)

Link

https://www.nature.com/articles/s42003-022-03036-1.pdf

Reference95 articles.

1. Turing, A. M. Parsing the Turing Test 23–65 (Springer, 2009).

2. Chomsky, N. Language and Mind (Cambridge University Press, 2006).

3. Dehaene, S., Yann, L. & Girardon, J. La plus belle histoire de l’intelligence: des origines aux neurones artificiels: vers une nouvelle étape de l’évolution (Robert Laffont, 2018).

4. Vaswani, A. et al. Attention is all you need. In Proceedings on NIPS (Cornell University, 2017).

5. Devlin, J., Chang, M., Lee, K. & Toutanova, K. BERT: pre-training of deep bidirectional transformers for language understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) (2019).

Cited by 76 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Is word order considered by foundation models? A comparative task-oriented analysis;Expert Systems with Applications;2024-05

2. Sentence-level embeddings reveal dissociable word- and sentence-level cortical representation across coarse- and fine-grained levels of meaning;Brain and Language;2024-03

3. Emergence of syntax and word prediction in an artificial neural circuit of the cerebellum;Nature Communications;2024-01-31

4. Do Topographic Deep ANN Models of the Primate Ventral Stream Predict the Perceptual Effects of Direct IT Cortical Interventions?;2024-01-09

5. Driving and suppressing the human language network using large language models;Nature Human Behaviour;2024-01-03