Understanding models understanding language-Reference-Cited by-同舟云学术

Understanding models understanding language

Published:2022-10-27 Issue:6 Volume:200 Page:
ISSN:1573-0964
Container-title:Synthese
language:en
Short-container-title:Synthese

Author:

Søgaard Anders^ORCID

Abstract

AbstractLandgrebe and Smith (Synthese 198(March):2061–2081, 2021) present an unflattering diagnosis of recent advances in what they call language-centric artificial intelligence—perhaps more widely known as natural language processing: The models that are currently employed do not have sufficient expressivity, will not generalize, and are fundamentally unable to induce linguistic semantics, they say. The diagnosis is mainly derived from an analysis of the widely used Transformer architecture. Here I address a number of misunderstandings in their analysis, and present what I take to be a more adequate analysis of the ability of Transformer models to learn natural language semantics. To avoid confusion, I distinguish between inferential and referential semantics. Landgrebe and Smith (2021)’s analysis of the Transformer architecture’s expressivity and generalization concerns inferential semantics. This part of their diagnosis is shown to rely on misunderstandings of technical properties of Transformers. Landgrebe and Smith (2021) also claim that referential semantics is unobtainable for Transformer models. In response, I present a non-technical discussion of techniques for grounding Transformer models, giving them referential semantics, even in the absence of supervision. I also present a simple thought experiment to highlight the mechanisms that would lead to referential semantics, and discuss in what sense models that are grounded in this way, can be said to understand language. Finally, I discuss the approach Landgrebe and Smith (2021) advocate for, namely manual specification of formal grammars that associate linguistic expressions with logical form.

Publisher

Springer Science and Business Media LLC

Subject

General Social Sciences,Philosophy

Link

https://link.springer.com/content/pdf/10.1007/s11229-022-03931-4.pdf

Reference57 articles.

1. Abdou, M., Kulmizev, A., Hershcovich, D., Frank, S., Pavlick, E., & Søgaard, A. (2021). Can language models encode perceptual structure without grounding? A case study in color. In: Proceedings of the 25th conference on computational natural language learning (pp. 109–132), Online. Association for Computational Linguistics.

2. Aldarmaki, H., Mohan, M., & Diab, M. (2018). Unsupervised word mapping using structural similarities in monolingual embeddings. Transactions of the Association for Computational Linguistics, 6, 185–196.

3. Artetxe, M., Labaka, G., & Agirre, E. (2017). Learning bilingual word embeddings with (almost) no bilingual data. In: Proceedings of the 55th annual meeting of the Association for Computational Linguistics (Vol. 1: Long Papers) (pp. 451–462), Vancouver, Canada. Association for Computational Linguistics.

4. Babu, A., Shrivastava, A., Aghajanyan, A., Aly, A., Fan, A., & Ghazvininejad, M. (2021). Non-autoregressive semantic parsing for compositional task-oriented dialog. In: Proceedings of the 2021 conference of the North American chapter of the Association for Computational Linguistics: Human language technologies (pp. 2969–2978), Online. Association for Computational Linguistics.

5. Bender, E. M. & Koller, A. (2020). Climbing towards NLU: On meaning, form, and understanding in the age of data. In: Proceedings of the 58th annual meeting of the Association for Computational Linguistics (pp. 5185–5198), Online. Association for Computational Linguistics.

Cited by 4 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. The Simulative Role of Neural Language Models in Brain Language Processing;Philosophies;2024-08-29

2. Contrasting Linguistic Patterns in Human and LLM-Generated News Text;Artificial Intelligence Review;2024-08-23

3. Does Artificial Intelligence Speak Our Language?: A Gadamerian Assessment of Generative Language Models;Political Research Quarterly;2024-03-28

4. Big Data and (the New?) Reality;American, British and Canadian Studies;2023-12-01