LINSPECTOR: Multilingual Probing Tasks for Word Representations-Reference-Cited by-同舟云学术

LINSPECTOR: Multilingual Probing Tasks for Word Representations

Published:2020-06 Issue:2 Volume:46 Page:335-385
ISSN:0891-2017
Container-title:Computational Linguistics
language:en
Short-container-title:Computational Linguistics

Author:

Şahin Gözde Gül¹,Vania Clara²,Kuznetsov Ilia³,Gurevych Iryna³

Affiliation:

1. AIPHES and UKP Lab / TU Darmstadt, Technische Universität Darmstadt, Department of Computer Science.

2. New York University.

3. AIPHES and UKP Lab / TU Darmstadt.

Abstract

Despite an ever-growing number of word representation models introduced for a large number of languages, there is a lack of a standardized technique to provide insights into what is captured by these models. Such insights would help the community to get an estimate of the downstream task performance, as well as to design more informed neural architectures, while avoiding extensive experimentation that requires substantial computational resources not all researchers have access to. A recent development in NLP is to use simple classification tasks, also called probing tasks, that test for a single linguistic feature such as part-of-speech. Existing studies mostly focus on exploring the linguistic information encoded by the continuous representations of English text. However, from a typological perspective the morphologically poor English is rather an outlier: The information encoded by the word order and function words in English is often stored on a subword, morphological level in other languages. To address this, we introduce 15 type-level probing tasks such as case marking, possession, word length, morphological tag count, and pseudoword identification for 24 languages. We present a reusable methodology for creation and evaluation of such tests in a multilingual setting, which is challenging because of a lack of resources, lower quality of tools, and differences among languages. We then present experiments on several diverse multilingual word embedding models, in which we relate the probing task performance for a diverse set of languages to a range of five classic NLP tasks: POS-tagging, dependency parsing, semantic role labeling, named entity recognition, and natural language inference. We find that a number of probing tests have significantly high positive correlation to the downstream tasks, especially for morphologically rich languages. We show that our tests can be used to explore word embeddings or black-box neural models for linguistic cues in a multilingual setting. We release the probing data sets and the evaluation suite LINSPECTOR with https://github.com/UKPLab/linspector .

Publisher

MIT Press - Journals

Subject

Artificial Intelligence,Computer Science Applications,Linguistics and Language,Language and Linguistics

Link

https://www.mitpressjournals.org/doi/pdf/10.1162/coli_a_00376

Reference83 articles.

1. Common pitfalls in statistical analysis: The use of correlation techniques

2. Compositional Representation of Morphologically-Rich Input for Neural Machine Translation

3. What do Neural Machine Translation Models Learn about Morphology?

4. Analysis Methods in Neural Language Processing: A Survey

Cited by 6 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Assessing linguistic generalisation in language models: a dataset for Brazilian Portuguese;Language Resources and Evaluation;2023-06-02

2. Morphosyntactic probing of multilingual BERT models;Natural Language Engineering;2023-05-25

3. On Robustness and Sensitivity of a Neural Language Model: A Case Study on Italian L1 Learner Errors;IEEE/ACM Transactions on Audio, Speech, and Language Processing;2023

4. Cross lingual transfer learning for sentiment analysis of Italian TripAdvisor reviews;Expert Systems with Applications;2022-12

5. Investigating Language Relationships in Multilingual Sentence Encoders Through the Lens of Linguistic Typology;Computational Linguistics;2022