Affiliation:
1. Center for Information and Language Processing (CIS), LMU Munich. kalouli@cis.lmu.de
2. Shanghai Jiao Tong University School of Foreign Languages. hu.hai@sjtu.edu.cn
3. Indiana University Bloomington Department of Philosophy. afwebb@iu.edu
4. Indiana University Bloomington Department of Mathematics. lmoss@indiana.edu
5. Topos Institute. valeria.depaiva@gmail.com
Abstract
AbstractAgainst the backdrop of the ever-improving Natural Language Inference (NLI) models, recent efforts have focused on the suitability of the current NLI datasets and on the feasibility of the NLI task as it is currently approached. Many of the recent studies have exposed the inherent human disagreements of the inference task and have proposed a shift from categorical labels to human subjective probability assessments, capturing human uncertainty. In this work, we show how neither the current task formulation nor the proposed uncertainty gradient are entirely suitable for solving the NLI challenges. Instead, we propose an ordered sense space annotation, which distinguishes between logical and common-sense inference. One end of the space captures non-sensical inferences, while the other end represents strictly logical scenarios. In the middle of the space, we find a continuum of common-sense, namely, the subjective and graded opinion of a “person on the street.” To arrive at the proposed annotation scheme, we perform a careful investigation of the SICK corpus and we create a taxonomy of annotation issues and guidelines. We re-annotate the corpus with the proposed annotation scheme, utilizing four symbolic inference systems, and then perform a thorough evaluation of the scheme by fine-tuning and testing commonly used pre-trained language models on the re-annotated SICK within various settings. We also pioneer a crowd annotation of a small portion of the MultiNLI corpus, showcasing that it is possible to adapt our scheme for annotation by non-experts on another NLI corpus. Our work shows the efficiency and benefits of the proposed mechanism and opens the way for a careful NLI task refinement.
Subject
Artificial Intelligence,Computer Science Applications,Linguistics and Language,Language and Linguistics
Reference76 articles.
1. Towards a wide-coverage tableau method for natural logic;Abzianidze,2014
2. A tableau prover for natural logic and language;Abzianidze,2015
3. Abzianidze, Lasha . 2016. A Natural Proof System for Natural Language. Ph.D. thesis, Tilburg University.
4. LangPro: Natural language theorem prover;Abzianidze,2017
5. The second PASCAL recognising textual entailment challenge;Bar-Haim,2006
Cited by
3 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献