Curing the SICK and Other NLI Maladies

Aikaterini-Lida Kalouli; Hai Hu; Alexander F. Webb; Lawrence S. Moss; Valeria de Paiva

doi:10.1162/coli_a_00465

Computational Linguistics (Mar 2023)

Curing the SICK and Other NLI Maladies

Aikaterini-Lida Kalouli,
Hai Hu,
Alexander F. Webb,
Lawrence S. Moss,
Valeria de Paiva

Affiliations

Aikaterini-Lida Kalouli
Hai Hu
Alexander F. Webb
Lawrence S. Moss
Valeria de Paiva

DOI: https://doi.org/10.1162/coli_a_00465
Journal volume & issue: Vol. 49, no. 1

Abstract

Read online

Against the backdrop of the ever-improving Natural Language Inference (NLI) models, recent efforts have focused on the suitability of the current NLI datasets and on the feasibility of the NLI task as it is currently approached. Many of the recent studies have exposed the inherent human disagreements of the inference task and have proposed a shift from categorical labels to human subjective probability assessments, capturing human uncertainty. In this work, we show how neither the current task formulation nor the proposed uncertainty gradient are entirely suitable for solving the NLI challenges. Instead, we propose an ordered sense space annotation, which distinguishes between logical and common-sense inference. One end of the space captures non-sensical inferences, while the other end represents strictly logical scenarios. In the middle of the space, we find a continuum of common-sense, namely, the subjective and graded opinion of a “person on the street.” To arrive at the proposed annotation scheme, we perform a careful investigation of the SICK corpus and we create a taxonomy of annotation issues and guidelines. We re-annotate the corpus with the proposed annotation scheme, utilizing four symbolic inference systems, and then perform a thorough evaluation of the scheme by fine-tuning and testing commonly used pre-trained language models on the re-annotated SICK within various settings. We also pioneer a crowd annotation of a small portion of the MultiNLI corpus, showcasing that it is possible to adapt our scheme for annotation by non-experts on another NLI corpus. Our work shows the efficiency and benefits of the proposed mechanism and opens the way for a careful NLI task refinement.

Published in Computational Linguistics

ISSN: 0891-2017 (Print); 1530-9312 (Online)
Publisher: The MIT Press
Country of publisher: United States
LCC subjects: Language and Literature: Philology. Linguistics: Computational linguistics. Natural language processing
Website: https://direct.mit.edu/coli

About the journal