Ecologically Valid Explanations for Label Variation in NLI

de Marneffe, Marie-Catherine; Jiang, Nan-Jiang; Tan, Chenhao

Ecologically Valid Explanations for Label Variation in NLI

Authors: Marie-Catherine de Marneffe
Nan-Jiang Jiang
Chenhao Tan
Publication date: 20 October 2023
Publisher

Abstract

Human label variation, or annotation disagreement, exists in many natural language processing (NLP) tasks, including natural language inference (NLI). To gain direct evidence of how NLI label variation arises, we build LiveNLI, an English dataset of 1,415 ecologically valid explanations (annotators explain the NLI labels they chose) for 122 MNLI items (at least 10 explanations per item). The LiveNLI explanations confirm that people can systematically vary on their interpretation and highlight within-label variation: annotators sometimes choose the same label for different reasons. This suggests that explanations are crucial for navigating label interpretations in general. We few-shot prompt large language models to generate explanations but the results are inconsistent: they sometimes produces valid and informative explanations, but it also generates implausible ones that do not support the label, highlighting directions for improvement.Comment: Findings at EMNLP 2023. Overlap with previous version arXiv:2304.1244

Similar works

Full text

Available Versions

arXiv.org e-Print Archive

oai:arXiv.org:2310.13850

Last time updated on 16/01/2024