DeepAI AI Chat
Log In Sign Up

SherLIiC: A Typed Event-Focused Lexical Inference Benchmark for Evaluating Natural Language Inference

06/04/2019
by   Martin Schmitt, et al.
0

We present SherLIiC, a testbed for lexical inference in context (LIiC), consisting of 3985 manually annotated inference rule candidates (InfCands), accompanied by (i) 960k unlabeled InfCands, and (ii) 190k typed textual relations between Freebase entities extracted from the large entity-linked corpus ClueWeb09. Each InfCand consists of one of these relations, expressed as a lemmatized dependency path, and two argument placeholders, each linked to one or more Freebase types. Due to our candidate selection process based on strong distributional evidence, SherLIiC is much harder than existing testbeds because distributional evidence is of little utility in the classification of InfCands. We also show that, due to its construction, many of SherLIiC's correct InfCands are novel and missing from existing rule bases. We evaluate a number of strong baselines on SherLIiC, ranging from semantic vector space models to state of the art neural models of natural language inference (NLI). We show that SherLIiC poses a tough challenge to existing NLI systems.

READ FULL TEXT

page 1

page 2

page 3

page 4

02/24/2020

Using Distributional Thesaurus Embedding for Co-hyponymy Detection

Discriminating lexical relations among distributionally similar words ha...
02/10/2021

Language Models for Lexical Inference in Context

Lexical inference in context (LIiC) is the task of recognizing textual e...
08/05/2018

Instantiation

In computational linguistics, a large body of work exists on distributed...
10/29/2020

Learning as Abduction: Trainable Natural Logic Theorem Prover for Natural Language Inference

Tackling Natural Language Inference with a logic-based method is becomin...
02/13/2018

Network Features Based Co-hyponymy Detection

Distinguishing lexical relations has been a long term pursuit in natural...
10/30/2018

Stress-Testing Neural Models of Natural Language Inference with Multiply-Quantified Sentences

Standard evaluations of deep learning models for semantics using natural...