University of Illinois at Urbana-Champaign
Textual entailment from image caption denotations
Abstract
dc:descriptionUnderstanding the meaning of linguistic expressions is a fundamental task of natural language processing. While distributed representations have become a powerful technique for modeling lexical semantics, but they have traditionally relied on ungrounded text corpora to identify semantically similar words. In contrast, this thesis explicitly models the denotation of linguistic expressions by building representations from grounded image captions. This allows us to use descriptions of the world to learn connections that would be difficult to identify in text-based corpora. In particular, we explore novel approaches to entailment that capture everyday world knowledge missing from other NLP tasks, on both existing datasets and our own new dataset. We also present a novel embedding model that produces phrase representations that are informed by our grounded representation. We conclude with an analysis of how grounded embeddings differ from standard distributional embeddings and suggestions for future refinement of this approach.
Degree
thesis:*- Name thesis:degree_name
- Ph.D.
- Level thesis:degree_level
- Dissertation
- Discipline thesis:degree_discipline
- Computer Science
- Grantor
- University of Illinois at Urbana-Champaign
- Year dc:date
- 2019
Author and committee
dc:creator, dc:contributor.*- Author dc:creator
-
- Lai, Alice Yingming
- Contributors dc:contributor
-
- Hockenmaier, Julia
- Erk, Katrin
- Roth, Dan
- Zhai, ChengXiang
Subjects
dc:subject × 3Rights
dc:rights- Statement dc:rights
-
- Copyright 2018 Alice Lai
- Language dc:language
- en
Identifiers
dc:identifier.*- Handle dc:identifier
- http://hdl.handle.net/2142/102460
- OAI identifier oai:identifier
- oai:www.ideals.illinois.edu:2142/102460