Discourse-Level Language Understanding with Deep Learning

Iyyer, Mohit Nagaraja

Discourse-Level Language Understanding with Deep Learning

dc.contributor.advisor	Boyd-Graber, Jordan	en_US
dc.contributor.advisor	Daumé, Hal	en_US
dc.contributor.author	Iyyer, Mohit Nagaraja	en_US
dc.contributor.department	Computer Science	en_US
dc.contributor.publisher	Digital Repository at the University of Maryland	en_US
dc.contributor.publisher	University of Maryland (College Park, Md.)	en_US
dc.date.accessioned	2017-10-14T05:30:16Z
dc.date.available	2017-10-14T05:30:16Z
dc.date.issued	2017	en_US
dc.description.abstract	Designing computational models that can understand language at a human level is a foundational goal in the field of natural language processing (NLP). Given a sentence, machines are capable of translating it into many different languages, generating a corresponding syntactic parse tree, marking words that refer to people or places, and much more. These tasks are solved by statistical machine learning algorithms, which leverage patterns in large datasets to build predictive models. Many recent advances in NLP are due to deep learning models (parameterized as neural networks), which bypass user-specified features in favor of building representations of language directly from the text. Despite many deep learning-fueled advances at the word and sentence level, however, computers still struggle to understand high-level discourse structure in language, or the way in which authors combine and order different units of text (e.g., sentences, paragraphs, chapters) to express a coherent message or narrative. Part of the reason is data-related, as there are no existing datasets for many contextual language-based problems, and some tasks are too complex to be framed as supervised learning problems; for the latter type, we must either resort to unsupervised learning or devise training objectives that simulate the supervised setting. Another reason is architectural: neural networks designed for sentence-level tasks require additional functionality, interpretability, and efficiency to operate at the discourse level. In this thesis, I design deep learning architectures for three NLP tasks that require integrating information across high-level linguistic context: question answering, fictional relationship understanding, and comic book narrative modeling. While these tasks are very different from each other on the surface, I show that similar neural network modules can be used in each case to form contextual representations.	en_US
dc.identifier	https://doi.org/10.13016/M2930NW6W
dc.identifier.uri	http://hdl.handle.net/1903/20159
dc.language.iso	en	en_US
dc.subject.pqcontrolled	Computer science	en_US
dc.subject.pquncontrolled	artificial intelligence	en_US
dc.subject.pquncontrolled	creative language	en_US
dc.subject.pquncontrolled	deep learning	en_US
dc.subject.pquncontrolled	machine learning	en_US
dc.subject.pquncontrolled	natural language processing	en_US
dc.subject.pquncontrolled	question answering	en_US
dc.title	Discourse-Level Language Understanding with Deep Learning	en_US
dc.type	Dissertation	en_US

Files

Original bundle

Now showing 1 - 1 of 1

Name:: Iyyer_umd_0117E_18370.pdf
Size:: 22.74 MB
Format:: Adobe Portable Document Format

Download

Collections

UMD Theses and Dissertations
Computer Science Theses and Dissertations