{"id":{"repo_id":"wfu","oai_identifier":"oai:wakespace.lib.wfu.edu:10339/38554"},"canonical_url":"https://search.dev.ndltd.org/etd/wfu/oai:wakespace.lib.wfu.edu:10339/38554","repository":{"repo_id":"wfu","name":"Wake Forest University","base_url":"https://wakespace.lib.wfu.edu/oai/request"},"display":{"title":"Efficient Information Extraction Using Statistical Relational Learning","abstract":"Information extraction has gained significant importance due to the dramatic increase of information stored in the form of natural language text. In this thesis we explore a machine learning-based approach to support a natural language processing (NLP) algorithm, and an application of information extraction. One of the challenges of learning-based approaches is the requirement of human annotated examples. Current successful approaches alleviate this problem by employing some form of distant supervision. In this work, we take a different approach -- we create weakly supervised examples for relations by using commonsense knowledge. The key innovation is that this commonsense knowledge is completely independent of the natural language text. This helps when learning the full model for information extraction as against simply learning the parameters of a known model. We demonstrate on two domains that this form of weak supervision yields superior results when learning structure compared to simply using the gold standard labels. In the second part of this thesis, we consider the problem of Adverse Drug Events (ADEs) discovery. Several methods have been proposed for ADE discovery, exploiting various information sources such as health data, social network data, and scientific literature. We propose a NLP-based method that exploits scientific literature to quantitatively evaluate proposed ADEs. We validate our approach on a common ADE dataset, where we find better agreement than state-of-the-art ADE discovery methods.","abstract_html":"Information extraction has gained significant importance due to the dramatic increase of information stored in the form of natural language text. In this thesis we explore a machine learning-based approach to support a natural language processing (NLP) algorithm, and an application of information extraction. One of the challenges of learning-based approaches is the requirement of human annotated examples. Current successful approaches alleviate this problem by employing some form of distant supervision. In this work, we take a different approach -- we create weakly supervised examples for relations by using commonsense knowledge. The key innovation is that this commonsense knowledge is completely independent of the natural language text. This helps when learning the full model for information extraction as against simply learning the parameters of a known model. We demonstrate on two domains that this form of weak supervision yields superior results when learning structure compared to simply using the gold standard labels. In the second part of this thesis, we consider the problem of Adverse Drug Events (ADEs) discovery. Several methods have been proposed for ADE discovery, exploiting various information sources such as health data, social network data, and scientific literature. We propose a NLP-based method that exploits scientific literature to quantitatively evaluate proposed ADEs. We validate our approach on a common ADE dataset, where we find better agreement than state-of-the-art ADE discovery methods.","abstract_has_math":false,"creators":["Picado Leiva, Jose Manuel"],"institution":"Wake Forest University","degree_name":null,"degree_level":null,"degree_discipline":null,"degree_department":null,"school":null,"contributors":[],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2013,"date_issued":"2013","date_published":"2013","updated_at":"2026-07-27T22:01:33Z","subjects":["information extraction"],"languages":["en"],"rights":[],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"http://hdl.handle.net/10339/38554","outbound_label":"Handle","outbound_source":"dc:identifier.uri"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:creator","label":"Author","values":["Picado Leiva, Jose Manuel"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.accessioned","label":"Dc Date Accessioned","values":["2013-06-06T21:19:34Z"]},{"key":"dc:date.available","label":"Dc Date Available","values":["2013-06-06T21:19:34Z"]},{"key":"dc:date.issued","label":"Date","values":["2013"]},{"key":"dc:publisher","label":"Institution","values":["Wake Forest University"]},{"key":"dc:type","label":"Dc Type","values":["Thesis"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["information extraction"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language.iso","label":"Language (ISO)","values":["en"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.uri","label":"Identifier URI","values":["http://hdl.handle.net/10339/38554"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["Information extraction has gained significant importance due to the dramatic increase of information stored in the form of natural language text. In this thesis we explore a machine learning-based approach to support a natural language processing (NLP) algorithm, and an application of information extraction. One of the challenges of learning-based approaches is the requirement of human annotated examples. Current successful approaches alleviate this problem by employing some form of distant supervision. In this work, we take a different approach -- we create weakly supervised examples for relations by using commonsense knowledge. The key innovation is that this commonsense knowledge is completely independent of the natural language text. This helps when learning the full model for information extraction as against simply learning the parameters of a known model. We demonstrate on two domains that this form of weak supervision yields superior results when learning structure compared to simply using the gold standard labels. In the second part of this thesis, we consider the problem of Adverse Drug Events (ADEs) discovery. Several methods have been proposed for ADE discovery, exploiting various information sources such as health data, social network data, and scientific literature. We propose a NLP-based method that exploits scientific literature to quantitatively evaluate proposed ADEs. We validate our approach on a common ADE dataset, where we find better agreement than state-of-the-art ADE discovery methods."]},{"key":"dc:title","label":"Title","values":["Efficient Information Extraction Using Statistical Relational Learning"]}]}],"canonical_facts":{"dc:creator":["Picado Leiva, Jose Manuel"],"dc:date.accessioned":["2013-06-06T21:19:34Z"],"dc:date.available":["2013-06-06T21:19:34Z"],"dc:date.issued":["2013"],"dc:description.abstract":["Information extraction has gained significant importance due to the dramatic increase of information stored in the form of natural language text. In this thesis we explore a machine learning-based approach to support a natural language processing (NLP) algorithm, and an application of information extraction. One of the challenges of learning-based approaches is the requirement of human annotated examples. Current successful approaches alleviate this problem by employing some form of distant supervision. In this work, we take a different approach -- we create weakly supervised examples for relations by using commonsense knowledge. The key innovation is that this commonsense knowledge is completely independent of the natural language text. This helps when learning the full model for information extraction as against simply learning the parameters of a known model. We demonstrate on two domains that this form of weak supervision yields superior results when learning structure compared to simply using the gold standard labels. In the second part of this thesis, we consider the problem of Adverse Drug Events (ADEs) discovery. Several methods have been proposed for ADE discovery, exploiting various information sources such as health data, social network data, and scientific literature. We propose a NLP-based method that exploits scientific literature to quantitatively evaluate proposed ADEs. We validate our approach on a common ADE dataset, where we find better agreement than state-of-the-art ADE discovery methods."],"dc:identifier.uri":["http://hdl.handle.net/10339/38554"],"dc:language.iso":["en"],"dc:publisher":["Wake Forest University"],"dc:subject":["information extraction"],"dc:title":["Efficient Information Extraction Using Statistical Relational Learning"],"dc:type":["Thesis"]},"updated_at":"2026-07-27T22:01:33Z"}