{"id":{"repo_id":"arkansas","oai_identifier":"oai:scholarworks.uark.edu:etd-5153"},"canonical_url":"https://search.dev.ndltd.org/etd/arkansas/oai:scholarworks.uark.edu:etd-5153","repository":{"repo_id":"arkansas","name":"University of Arkansas","base_url":"https://scholarworks.uark.edu/do/oai/"},"display":{"title":"Shakespeare in the Eighteenth Century: Algorithm for Quotation Identification","abstract":"<p>Quoting a borrowed excerpt of text within another literary work was infrequently done prior to the beginning of the eighteenth century. However, quoting other texts, particularly Shakespeare, became quite common after that. Our work develops automatic approaches to identify that trend. Initial work focuses on identifying exact and modified sections of texts taken from works of Shakespeare in novels spanning the eighteenth century. We then introduce a novel approach to identifying modified quotes by adapting the Edit Distance metric, which is character based, to a word based approach. This paper offers an introduction to previous uses of this metric within a multitude of fields, describes the implementation of the different methodologies used for quote identification and then shows how a combination of both Edit Distance methods can help achieve a higher accuracy in quote identification than any one method implemented alone with an overall increase of 10%: from 0.638 and 0.609 to 0.737. Although we demonstrate our approach using Shakespeare quotes in eighteenth century novels, the techniques can be generalized to locate exact and/or partial matches between any set of text targets in any corpus. This work would be of value to literary scholars who want to track quotations over time and could also be applied to other languages.</p>","abstract_html":"&lt;p&gt;Quoting a borrowed excerpt of text within another literary work was infrequently done prior to the beginning of the eighteenth century. However, quoting other texts, particularly Shakespeare, became quite common after that. Our work develops automatic approaches to identify that trend. Initial work focuses on identifying exact and modified sections of texts taken from works of Shakespeare in novels spanning the eighteenth century. We then introduce a novel approach to identifying modified quotes by adapting the Edit Distance metric, which is character based, to a word based approach. This paper offers an introduction to previous uses of this metric within a multitude of fields, describes the implementation of the different methodologies used for quote identification and then shows how a combination of both Edit Distance methods can help achieve a higher accuracy in quote identification than any one method implemented alone with an overall increase of 10%: from 0.638 and 0.609 to 0.737. Although we demonstrate our approach using Shakespeare quotes in eighteenth century novels, the techniques can be generalized to locate exact and/or partial matches between any set of text targets in any corpus. This work would be of value to literary scholars who want to track quotations over time and could also be applied to other languages.&lt;/p&gt;","abstract_has_math":false,"creators":["Chiariglione, Marion Pauline"],"institution":null,"degree_name":"Master of Science in Computer Science (MS)","degree_level":"Thesis","degree_discipline":null,"degree_department":null,"school":null,"contributors":["Li, Qinghua","Luu, Khoa"],"advisors":["Gauch, Susan E."],"committee_chairs":[],"committee_members":[],"year":2020,"date_issued":"2020-05-01T07:00:00Z","date_published":"2020-05-01T07:00:00Z","updated_at":"2026-07-24T00:59:15Z","subjects":["Edit Distance","Information Retrieval","Quote Identification","String matching","Numerical Analysis and Scientific Computing","Theory and Algorithms"],"languages":[],"rights":[],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"https://scholarworks.uark.edu/etd/3580","outbound_label":"Repository record","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Li, Qinghua","Luu, Khoa"]},{"key":"dc:contributor.advisor","label":"Advisor","values":["Gauch, Susan E."]},{"key":"dc:creator","label":"Author","values":["Chiariglione, Marion Pauline"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2020"]},{"key":"dc:date.available","label":"Dc Date Available","values":["2020-06-11T07:00:00Z"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Thesis"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Master of Science in Computer Science (MS)"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Edit Distance","Information Retrieval","Quote Identification","String matching","Numerical Analysis and Scientific Computing","Theory and Algorithms"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["https://scholarworks.uark.edu/etd/3580"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["<p>Quoting a borrowed excerpt of text within another literary work was infrequently done prior to the beginning of the eighteenth century. However, quoting other texts, particularly Shakespeare, became quite common after that. Our work develops automatic approaches to identify that trend. Initial work focuses on identifying exact and modified sections of texts taken from works of Shakespeare in novels spanning the eighteenth century. We then introduce a novel approach to identifying modified quotes by adapting the Edit Distance metric, which is character based, to a word based approach. This paper offers an introduction to previous uses of this metric within a multitude of fields, describes the implementation of the different methodologies used for quote identification and then shows how a combination of both Edit Distance methods can help achieve a higher accuracy in quote identification than any one method implemented alone with an overall increase of 10%: from 0.638 and 0.609 to 0.737. Although we demonstrate our approach using Shakespeare quotes in eighteenth century novels, the techniques can be generalized to locate exact and/or partial matches between any set of text targets in any corpus. This work would be of value to literary scholars who want to track quotations over time and could also be applied to other languages.</p>"]},{"key":"dc:title","label":"Title","values":["Shakespeare in the Eighteenth Century: Algorithm for Quotation Identification"]}]}],"canonical_facts":{"dc:contributor":["Li, Qinghua","Luu, Khoa"],"dc:contributor.advisor":["Gauch, Susan E."],"dc:creator":["Chiariglione, Marion Pauline"],"dc:date":["2020"],"dc:date.available":["2020-06-11T07:00:00Z"],"dc:description.abstract":["<p>Quoting a borrowed excerpt of text within another literary work was infrequently done prior to the beginning of the eighteenth century. However, quoting other texts, particularly Shakespeare, became quite common after that. Our work develops automatic approaches to identify that trend. Initial work focuses on identifying exact and modified sections of texts taken from works of Shakespeare in novels spanning the eighteenth century. We then introduce a novel approach to identifying modified quotes by adapting the Edit Distance metric, which is character based, to a word based approach. This paper offers an introduction to previous uses of this metric within a multitude of fields, describes the implementation of the different methodologies used for quote identification and then shows how a combination of both Edit Distance methods can help achieve a higher accuracy in quote identification than any one method implemented alone with an overall increase of 10%: from 0.638 and 0.609 to 0.737. Although we demonstrate our approach using Shakespeare quotes in eighteenth century novels, the techniques can be generalized to locate exact and/or partial matches between any set of text targets in any corpus. This work would be of value to literary scholars who want to track quotations over time and could also be applied to other languages.</p>"],"dc:identifier":["https://scholarworks.uark.edu/etd/3580"],"dc:subject":["Edit Distance","Information Retrieval","Quote Identification","String matching","Numerical Analysis and Scientific Computing","Theory and Algorithms"],"dc:title":["Shakespeare in the Eighteenth Century: Algorithm for Quotation Identification"],"thesis:degree_level":["Thesis"],"thesis:degree_name":["Master of Science in Computer Science (MS)"]},"updated_at":"2026-07-24T00:59:15Z"}