{"id":{"repo_id":"uiuc","oai_identifier":"oai:www.ideals.illinois.edu:2142/124061"},"canonical_url":"https://search.dev.ndltd.org/etd/uiuc/oai:www.ideals.illinois.edu:2142/124061","repository":{"repo_id":"uiuc","name":"University of Illinois - Urbana-Champaign","base_url":"https://www.ideals.illinois.edu/oai-pmh"},"display":{"title":"A lexical and morphological analysis of the vernacularization of medical vocabulary in Augsburg from 1470–1500: the creation and application of the German Medical Incunabula Corpus","abstract":"Recent developments in Optical Character Recognition (OCR) software make it possible to produce accurate, semi-automated digital transliterations of the first European prints created with Gutenberg’s 15th-century movable type printing press. Prior to these developments, the digital recognition of these first prints or incunabula is challenging due to the difficulty in training models using inconsistent and complex typesets. This research documents the employment of this new OCR technology in the creation of a resource of 15th-century medical texts printed in Augsburg called the German Medical Incunabula Corpus (GeMedIC). The study first introduces key prior research and the socio-historical background information necessary in preparation for various linguistic studies. It then explains the process of GeMedIC’s digital creation and the requirements for its compilation. Thereafter, the raw text corpus is applied to multiple questions in historical linguistics. The first section of the analysis comprises a lexicological study in which a glossary of the key nouns used in GeMedIC is created and explored further using corpus linguistic methods. The next portion of the study focuses on the word formation processes of the key vernacular terminology in the corpus unattested prior to the Middle High German period. This section then discusses the morphological features of these unattested nominal compounds and derivations in GeMedIC along with their inflectional phenomena. The research concludes by exploring textual variation and multilingualism within the corpus in various sections. The first section consists of a glossary of key nouns falling into the categories of foreign words, loan words, loan translations, and loan renderings. This is followed by assorted studies measuring the amount of Latin in GeMedIC and then delves into the possible motivation behind in-text translations, intertextuality, and code-switching within the corpus. The closing chapter emphasizes the continued use of GeMedIC to not only answer scientific questions in linguistics, but also those in history and historical medicine. This research ultimately documents the creation of a specialized corpus of German medical jargon and explores its lexicon in a time in which Latin remains the lingua franca for the genre.","abstract_html":"Recent developments in Optical Character Recognition (OCR) software make it possible to produce accurate, semi-automated digital transliterations of the first European prints created with Gutenberg’s 15th-century movable type printing press. Prior to these developments, the digital recognition of these first prints or incunabula is challenging due to the difficulty in training models using inconsistent and complex typesets. This research documents the employment of this new OCR technology in the creation of a resource of 15th-century medical texts printed in Augsburg called the German Medical Incunabula Corpus (GeMedIC). The study first introduces key prior research and the socio-historical background information necessary in preparation for various linguistic studies. It then explains the process of GeMedIC’s digital creation and the requirements for its compilation. Thereafter, the raw text corpus is applied to multiple questions in historical linguistics. The first section of the analysis comprises a lexicological study in which a glossary of the key nouns used in GeMedIC is created and explored further using corpus linguistic methods. The next portion of the study focuses on the word formation processes of the key vernacular terminology in the corpus unattested prior to the Middle High German period. This section then discusses the morphological features of these unattested nominal compounds and derivations in GeMedIC along with their inflectional phenomena. The research concludes by exploring textual variation and multilingualism within the corpus in various sections. The first section consists of a glossary of key nouns falling into the categories of foreign words, loan words, loan translations, and loan renderings. This is followed by assorted studies measuring the amount of Latin in GeMedIC and then delves into the possible motivation behind in-text translations, intertextuality, and code-switching within the corpus. The closing chapter emphasizes the continued use of GeMedIC to not only answer scientific questions in linguistics, but also those in history and historical medicine. This research ultimately documents the creation of a specialized corpus of German medical jargon and explores its lexicon in a time in which Latin remains the lingua franca for the genre.","abstract_has_math":false,"creators":["Robins, Jenny"],"institution":"Ludwig-Maximilians-Universität München","degree_name":"Ph.D. (doctoral)","degree_level":"Dissertation","degree_discipline":null,"degree_department":null,"school":null,"contributors":["Schallert, Oliver","Habermann, Mechthild"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2023,"date_issued":"2023-02-22","date_published":"2023-02-22","updated_at":"2026-07-22T22:25:00Z","subjects":["Historical Sociolinguistics","Code Switching","Linguistics","Optical Character Recognition (OCR)","Incunabula","Medicine"],"languages":["eng","deu"],"rights":["Copyright 2023 by Jenny Robins"],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"https://hdl.handle.net/2142/124061","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Schallert, Oliver","Habermann, Mechthild"]},{"key":"dc:creator","label":"Author","values":["Robins, Jenny"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2023-02-22","2024-08-29T11:54:17-05:00"]},{"key":"dc:relation","label":"Dc Relation","values":["https://doi.org/10.5282/edoc.33612"]},{"key":"dc:type","label":"Dc Type","values":["Thesis"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Dissertation"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Ph.D. (doctoral)"]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["Ludwig-Maximilians-Universität München"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Historical Sociolinguistics","Code Switching","Linguistics","Optical Character Recognition (OCR)","Incunabula","Medicine"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["eng","deu"]},{"key":"dc:rights","label":"Dc Rights","values":["Copyright 2023 by Jenny Robins"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["https://hdl.handle.net/2142/124061"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["Recent developments in Optical Character Recognition (OCR) software make it possible to produce accurate, semi-automated digital transliterations of the first European prints created with Gutenberg’s 15th-century movable type printing press. Prior to these developments, the digital recognition of these first prints or incunabula is challenging due to the difficulty in training models using inconsistent and complex typesets. This research documents the employment of this new OCR technology in the creation of a resource of 15th-century medical texts printed in Augsburg called the German Medical Incunabula Corpus (GeMedIC). The study first introduces key prior research and the socio-historical background information necessary in preparation for various linguistic studies. It then explains the process of GeMedIC’s digital creation and the requirements for its compilation. Thereafter, the raw text corpus is applied to multiple questions in historical linguistics. The first section of the analysis comprises a lexicological study in which a glossary of the key nouns used in GeMedIC is created and explored further using corpus linguistic methods. The next portion of the study focuses on the word formation processes of the key vernacular terminology in the corpus unattested prior to the Middle High German period. This section then discusses the morphological features of these unattested nominal compounds and derivations in GeMedIC along with their inflectional phenomena. The research concludes by exploring textual variation and multilingualism within the corpus in various sections. The first section consists of a glossary of key nouns falling into the categories of foreign words, loan words, loan translations, and loan renderings. This is followed by assorted studies measuring the amount of Latin in GeMedIC and then delves into the possible motivation behind in-text translations, intertextuality, and code-switching within the corpus. The closing chapter emphasizes the continued use of GeMedIC to not only answer scientific questions in linguistics, but also those in history and historical medicine. This research ultimately documents the creation of a specialized corpus of German medical jargon and explores its lexicon in a time in which Latin remains the lingua franca for the genre."]},{"key":"dc:title","label":"Title","values":["A lexical and morphological analysis of the vernacularization of medical vocabulary in Augsburg from 1470–1500: the creation and application of the German Medical Incunabula Corpus"]}]}],"canonical_facts":{"dc:contributor":["Schallert, Oliver","Habermann, Mechthild"],"dc:creator":["Robins, Jenny"],"dc:date":["2023-02-22","2024-08-29T11:54:17-05:00"],"dc:description":["Recent developments in Optical Character Recognition (OCR) software make it possible to produce accurate, semi-automated digital transliterations of the first European prints created with Gutenberg’s 15th-century movable type printing press. Prior to these developments, the digital recognition of these first prints or incunabula is challenging due to the difficulty in training models using inconsistent and complex typesets. This research documents the employment of this new OCR technology in the creation of a resource of 15th-century medical texts printed in Augsburg called the German Medical Incunabula Corpus (GeMedIC). The study first introduces key prior research and the socio-historical background information necessary in preparation for various linguistic studies. It then explains the process of GeMedIC’s digital creation and the requirements for its compilation. Thereafter, the raw text corpus is applied to multiple questions in historical linguistics. The first section of the analysis comprises a lexicological study in which a glossary of the key nouns used in GeMedIC is created and explored further using corpus linguistic methods. The next portion of the study focuses on the word formation processes of the key vernacular terminology in the corpus unattested prior to the Middle High German period. This section then discusses the morphological features of these unattested nominal compounds and derivations in GeMedIC along with their inflectional phenomena. The research concludes by exploring textual variation and multilingualism within the corpus in various sections. The first section consists of a glossary of key nouns falling into the categories of foreign words, loan words, loan translations, and loan renderings. This is followed by assorted studies measuring the amount of Latin in GeMedIC and then delves into the possible motivation behind in-text translations, intertextuality, and code-switching within the corpus. The closing chapter emphasizes the continued use of GeMedIC to not only answer scientific questions in linguistics, but also those in history and historical medicine. This research ultimately documents the creation of a specialized corpus of German medical jargon and explores its lexicon in a time in which Latin remains the lingua franca for the genre."],"dc:identifier":["https://hdl.handle.net/2142/124061"],"dc:language":["eng","deu"],"dc:relation":["https://doi.org/10.5282/edoc.33612"],"dc:rights":["Copyright 2023 by Jenny Robins"],"dc:subject":["Historical Sociolinguistics","Code Switching","Linguistics","Optical Character Recognition (OCR)","Incunabula","Medicine"],"dc:title":["A lexical and morphological analysis of the vernacularization of medical vocabulary in Augsburg from 1470–1500: the creation and application of the German Medical Incunabula Corpus"],"dc:type":["Thesis"],"thesis:degree_level":["Dissertation"],"thesis:degree_name":["Ph.D. (doctoral)"],"thesis:institution_name":["Ludwig-Maximilians-Universität München"]},"updated_at":"2026-07-22T22:25:00Z"}