{"id":{"repo_id":"arizona-thes","oai_identifier":"oai:repository.arizona.edu:10150/648598"},"canonical_url":"https://search.dev.ndltd.org/etd/arizona-thes/oai:repository.arizona.edu:10150/648598","repository":{"repo_id":"arizona-thes","name":"University of Arizona","base_url":"https://repository.arizona.edu/oai/request"},"display":{"title":"The Lexicon is Shaped for Incremental Processing in a Noisy Channel","abstract":"Human language is a constantly evolving system, with properties of a language’s grammar being shaped, among other things, to benefit efficient communication (Zipf, 1949; Köhler, 1987; Gibson et al., 2019). One important area of a language’s grammar is its lexicon, or set of words and their corresponding phonological forms, and there is a great deal of evidence that the lexicons of the world’s languages are structured to be similar to abstract, maximally efficient communicative codes (e.g., Zipf 1935; Ferrer-i Cancho and Solé 2003; Piantadosi et al. 2009, 2011; Mahowald et al. 2018. In this dissertation, I will present additional evidence that the lexicons of natural languages are structured for efficiency, moving past abstract codes and focusing on how listeners process and identify words in speech. Primarily, I will show that the distribution of inter-lexical contrasts, i.e., phonemes, in a language is such that the average Shannon information (Shannon, 1948) of phonemic contrasts is greater than would be expected otherwise, using a typologically and geographically diverse dataset of 25 languages. In addition, I will show that the increased informativeness of lexical contrasts does not interfere with a lexicon’s potential for accurate communication. Together, these offer strong support that languages evolve to be efficient communication systems, tailored to human users.","abstract_html":"Human language is a constantly evolving system, with properties of a language’s grammar being shaped, among other things, to benefit efficient communication (Zipf, 1949; Köhler, 1987; Gibson et al., 2019). One important area of a language’s grammar is its lexicon, or set of words and their corresponding phonological forms, and there is a great deal of evidence that the lexicons of the world’s languages are structured to be similar to abstract, maximally efficient communicative codes (e.g., Zipf 1935; Ferrer-i Cancho and Solé 2003; Piantadosi et al. 2009, 2011; Mahowald et al. 2018. In this dissertation, I will present additional evidence that the lexicons of natural languages are structured for efficiency, moving past abstract codes and focusing on how listeners process and identify words in speech. Primarily, I will show that the distribution of inter-lexical contrasts, i.e., phonemes, in a language is such that the average Shannon information (Shannon, 1948) of phonemic contrasts is greater than would be expected otherwise, using a typologically and geographically diverse dataset of 25 languages. In addition, I will show that the increased informativeness of lexical contrasts does not interfere with a lexicon’s potential for accurate communication. Together, these offer strong support that languages evolve to be efficient communication systems, tailored to human users.","abstract_has_math":false,"creators":["King, Adam"],"institution":"The University of Arizona.","degree_name":"Ph.D.","degree_level":"doctoral","degree_discipline":"Graduate College","degree_department":null,"school":null,"contributors":[],"advisors":["Wedel, Andrew"],"committee_chairs":[],"committee_members":["Hammond, Michael","Fezechkina, Maryia"],"year":2020,"date_issued":"2020","date_published":"2020","updated_at":"2026-07-24T00:56:22Z","subjects":["corpus linguistics","efficient communication","incremental processing","language evolution","word processing","Zipf's law of abbreviation"],"languages":["en"],"rights":["Copyright © is held by the author. Digital access to this material is made possible by the University Libraries, University of Arizona. Further transmission, reproduction, presentation (such as public display or performance) of protected items is prohibited except with permission of the author."],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"http://hdl.handle.net/10150/648598","outbound_label":"Handle","outbound_source":"dc:identifier.uri"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor.advisor","label":"Advisor","values":["Wedel, Andrew"]},{"key":"dc:contributor.committeemember","label":"Committee Member","values":["Hammond, Michael","Fezechkina, Maryia"]},{"key":"dc:creator","label":"Author","values":["King, Adam"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.accessioned","label":"Dc Date Accessioned","values":["2020-11-26T02:19:42Z"]},{"key":"dc:date.available","label":"Dc Date Available","values":["2020-11-26T02:19:42Z"]},{"key":"dc:date.issued","label":"Date","values":["2020"]},{"key":"dc:publisher","label":"Institution","values":["The University of Arizona."]},{"key":"dc:type","label":"Dc Type","values":["text","Electronic Dissertation"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Graduate College","Linguistics"]},{"key":"thesis:degree_level","label":"Degree Level","values":["doctoral"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Ph.D."]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Arizona"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["corpus linguistics","efficient communication","incremental processing","language evolution","word processing","Zipf's law of abbreviation"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language.iso","label":"Language (ISO)","values":["en"]},{"key":"dc:rights","label":"Dc Rights","values":["Copyright © is held by the author. Digital access to this material is made possible by the University Libraries, University of Arizona. Further transmission, reproduction, presentation (such as public display or performance) of protected items is prohibited except with permission of the author."]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.uri","label":"Identifier URI","values":["http://hdl.handle.net/10150/648598"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["Human language is a constantly evolving system, with properties of a language’s grammar being shaped, among other things, to benefit efficient communication (Zipf, 1949; Köhler, 1987; Gibson et al., 2019). One important area of a language’s grammar is its lexicon, or set of words and their corresponding phonological forms, and there is a great deal of evidence that the lexicons of the world’s languages are structured to be similar to abstract, maximally efficient communicative codes (e.g., Zipf 1935; Ferrer-i Cancho and Solé 2003; Piantadosi et al. 2009, 2011; Mahowald et al. 2018. In this dissertation, I will present additional evidence that the lexicons of natural languages are structured for efficiency, moving past abstract codes and focusing on how listeners process and identify words in speech. Primarily, I will show that the distribution of inter-lexical contrasts, i.e., phonemes, in a language is such that the average Shannon information (Shannon, 1948) of phonemic contrasts is greater than would be expected otherwise, using a typologically and geographically diverse dataset of 25 languages. In addition, I will show that the increased informativeness of lexical contrasts does not interfere with a lexicon’s potential for accurate communication. Together, these offer strong support that languages evolve to be efficient communication systems, tailored to human users."]},{"key":"dc:title","label":"Title","values":["The Lexicon is Shaped for Incremental Processing in a Noisy Channel"]}]}],"canonical_facts":{"dc:contributor.advisor":["Wedel, Andrew"],"dc:contributor.committeemember":["Hammond, Michael","Fezechkina, Maryia"],"dc:creator":["King, Adam"],"dc:date.accessioned":["2020-11-26T02:19:42Z"],"dc:date.available":["2020-11-26T02:19:42Z"],"dc:date.issued":["2020"],"dc:description.abstract":["Human language is a constantly evolving system, with properties of a language’s grammar being shaped, among other things, to benefit efficient communication (Zipf, 1949; Köhler, 1987; Gibson et al., 2019). One important area of a language’s grammar is its lexicon, or set of words and their corresponding phonological forms, and there is a great deal of evidence that the lexicons of the world’s languages are structured to be similar to abstract, maximally efficient communicative codes (e.g., Zipf 1935; Ferrer-i Cancho and Solé 2003; Piantadosi et al. 2009, 2011; Mahowald et al. 2018. In this dissertation, I will present additional evidence that the lexicons of natural languages are structured for efficiency, moving past abstract codes and focusing on how listeners process and identify words in speech. Primarily, I will show that the distribution of inter-lexical contrasts, i.e., phonemes, in a language is such that the average Shannon information (Shannon, 1948) of phonemic contrasts is greater than would be expected otherwise, using a typologically and geographically diverse dataset of 25 languages. In addition, I will show that the increased informativeness of lexical contrasts does not interfere with a lexicon’s potential for accurate communication. Together, these offer strong support that languages evolve to be efficient communication systems, tailored to human users."],"dc:identifier.uri":["http://hdl.handle.net/10150/648598"],"dc:language.iso":["en"],"dc:publisher":["The University of Arizona."],"dc:rights":["Copyright © is held by the author. Digital access to this material is made possible by the University Libraries, University of Arizona. Further transmission, reproduction, presentation (such as public display or performance) of protected items is prohibited except with permission of the author."],"dc:subject":["corpus linguistics","efficient communication","incremental processing","language evolution","word processing","Zipf's law of abbreviation"],"dc:title":["The Lexicon is Shaped for Incremental Processing in a Noisy Channel"],"dc:type":["text","Electronic Dissertation"],"thesis:degree_discipline":["Graduate College","Linguistics"],"thesis:degree_level":["doctoral"],"thesis:degree_name":["Ph.D."],"thesis:institution_name":["University of Arizona"]},"updated_at":"2026-07-24T00:56:22Z"}