{"id":{"repo_id":"uiuc","oai_identifier":"oai:www.ideals.illinois.edu:2142/121446"},"canonical_url":"https://search.dev.ndltd.org/etd/uiuc/oai:www.ideals.illinois.edu:2142/121446","repository":{"repo_id":"uiuc","name":"University of Illinois - Urbana-Champaign","base_url":"https://www.ideals.illinois.edu/oai-pmh"},"display":{"title":"The learning the sounds of Japanese: Experimental and computational approaches","abstract":"Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2023-12-04 without embargo terms","abstract_html":"Submission original under an indefinite embargo labeled &#x27;Open Access&#x27;. The submission was exported from vireo on 2023-12-04 without embargo terms","abstract_has_math":false,"creators":["Silva Fonseca, Marco Aurelio"],"institution":"University of Illinois at Urbana-Champaign","degree_name":"Ph.D.","degree_level":"Dissertation","degree_discipline":"Linguistics","degree_department":null,"school":null,"contributors":["Hualde, Jose Ignacio","Scwhartz, Lane","Montrul, Silvina","Shosted, Ryan"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2023,"date_issued":"2023-08","date_published":"2023-08","updated_at":"2026-07-22T22:24:57Z","subjects":["Second Language Acquisition","Phonetics","Phonology","Neural Networks","Machine Learning","Japanese."],"languages":["en","eng"],"rights":["Copyright 2023 Marco Aurelio Silva Fonseca"],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"https://hdl.handle.net/2142/121446","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Hualde, Jose Ignacio","Scwhartz, Lane","Montrul, Silvina","Shosted, Ryan"]},{"key":"dc:creator","label":"Author","values":["Silva Fonseca, Marco Aurelio"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2023-08","2023-07-05"]},{"key":"dc:type","label":"Dc Type","values":["text"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Linguistics"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Dissertation"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Ph.D."]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Illinois at Urbana-Champaign"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Second Language Acquisition","Phonetics","Phonology","Neural Networks","Machine Learning","Japanese."]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["en","eng"]},{"key":"dc:rights","label":"Dc Rights","values":["Copyright 2023 Marco Aurelio Silva Fonseca"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["https://hdl.handle.net/2142/121446"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2023-12-04 without embargo terms","The student, Marco Aurelio Silva Fonseca, accepted the attached license on 2023-06-30 at 14:57.","The student, Marco Aurelio Silva Fonseca, submitted this Dissertation for approval on 2023-06-30 at 15:16.","This Dissertation was approved for publication on 2023-07-05 at 10:39.","DSpace SAF Submission Ingestion Package generated from Vireo submission #19490 on 2023-12-04 at 17:00:22","In this dissertation I analyze the learning of the sounds of Japanese from experimental and computational perspectives. More specifically, I conducted a perception experiment with second language learners who are native speakers of English (henceforth L2 learners) and modeling experiments using long-short term memory recurrent neural networks (LSTM RNNs). By using experimental machine learning to model sound phenomena and their acquisition, this dissertation provides a novel methodology to place second language acquisition and phonology research within an interdisciplinary field. In chapter 2, I present an experiment that I conducted with 22 advanced L2 learners and a control group of 17 native speakers of Japanese. In this experiment I analyzed the perception of three types of lexical contrasts: a) voiceless versus voiced stops: kakkou “outfit” versus gakkou “school”; b) long versus short vowels: biru “building” versus biiru “beer”; and c) words with lexical pitch accent on their last syllable versus first syllable: am ́e “candy” versus ́ame “rain”. The results of the experiment show that even though L2 learners perform as well as native speakers in the ABX discrimination, their performance is worse than native speakers in the lexical assignment task. Accuracy was particularly low for the pitch accent contrast. This indicates that even though L2 learners can hear the difference between contrasts that manifest phonetically differently in their language, it does not mean that they fully acquired such contrasts in their mental lexicon. In chapter 3, I investigate the role of distinctive features vs. phoneme embeddings to model the sounds of English and Japanese using long-short term memory recurrent neural networks (LSTM RNNs). Besides building English and Japanese networks, I additionally built two different types of bilingual networks: simultaneous (trained with English and Japanese at the same time) and consecutive (first trained in English and then in Japanese) with a variety of English/Japanese ratios. For both monolingual and bilingual models, feature-na ̈ıve networks outperformed feature-aware networks. These results corroborate previous research (Mirea and Bicknell 2019), and provide additional evidence using Japanese and bilingual networks. In chapter 4 I focus on the model’s performance related to two phonological phenomena of Japanese. By doing so, I propose a novel approach to model the phonology of Japanese. More specifically, I used LSTM RNNS to model nasal assimilation and loanword geminate devoicing. While nasal assimilation is a local phenomenon (i.e., it affects sounds adjacent to each other), loanword geminate devoicing is a long-distance phenomenon (i.e., it affects sounds across syllable boundaries). The results indicate that both bilingual and monolingual networks can learn nasal assimilation but not loanword geminate devoicing. This could be because LSTM RNNs are not efficient to model long-distance phenomena, but could also be because of the nature of the training data. In summary, this dissertation investigates how L2 learners and LSTM RNNs learn the sounds of Japanese. The results of chapter 2 revealed that L2 learners find it difficult to assign meaning to words that differ in pitch accent; chapter 3 reveals that LSTM RNNs without distinctive features in their architecture outperform those with them (in terms of cross-entropy values assigned to test data set); and chapter 4 revealed that LSTM RNNs can learn nasal assimilation but cannot learn geminate loanword devoicing. By employing both experimental and computational models, this dissertation provides evidence for the development of new techniques to conduct research on the acquisition of sounds by L2 learners."]},{"key":"dc:format","label":"Dc Format","values":["application/pdf"]},{"key":"dc:title","label":"Title","values":["The learning the sounds of Japanese: Experimental and computational approaches"]}]}],"canonical_facts":{"dc:contributor":["Hualde, Jose Ignacio","Scwhartz, Lane","Montrul, Silvina","Shosted, Ryan"],"dc:creator":["Silva Fonseca, Marco Aurelio"],"dc:date":["2023-08","2023-07-05"],"dc:description":["Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2023-12-04 without embargo terms","The student, Marco Aurelio Silva Fonseca, accepted the attached license on 2023-06-30 at 14:57.","The student, Marco Aurelio Silva Fonseca, submitted this Dissertation for approval on 2023-06-30 at 15:16.","This Dissertation was approved for publication on 2023-07-05 at 10:39.","DSpace SAF Submission Ingestion Package generated from Vireo submission #19490 on 2023-12-04 at 17:00:22","In this dissertation I analyze the learning of the sounds of Japanese from experimental and computational perspectives. More specifically, I conducted a perception experiment with second language learners who are native speakers of English (henceforth L2 learners) and modeling experiments using long-short term memory recurrent neural networks (LSTM RNNs). By using experimental machine learning to model sound phenomena and their acquisition, this dissertation provides a novel methodology to place second language acquisition and phonology research within an interdisciplinary field. In chapter 2, I present an experiment that I conducted with 22 advanced L2 learners and a control group of 17 native speakers of Japanese. In this experiment I analyzed the perception of three types of lexical contrasts: a) voiceless versus voiced stops: kakkou “outfit” versus gakkou “school”; b) long versus short vowels: biru “building” versus biiru “beer”; and c) words with lexical pitch accent on their last syllable versus first syllable: am ́e “candy” versus ́ame “rain”. The results of the experiment show that even though L2 learners perform as well as native speakers in the ABX discrimination, their performance is worse than native speakers in the lexical assignment task. Accuracy was particularly low for the pitch accent contrast. This indicates that even though L2 learners can hear the difference between contrasts that manifest phonetically differently in their language, it does not mean that they fully acquired such contrasts in their mental lexicon. In chapter 3, I investigate the role of distinctive features vs. phoneme embeddings to model the sounds of English and Japanese using long-short term memory recurrent neural networks (LSTM RNNs). Besides building English and Japanese networks, I additionally built two different types of bilingual networks: simultaneous (trained with English and Japanese at the same time) and consecutive (first trained in English and then in Japanese) with a variety of English/Japanese ratios. For both monolingual and bilingual models, feature-na ̈ıve networks outperformed feature-aware networks. These results corroborate previous research (Mirea and Bicknell 2019), and provide additional evidence using Japanese and bilingual networks. In chapter 4 I focus on the model’s performance related to two phonological phenomena of Japanese. By doing so, I propose a novel approach to model the phonology of Japanese. More specifically, I used LSTM RNNS to model nasal assimilation and loanword geminate devoicing. While nasal assimilation is a local phenomenon (i.e., it affects sounds adjacent to each other), loanword geminate devoicing is a long-distance phenomenon (i.e., it affects sounds across syllable boundaries). The results indicate that both bilingual and monolingual networks can learn nasal assimilation but not loanword geminate devoicing. This could be because LSTM RNNs are not efficient to model long-distance phenomena, but could also be because of the nature of the training data. In summary, this dissertation investigates how L2 learners and LSTM RNNs learn the sounds of Japanese. The results of chapter 2 revealed that L2 learners find it difficult to assign meaning to words that differ in pitch accent; chapter 3 reveals that LSTM RNNs without distinctive features in their architecture outperform those with them (in terms of cross-entropy values assigned to test data set); and chapter 4 revealed that LSTM RNNs can learn nasal assimilation but cannot learn geminate loanword devoicing. By employing both experimental and computational models, this dissertation provides evidence for the development of new techniques to conduct research on the acquisition of sounds by L2 learners."],"dc:format":["application/pdf"],"dc:identifier":["https://hdl.handle.net/2142/121446"],"dc:language":["en","eng"],"dc:rights":["Copyright 2023 Marco Aurelio Silva Fonseca"],"dc:subject":["Second Language Acquisition","Phonetics","Phonology","Neural Networks","Machine Learning","Japanese."],"dc:title":["The learning the sounds of Japanese: Experimental and computational approaches"],"dc:type":["text"],"thesis:degree_discipline":["Linguistics"],"thesis:degree_level":["Dissertation"],"thesis:degree_name":["Ph.D."],"thesis:institution_name":["University of Illinois at Urbana-Champaign"]},"updated_at":"2026-07-22T22:24:57Z"}