{"id":{"repo_id":"uiuc","oai_identifier":"oai:www.ideals.illinois.edu:2142/15967"},"canonical_url":"https://search.dev.ndltd.org/etd/uiuc/oai:www.ideals.illinois.edu:2142/15967","repository":{"repo_id":"uiuc","name":"University of Illinois - Urbana-Champaign","base_url":"https://www.ideals.illinois.edu/oai-pmh"},"display":{"title":"Automated methods for correcting errors in grammar and usage","abstract":"Over the last several decades, the number of electronic documents has increased dramatically. With the growing availability of computers, more and more people are using text editors. However, the development of automated methods for correcting mistakes in text has not progressed as far. Text editors usually employ basic spell checking techniques and address very few mistakes of other types. In this thesis, we propose two methods for correcting errors in grammar and usage. First, we propose a novel approach to the problem of training classifiers to detect and correct errors in text by selectively introducing mistakes into the training data and show that this method is superior to the traditional method of training using clean data. Second, we define high-level features and propose a method of correcting mistakes using these features. We combine the two methods and build a system for correcting mistakes in article usage made by non-native speakers of English.","abstract_html":"Over the last several decades, the number of electronic documents has increased dramatically. With the growing availability of computers, more and more people are using text editors. However, the development of automated methods for correcting mistakes in text has not progressed as far. Text editors usually employ basic spell checking techniques and address very few mistakes of other types. In this thesis, we propose two methods for correcting errors in grammar and usage. First, we propose a novel approach to the problem of training classifiers to detect and correct errors in text by selectively introducing mistakes into the training data and show that this method is superior to the traditional method of training using clean data. Second, we define high-level features and propose a method of correcting mistakes using these features. We combine the two methods and build a system for correcting mistakes in article usage made by non-native speakers of English.","abstract_has_math":false,"creators":["Rozovskaya, Alla"],"institution":"University of Illinois at Urbana-Champaign","degree_name":"M.S.","degree_level":"Thesis","degree_discipline":"Computer Science","degree_department":null,"school":null,"contributors":["Roth, Dan"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2010,"date_issued":"2010-05-18T18:52:57Z","date_published":"2010-05-18T18:52:57Z","updated_at":"2026-07-22T22:25:08Z","subjects":["Text correction","Correcting ESL mistakes","Errors in article usage","automated methods for correcting ESL mistakes","annotated ESL corpus","Error statistics","Features for article correction","English as a Second Language (ESL)"],"languages":["en"],"rights":["Copyright 2010 Alla Rozovskaya"],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"http://hdl.handle.net/2142/15967","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Roth, Dan"]},{"key":"dc:creator","label":"Author","values":["Rozovskaya, Alla"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2010-05-18T18:52:57Z","2012-05-19T10:00:11Z","2010-5"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Computer Science"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Thesis"]},{"key":"thesis:degree_name","label":"Degree Name","values":["M.S."]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Illinois at Urbana-Champaign"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Text correction","Correcting ESL mistakes","Errors in article usage","automated methods for correcting ESL mistakes","annotated ESL corpus","Error statistics","Features for article correction","English as a Second Language (ESL)"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["en"]},{"key":"dc:rights","label":"Dc Rights","values":["Copyright 2010 Alla Rozovskaya"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["http://hdl.handle.net/2142/15967"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["Over the last several decades, the number of electronic documents has increased dramatically. With the growing availability of computers, more and more people are using text editors. However, the development of automated methods for correcting mistakes in text has not progressed as far. Text editors usually employ basic spell checking techniques and address very few mistakes of other types. In this thesis, we propose two methods for correcting errors in grammar and usage. First, we propose a novel approach to the problem of training classifiers to detect and correct errors in text by selectively introducing mistakes into the training data and show that this method is superior to the traditional method of training using clean data. Second, we define high-level features and propose a method of correcting mistakes using these features. We combine the two methods and build a system for correcting mistakes in article usage made by non-native speakers of English.","Item withdrawn by Mark Zulauf (zulauf@illinois.edu) on 2010-04-29T18:11:24Z Item was in collections: University of Illinois Theses & Dissertations (ID: 1) No. of bitstreams: 1 Rozovskaya_Alla.pdf: 425941 bytes, checksum: a27d165975d6f71b30516f5fa8067a17 (MD5)","Made available in DSpace on 2010-05-18T18:52:57Z (GMT). No. of bitstreams: 3 Rozovskaya_Alla.pdf: 425941 bytes, checksum: a27d165975d6f71b30516f5fa8067a17 (MD5) 1_Rozovskaya_Alla.pdf: 426071 bytes, checksum: c6b0953781f86c705d9839afcf6ac662 (MD5) license.txt: 4065 bytes, checksum: 953e9604a4b02a4146d4d7e9e4be3f2c (MD5)","Item marked as restricted to the 'UIUC Users [automated]' Group (id=2) by William Ingram (wingram2@illinois.edu) on 2010-05-18T18:54:48Z Item is restricted until 2012-05-18T18:54:47Z","Item reinstated by Sarah Shreeves (sshreeve@illinois.edu) on 2012-05-19T10:00:11Z Item was in collections: University of Illinois Dissertations and Theses (ID: 204) Dissertations and Theses - Computer Science (ID: 587) No. of bitstreams: 4 1_Rozovskaya_Alla.pdf.txt: 136988 bytes, checksum: 3050b73e459aeb38a43e8c65af8c0e4b (MD5) Rozovskaya_Alla.pdf: 425941 bytes, checksum: a27d165975d6f71b30516f5fa8067a17 (MD5) 1_Rozovskaya_Alla.pdf: 426071 bytes, checksum: c6b0953781f86c705d9839afcf6ac662 (MD5) license.txt: 4065 bytes, checksum: 953e9604a4b02a4146d4d7e9e4be3f2c (MD5)","Item released from any restrictions by Sarah Shreeves (sshreeve@illinois.edu) on 2012-05-19T10:00:11Z"]},{"key":"dc:title","label":"Title","values":["Automated methods for correcting errors in grammar and usage"]}]}],"canonical_facts":{"dc:contributor":["Roth, Dan"],"dc:creator":["Rozovskaya, Alla"],"dc:date":["2010-05-18T18:52:57Z","2012-05-19T10:00:11Z","2010-5"],"dc:description":["Over the last several decades, the number of electronic documents has increased dramatically. With the growing availability of computers, more and more people are using text editors. However, the development of automated methods for correcting mistakes in text has not progressed as far. Text editors usually employ basic spell checking techniques and address very few mistakes of other types. In this thesis, we propose two methods for correcting errors in grammar and usage. First, we propose a novel approach to the problem of training classifiers to detect and correct errors in text by selectively introducing mistakes into the training data and show that this method is superior to the traditional method of training using clean data. Second, we define high-level features and propose a method of correcting mistakes using these features. We combine the two methods and build a system for correcting mistakes in article usage made by non-native speakers of English.","Item withdrawn by Mark Zulauf (zulauf@illinois.edu) on 2010-04-29T18:11:24Z Item was in collections: University of Illinois Theses & Dissertations (ID: 1) No. of bitstreams: 1 Rozovskaya_Alla.pdf: 425941 bytes, checksum: a27d165975d6f71b30516f5fa8067a17 (MD5)","Made available in DSpace on 2010-05-18T18:52:57Z (GMT). No. of bitstreams: 3 Rozovskaya_Alla.pdf: 425941 bytes, checksum: a27d165975d6f71b30516f5fa8067a17 (MD5) 1_Rozovskaya_Alla.pdf: 426071 bytes, checksum: c6b0953781f86c705d9839afcf6ac662 (MD5) license.txt: 4065 bytes, checksum: 953e9604a4b02a4146d4d7e9e4be3f2c (MD5)","Item marked as restricted to the 'UIUC Users [automated]' Group (id=2) by William Ingram (wingram2@illinois.edu) on 2010-05-18T18:54:48Z Item is restricted until 2012-05-18T18:54:47Z","Item reinstated by Sarah Shreeves (sshreeve@illinois.edu) on 2012-05-19T10:00:11Z Item was in collections: University of Illinois Dissertations and Theses (ID: 204) Dissertations and Theses - Computer Science (ID: 587) No. of bitstreams: 4 1_Rozovskaya_Alla.pdf.txt: 136988 bytes, checksum: 3050b73e459aeb38a43e8c65af8c0e4b (MD5) Rozovskaya_Alla.pdf: 425941 bytes, checksum: a27d165975d6f71b30516f5fa8067a17 (MD5) 1_Rozovskaya_Alla.pdf: 426071 bytes, checksum: c6b0953781f86c705d9839afcf6ac662 (MD5) license.txt: 4065 bytes, checksum: 953e9604a4b02a4146d4d7e9e4be3f2c (MD5)","Item released from any restrictions by Sarah Shreeves (sshreeve@illinois.edu) on 2012-05-19T10:00:11Z"],"dc:identifier":["http://hdl.handle.net/2142/15967"],"dc:language":["en"],"dc:rights":["Copyright 2010 Alla Rozovskaya"],"dc:subject":["Text correction","Correcting ESL mistakes","Errors in article usage","automated methods for correcting ESL mistakes","annotated ESL corpus","Error statistics","Features for article correction","English as a Second Language (ESL)"],"dc:title":["Automated methods for correcting errors in grammar and usage"],"thesis:degree_discipline":["Computer Science"],"thesis:degree_level":["Thesis"],"thesis:degree_name":["M.S."],"thesis:institution_name":["University of Illinois at Urbana-Champaign"]},"updated_at":"2026-07-22T22:25:08Z"}