{"id":{"repo_id":"uiuc","oai_identifier":"oai:www.ideals.illinois.edu:2142/97787"},"canonical_url":"https://search.dev.ndltd.org/etd/uiuc/oai:www.ideals.illinois.edu:2142/97787","repository":{"repo_id":"uiuc","name":"University of Illinois - Urbana-Champaign","base_url":"https://www.ideals.illinois.edu/oai-pmh"},"display":{"title":"Comprehensive evaluation of error correction methods for high-throughput sequencing data","abstract":"The advent of DNA and RNA sequencing has significantly revolutionized the study of genomics and molecular biology. Development of high-throughput sequencing technologies have brought about a quick and cheaper way to sequence genomes. Different technologies use different underlying methods for sequencing and are prone to different error rates. Though many tools exist for error correction in high-throughput sequencing data, no standard technology-independent method is available yet to evaluate the accuracy and effectiveness of these error correction tools. In order to supply a standard way to evaluate error correction methods for DNA and RNA sequencing, this thesis presents a Software Package for Error Correction Tool Assessment on nuCLEic acid sequences (SPECTACLE). SPECTACLE can evaluate corrected DNA and RNA reads from many underlying sequencing technologies and differentiate heterozygous alleles from sequencing errors. The work provides some key insights on many factors that stress the challenges in error correction by compiling high-throughput sequencing read sets from technologies like Illumina, PacBio and ONT. The performances of 23 different error correction tools have been analyzed using SPECTACLE and the compiled datasets. This thesis also provides unique and helpful insights into the strengths and weaknesses of various error correction tools and aims to establish a standard platform for evaluating error correction tools in the future.","abstract_html":"The advent of DNA and RNA sequencing has significantly revolutionized the study of genomics and molecular biology. Development of high-throughput sequencing technologies have brought about a quick and cheaper way to sequence genomes. Different technologies use different underlying methods for sequencing and are prone to different error rates. Though many tools exist for error correction in high-throughput sequencing data, no standard technology-independent method is available yet to evaluate the accuracy and effectiveness of these error correction tools. In order to supply a standard way to evaluate error correction methods for DNA and RNA sequencing, this thesis presents a Software Package for Error Correction Tool Assessment on nuCLEic acid sequences (SPECTACLE). SPECTACLE can evaluate corrected DNA and RNA reads from many underlying sequencing technologies and differentiate heterozygous alleles from sequencing errors. The work provides some key insights on many factors that stress the challenges in error correction by compiling high-throughput sequencing read sets from technologies like Illumina, PacBio and ONT. The performances of 23 different error correction tools have been analyzed using SPECTACLE and the compiled datasets. This thesis also provides unique and helpful insights into the strengths and weaknesses of various error correction tools and aims to establish a standard platform for evaluating error correction tools in the future.","abstract_has_math":false,"creators":["Manikandan, Gowthami Jayashri"],"institution":"University of Illinois at Urbana-Champaign","degree_name":"M.S.","degree_level":"Thesis","degree_discipline":"Electrical & Computer Engr","degree_department":null,"school":null,"contributors":["Chen, Deming"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2017,"date_issued":"2017-08-10T20:33:25Z","date_published":"2017-08-10T20:33:25Z","updated_at":"2026-07-22T22:24:34Z","subjects":["Computational genomics","Algorithms","Parallel computing"],"languages":["en"],"rights":["Copyright 2017 Gowthami Jayashri Manikandan"],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"http://hdl.handle.net/2142/97787","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Chen, Deming"]},{"key":"dc:creator","label":"Author","values":["Manikandan, Gowthami Jayashri"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2017-08-10T20:33:25Z","2019-08-11T09:15:24Z","2017-04-27","2017-05"]},{"key":"dc:type","label":"Dc Type","values":["text"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Electrical & Computer Engr"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Thesis"]},{"key":"thesis:degree_name","label":"Degree Name","values":["M.S."]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Illinois at Urbana-Champaign"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Computational genomics","Algorithms","Parallel computing"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["en"]},{"key":"dc:rights","label":"Dc Rights","values":["Copyright 2017 Gowthami Jayashri Manikandan"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["http://hdl.handle.net/2142/97787"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["The advent of DNA and RNA sequencing has significantly revolutionized the study of genomics and molecular biology. Development of high-throughput sequencing technologies have brought about a quick and cheaper way to sequence genomes. Different technologies use different underlying methods for sequencing and are prone to different error rates. Though many tools exist for error correction in high-throughput sequencing data, no standard technology-independent method is available yet to evaluate the accuracy and effectiveness of these error correction tools. In order to supply a standard way to evaluate error correction methods for DNA and RNA sequencing, this thesis presents a Software Package for Error Correction Tool Assessment on nuCLEic acid sequences (SPECTACLE). SPECTACLE can evaluate corrected DNA and RNA reads from many underlying sequencing technologies and differentiate heterozygous alleles from sequencing errors. The work provides some key insights on many factors that stress the challenges in error correction by compiling high-throughput sequencing read sets from technologies like Illumina, PacBio and ONT. The performances of 23 different error correction tools have been analyzed using SPECTACLE and the compiled datasets. This thesis also provides unique and helpful insights into the strengths and weaknesses of various error correction tools and aims to establish a standard platform for evaluating error correction tools in the future.","Submission published under a 24 month embargo labeled 'U of I Access', the embargo will last until 2019-05-01","The student, Gowthami Jayashri Manikandan, accepted the attached license on 2017-04-25 at 16:19.","The student, Gowthami Jayashri Manikandan, submitted this Thesis for approval on 2017-04-25 at 16:29.","This Thesis was approved for publication on 2017-04-27 at 16:31.","DSpace SAF Submission Ingestion Package generated from Vireo submission #11062 on 2017-08-10 at 15:07:05","Made available in DSpace on 2017-08-10T20:33:25Z (GMT). No. of bitstreams: 3 MANIKANDAN-THESIS-2017.pdf: 1904607 bytes, checksum: a837fc9657f487fe35aba3f9be7b1886 (MD5) Gowthami_thesis_updated.docx: 1632427 bytes, checksum: 8f4877612eeb9f92256f03921f038e23 (MD5) LICENSE.txt: 4225 bytes, checksum: 2c6c92e0015a13d31c18b8ba34484a54 (MD5) Previous issue date: 2017-04-27","Embargo set by: Colleen Fallaw for item 102840 Lift date: 2019-08-10T21:27:21Z Reason: Author requested U of Illinois access only (OA after 2yrs) in Vireo ETD system","U of I Only Restriction Lifted for Item 102840 on 2019-08-11T09:15:24Z."]},{"key":"dc:format","label":"Dc Format","values":["application/pdf"]},{"key":"dc:title","label":"Title","values":["Comprehensive evaluation of error correction methods for high-throughput sequencing data"]}]}],"canonical_facts":{"dc:contributor":["Chen, Deming"],"dc:creator":["Manikandan, Gowthami Jayashri"],"dc:date":["2017-08-10T20:33:25Z","2019-08-11T09:15:24Z","2017-04-27","2017-05"],"dc:description":["The advent of DNA and RNA sequencing has significantly revolutionized the study of genomics and molecular biology. Development of high-throughput sequencing technologies have brought about a quick and cheaper way to sequence genomes. Different technologies use different underlying methods for sequencing and are prone to different error rates. Though many tools exist for error correction in high-throughput sequencing data, no standard technology-independent method is available yet to evaluate the accuracy and effectiveness of these error correction tools. In order to supply a standard way to evaluate error correction methods for DNA and RNA sequencing, this thesis presents a Software Package for Error Correction Tool Assessment on nuCLEic acid sequences (SPECTACLE). SPECTACLE can evaluate corrected DNA and RNA reads from many underlying sequencing technologies and differentiate heterozygous alleles from sequencing errors. The work provides some key insights on many factors that stress the challenges in error correction by compiling high-throughput sequencing read sets from technologies like Illumina, PacBio and ONT. The performances of 23 different error correction tools have been analyzed using SPECTACLE and the compiled datasets. This thesis also provides unique and helpful insights into the strengths and weaknesses of various error correction tools and aims to establish a standard platform for evaluating error correction tools in the future.","Submission published under a 24 month embargo labeled 'U of I Access', the embargo will last until 2019-05-01","The student, Gowthami Jayashri Manikandan, accepted the attached license on 2017-04-25 at 16:19.","The student, Gowthami Jayashri Manikandan, submitted this Thesis for approval on 2017-04-25 at 16:29.","This Thesis was approved for publication on 2017-04-27 at 16:31.","DSpace SAF Submission Ingestion Package generated from Vireo submission #11062 on 2017-08-10 at 15:07:05","Made available in DSpace on 2017-08-10T20:33:25Z (GMT). No. of bitstreams: 3 MANIKANDAN-THESIS-2017.pdf: 1904607 bytes, checksum: a837fc9657f487fe35aba3f9be7b1886 (MD5) Gowthami_thesis_updated.docx: 1632427 bytes, checksum: 8f4877612eeb9f92256f03921f038e23 (MD5) LICENSE.txt: 4225 bytes, checksum: 2c6c92e0015a13d31c18b8ba34484a54 (MD5) Previous issue date: 2017-04-27","Embargo set by: Colleen Fallaw for item 102840 Lift date: 2019-08-10T21:27:21Z Reason: Author requested U of Illinois access only (OA after 2yrs) in Vireo ETD system","U of I Only Restriction Lifted for Item 102840 on 2019-08-11T09:15:24Z."],"dc:format":["application/pdf"],"dc:identifier":["http://hdl.handle.net/2142/97787"],"dc:language":["en"],"dc:rights":["Copyright 2017 Gowthami Jayashri Manikandan"],"dc:subject":["Computational genomics","Algorithms","Parallel computing"],"dc:title":["Comprehensive evaluation of error correction methods for high-throughput sequencing data"],"dc:type":["text"],"thesis:degree_discipline":["Electrical & Computer Engr"],"thesis:degree_level":["Thesis"],"thesis:degree_name":["M.S."],"thesis:institution_name":["University of Illinois at Urbana-Champaign"]},"updated_at":"2026-07-22T22:24:34Z"}