{"id":{"repo_id":"south-carolina","oai_identifier":"oai:scholarcommons.sc.edu:etd-1315"},"canonical_url":"https://search.dev.ndltd.org/etd/south-carolina/oai:scholarcommons.sc.edu:etd-1315","repository":{"repo_id":"south-carolina","name":"University of South Carolina","base_url":"https://scholarcommons.sc.edu/do/oai/"},"display":{"title":"Molecular Evolution of Viruses","abstract":"<p>Molecular evolution is the process of evolution at the scale of DNA, RNA and proteins. Our goal was to study molecular evolution of viruses with special reference to Influenza A Virus.</p> <p>The recent Influenza A/H1N1(2009) outbreak has been the focus of intense research because of its high level of infectivity across the globe. Analysis of nucleotide sequence polymorphism in the genomic segments encoding the two most immunologically important proteins of influenza A, neuraminidase (NA) and hemagglutinin (HA), showed that the H1N1 (2009) has resulted from the spread of an HA segment of recent origin and low diversity through a population of ancient and much more diverse NA segments. We also studied about the selection pattern of CTL epitope and CTL non-epitope regions of different proteins of Influenza and found that natural selection is the main pattern of selection with epitope regions being more conserved than non-epitope regions.</p> <p>This dissertation also focuses on implementation of a methodology to study phylogeny of viral protein sequences based on a technique called chaos game representation. Here we present a method for using CGR to explore evolutionary relationships of protein sequences based on amino acid properties and illustrate the approach with complete sets of protein translations from viral genomes. In an analysis of complete polyprotein sequences from the viral family Flaviviridae, the CGR method was able to cluster members of major viral groups together, but relations within groups were not well resolved in comparison to an alignment-based phylogeny. We applied the method to members of five different families of ssRNA positive-strand viruses, and each family formed a distinct cluster. We also present a method of testing the reliability of clustering in CGR-based trees, which involves the use of multiple random starting points for the CGR process.</p>","abstract_html":"&lt;p&gt;Molecular evolution is the process of evolution at the scale of DNA, RNA and proteins. Our goal was to study molecular evolution of viruses with special reference to Influenza A Virus.&lt;/p&gt; &lt;p&gt;The recent Influenza A/H1N1(2009) outbreak has been the focus of intense research because of its high level of infectivity across the globe. Analysis of nucleotide sequence polymorphism in the genomic segments encoding the two most immunologically important proteins of influenza A, neuraminidase (NA) and hemagglutinin (HA), showed that the H1N1 (2009) has resulted from the spread of an HA segment of recent origin and low diversity through a population of ancient and much more diverse NA segments. We also studied about the selection pattern of CTL epitope and CTL non-epitope regions of different proteins of Influenza and found that natural selection is the main pattern of selection with epitope regions being more conserved than non-epitope regions.&lt;/p&gt; &lt;p&gt;This dissertation also focuses on implementation of a methodology to study phylogeny of viral protein sequences based on a technique called chaos game representation. Here we present a method for using CGR to explore evolutionary relationships of protein sequences based on amino acid properties and illustrate the approach with complete sets of protein translations from viral genomes. In an analysis of complete polyprotein sequences from the viral family Flaviviridae, the CGR method was able to cluster members of major viral groups together, but relations within groups were not well resolved in comparison to an alignment-based phylogeny. We applied the method to members of five different families of ssRNA positive-strand viruses, and each family formed a distinct cluster. We also present a method of testing the reliability of clustering in CGR-based trees, which involves the use of multiple random starting points for the CGR process.&lt;/p&gt;","abstract_has_math":false,"creators":["Bhoumik, Priyasma"],"institution":null,"degree_name":"Ph.D.","degree_level":"Campus Access Dissertation","degree_discipline":"Biological Sciences","degree_department":null,"school":null,"contributors":["Austin L. Hughes"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2010,"date_issued":"2010-01-01T08:00:00Z","date_published":"2010-01-01T08:00:00Z","updated_at":"2026-07-24T04:37:08Z","subjects":["Life Sciences","Medicine and Health Sciences","Physical Sciences and Mathematics","Chaos Game Representation","Evolution","Influenza Virus","Phylogenetic Analysis"],"languages":[],"rights":["© 2010, Priyasma Bhoumik"],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"https://scholarcommons.sc.edu/etd/314","outbound_label":"Repository record","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Austin L. Hughes"]},{"key":"dc:creator","label":"Author","values":["Bhoumik, Priyasma"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"thesis:degree_discipline","label":"Discipline","values":["Biological Sciences"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Campus Access Dissertation"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Ph.D."]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Life Sciences","Medicine and Health Sciences","Physical Sciences and Mathematics","Chaos Game Representation","Evolution","Influenza Virus","Phylogenetic Analysis"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:rights","label":"Dc Rights","values":["© 2010, Priyasma Bhoumik"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["https://scholarcommons.sc.edu/etd/314"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["<p>Molecular evolution is the process of evolution at the scale of DNA, RNA and proteins. Our goal was to study molecular evolution of viruses with special reference to Influenza A Virus.</p> <p>The recent Influenza A/H1N1(2009) outbreak has been the focus of intense research because of its high level of infectivity across the globe. Analysis of nucleotide sequence polymorphism in the genomic segments encoding the two most immunologically important proteins of influenza A, neuraminidase (NA) and hemagglutinin (HA), showed that the H1N1 (2009) has resulted from the spread of an HA segment of recent origin and low diversity through a population of ancient and much more diverse NA segments. We also studied about the selection pattern of CTL epitope and CTL non-epitope regions of different proteins of Influenza and found that natural selection is the main pattern of selection with epitope regions being more conserved than non-epitope regions.</p> <p>This dissertation also focuses on implementation of a methodology to study phylogeny of viral protein sequences based on a technique called chaos game representation. Here we present a method for using CGR to explore evolutionary relationships of protein sequences based on amino acid properties and illustrate the approach with complete sets of protein translations from viral genomes. In an analysis of complete polyprotein sequences from the viral family Flaviviridae, the CGR method was able to cluster members of major viral groups together, but relations within groups were not well resolved in comparison to an alignment-based phylogeny. We applied the method to members of five different families of ssRNA positive-strand viruses, and each family formed a distinct cluster. We also present a method of testing the reliability of clustering in CGR-based trees, which involves the use of multiple random starting points for the CGR process.</p>"]},{"key":"dc:title","label":"Title","values":["Molecular Evolution of Viruses"]}]}],"canonical_facts":{"dc:contributor":["Austin L. Hughes"],"dc:creator":["Bhoumik, Priyasma"],"dc:description.abstract":["<p>Molecular evolution is the process of evolution at the scale of DNA, RNA and proteins. Our goal was to study molecular evolution of viruses with special reference to Influenza A Virus.</p> <p>The recent Influenza A/H1N1(2009) outbreak has been the focus of intense research because of its high level of infectivity across the globe. Analysis of nucleotide sequence polymorphism in the genomic segments encoding the two most immunologically important proteins of influenza A, neuraminidase (NA) and hemagglutinin (HA), showed that the H1N1 (2009) has resulted from the spread of an HA segment of recent origin and low diversity through a population of ancient and much more diverse NA segments. We also studied about the selection pattern of CTL epitope and CTL non-epitope regions of different proteins of Influenza and found that natural selection is the main pattern of selection with epitope regions being more conserved than non-epitope regions.</p> <p>This dissertation also focuses on implementation of a methodology to study phylogeny of viral protein sequences based on a technique called chaos game representation. Here we present a method for using CGR to explore evolutionary relationships of protein sequences based on amino acid properties and illustrate the approach with complete sets of protein translations from viral genomes. In an analysis of complete polyprotein sequences from the viral family Flaviviridae, the CGR method was able to cluster members of major viral groups together, but relations within groups were not well resolved in comparison to an alignment-based phylogeny. We applied the method to members of five different families of ssRNA positive-strand viruses, and each family formed a distinct cluster. We also present a method of testing the reliability of clustering in CGR-based trees, which involves the use of multiple random starting points for the CGR process.</p>"],"dc:identifier":["https://scholarcommons.sc.edu/etd/314"],"dc:rights":["© 2010, Priyasma Bhoumik"],"dc:subject":["Life Sciences","Medicine and Health Sciences","Physical Sciences and Mathematics","Chaos Game Representation","Evolution","Influenza Virus","Phylogenetic Analysis"],"dc:title":["Molecular Evolution of Viruses"],"thesis:degree_discipline":["Biological Sciences"],"thesis:degree_level":["Campus Access Dissertation"],"thesis:degree_name":["Ph.D."]},"updated_at":"2026-07-24T04:37:08Z"}