{"id":{"repo_id":"milano","oai_identifier":"oai:air.unimi.it:2434/1125844"},"canonical_url":"https://search.dev.ndltd.org/etd/milano/oai:air.unimi.it:2434/1125844","repository":{"repo_id":"milano","name":"Università degli Studi di Milano","base_url":"https://air.unimi.it/oai/request"},"display":{"title":"NANOPORE SEQUENCING AS A NEW TOOL TO EXPLORE COMPLEX TRANSCRIPTOMES","abstract":"Long-read sequencing marked a revolution in the field of transcriptomics by helping to resolve isoform structure, unannotated splicing variants, complex loci and repetitive regions and proposing as a method for RNA modifications detection. This versatility permits its application to different biological problems. For example, we leveraged an adaptation of direct-RNA Nanopore sequencing (dRNA-Seq), named NRCeq, to obtain a comprehensive full-length annotation of the SARS-CoV-2 transcriptome. In parallel, we identified putative pseudouridylation sites on multiple sgRNAs, one of which was located in a well-known regulatory region and proposed to have a role in viral subgenomic RNAs translation. Furthermore, within the FANTOM6 consortium, we used an adaption of the cDNA-PCR Nanopore sequencing protocol, named CFC-seq, to enable the selection for full-length reads and annotate the human non-coding RNAs and eRNAs in neural and monocytic cells. We benchmarked SALA, a custom assembler for CFC-seq data, against assemblers available in literature outlining its intermediate performances between algorithms highly reliant on reference annotation and those primarily depending on input datasets. Finally, integrating multiple sequencing protocols, we have generated a Breast Cancer transcriptomic Panel, to investigate alternative splicing, retrotransposons and RNA modifications on coding and non-coding RNAs in multiple cell lines and organoids. We have proven our capacity to assemble breast cancer cell lines, to retrieve annotated and unannotated protein coding and lncRNAs, to detect m6A sites across different transcript biotypes to identify retrotransposons. Overall, these works provide indication that long-read sequencing is a valuable resource for profiling the transcriptional and epitranscriptional landscape of an organism.","abstract_html":"Long-read sequencing marked a revolution in the field of transcriptomics by helping to resolve isoform structure, unannotated splicing variants, complex loci and repetitive regions and proposing as a method for RNA modifications detection. This versatility permits its application to different biological problems. For example, we leveraged an adaptation of direct-RNA Nanopore sequencing (dRNA-Seq), named NRCeq, to obtain a comprehensive full-length annotation of the SARS-CoV-2 transcriptome. In parallel, we identified putative pseudouridylation sites on multiple sgRNAs, one of which was located in a well-known regulatory region and proposed to have a role in viral subgenomic RNAs translation. Furthermore, within the FANTOM6 consortium, we used an adaption of the cDNA-PCR Nanopore sequencing protocol, named CFC-seq, to enable the selection for full-length reads and annotate the human non-coding RNAs and eRNAs in neural and monocytic cells. We benchmarked SALA, a custom assembler for CFC-seq data, against assemblers available in literature outlining its intermediate performances between algorithms highly reliant on reference annotation and those primarily depending on input datasets. Finally, integrating multiple sequencing protocols, we have generated a Breast Cancer transcriptomic Panel, to investigate alternative splicing, retrotransposons and RNA modifications on coding and non-coding RNAs in multiple cell lines and organoids. We have proven our capacity to assemble breast cancer cell lines, to retrieve annotated and unannotated protein coding and lncRNAs, to detect m6A sites across different transcript biotypes to identify retrotransposons. Overall, these works provide indication that long-read sequencing is a valuable resource for profiling the transcriptional and epitranscriptional landscape of an organism.","abstract_has_math":false,"creators":["UGOLINI, CAMILLA"],"institution":"Università degli Studi di Milano","degree_name":null,"degree_level":null,"degree_discipline":null,"degree_department":null,"school":null,"contributors":["supervisor: F. Nicassio","; co-supervisor: T. Leonardi ; co-supervisor: M. Marzi ; coordinator: D. Pasini","Leonardi","Tommaso Marzi","Matteo","C. Ugolini","PASINI, DIEGO"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2025,"date_issued":"2025-01-21","date_published":"2025-01-21","updated_at":"2026-07-27T20:18:51Z","subjects":["long-read sequencing, nanopore, SARS-CoV-2, breast cancer","Settore MED/04 - Patologia Generale","Settore MED/05 - Patologia Clinica","Settore MEDS-02/A - Patologia generale","Settore MEDS-02/B - Patologia clinica"],"languages":["eng"],"rights":["info:eu-repo/semantics/embargoedAccess","license:Creative commons","license uri:http://creativecommons.org/licenses/by/4.0/"],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"https://hdl.handle.net/2434/1125844","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["supervisor: F. Nicassio","; co-supervisor: T. Leonardi ; co-supervisor: M. Marzi ; coordinator: D. Pasini","Leonardi","Tommaso Marzi","Matteo","C. Ugolini","PASINI, DIEGO"]},{"key":"dc:creator","label":"Author","values":["UGOLINI, CAMILLA"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2025-01-21"]},{"key":"dc:publisher","label":"Institution","values":["Università degli Studi di Milano","place:IFOM campus"]},{"key":"dc:relation","label":"Dc Relation","values":["numberofpages:275","alleditors:Leonardi, Tommaso Marzi, Matteo"]},{"key":"dc:type","label":"Dc Type","values":["info:eu-repo/semantics/doctoralThesis"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["long-read sequencing, nanopore, SARS-CoV-2, breast cancer","Settore MED/04 - Patologia Generale","Settore MED/05 - Patologia Clinica","Settore MEDS-02/A - Patologia generale","Settore MEDS-02/B - Patologia clinica"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["eng"]},{"key":"dc:rights","label":"Dc Rights","values":["info:eu-repo/semantics/embargoedAccess","license:Creative commons","license uri:http://creativecommons.org/licenses/by/4.0/"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["https://hdl.handle.net/2434/1125844"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["Long-read sequencing marked a revolution in the field of transcriptomics by helping to resolve isoform structure, unannotated splicing variants, complex loci and repetitive regions and proposing as a method for RNA modifications detection. This versatility permits its application to different biological problems. For example, we leveraged an adaptation of direct-RNA Nanopore sequencing (dRNA-Seq), named NRCeq, to obtain a comprehensive full-length annotation of the SARS-CoV-2 transcriptome. In parallel, we identified putative pseudouridylation sites on multiple sgRNAs, one of which was located in a well-known regulatory region and proposed to have a role in viral subgenomic RNAs translation. Furthermore, within the FANTOM6 consortium, we used an adaption of the cDNA-PCR Nanopore sequencing protocol, named CFC-seq, to enable the selection for full-length reads and annotate the human non-coding RNAs and eRNAs in neural and monocytic cells. We benchmarked SALA, a custom assembler for CFC-seq data, against assemblers available in literature outlining its intermediate performances between algorithms highly reliant on reference annotation and those primarily depending on input datasets. Finally, integrating multiple sequencing protocols, we have generated a Breast Cancer transcriptomic Panel, to investigate alternative splicing, retrotransposons and RNA modifications on coding and non-coding RNAs in multiple cell lines and organoids. We have proven our capacity to assemble breast cancer cell lines, to retrieve annotated and unannotated protein coding and lncRNAs, to detect m6A sites across different transcript biotypes to identify retrotransposons. Overall, these works provide indication that long-read sequencing is a valuable resource for profiling the transcriptional and epitranscriptional landscape of an organism."]},{"key":"dc:title","label":"Title","values":["NANOPORE SEQUENCING AS A NEW TOOL TO EXPLORE COMPLEX TRANSCRIPTOMES"]}]}],"canonical_facts":{"dc:contributor":["supervisor: F. Nicassio","; co-supervisor: T. Leonardi ; co-supervisor: M. Marzi ; coordinator: D. Pasini","Leonardi","Tommaso Marzi","Matteo","C. Ugolini","PASINI, DIEGO"],"dc:creator":["UGOLINI, CAMILLA"],"dc:date":["2025-01-21"],"dc:description":["Long-read sequencing marked a revolution in the field of transcriptomics by helping to resolve isoform structure, unannotated splicing variants, complex loci and repetitive regions and proposing as a method for RNA modifications detection. This versatility permits its application to different biological problems. For example, we leveraged an adaptation of direct-RNA Nanopore sequencing (dRNA-Seq), named NRCeq, to obtain a comprehensive full-length annotation of the SARS-CoV-2 transcriptome. In parallel, we identified putative pseudouridylation sites on multiple sgRNAs, one of which was located in a well-known regulatory region and proposed to have a role in viral subgenomic RNAs translation. Furthermore, within the FANTOM6 consortium, we used an adaption of the cDNA-PCR Nanopore sequencing protocol, named CFC-seq, to enable the selection for full-length reads and annotate the human non-coding RNAs and eRNAs in neural and monocytic cells. We benchmarked SALA, a custom assembler for CFC-seq data, against assemblers available in literature outlining its intermediate performances between algorithms highly reliant on reference annotation and those primarily depending on input datasets. Finally, integrating multiple sequencing protocols, we have generated a Breast Cancer transcriptomic Panel, to investigate alternative splicing, retrotransposons and RNA modifications on coding and non-coding RNAs in multiple cell lines and organoids. We have proven our capacity to assemble breast cancer cell lines, to retrieve annotated and unannotated protein coding and lncRNAs, to detect m6A sites across different transcript biotypes to identify retrotransposons. Overall, these works provide indication that long-read sequencing is a valuable resource for profiling the transcriptional and epitranscriptional landscape of an organism."],"dc:identifier":["https://hdl.handle.net/2434/1125844"],"dc:language":["eng"],"dc:publisher":["Università degli Studi di Milano","place:IFOM campus"],"dc:relation":["numberofpages:275","alleditors:Leonardi, Tommaso Marzi, Matteo"],"dc:rights":["info:eu-repo/semantics/embargoedAccess","license:Creative commons","license uri:http://creativecommons.org/licenses/by/4.0/"],"dc:subject":["long-read sequencing, nanopore, SARS-CoV-2, breast cancer","Settore MED/04 - Patologia Generale","Settore MED/05 - Patologia Clinica","Settore MEDS-02/A - Patologia generale","Settore MEDS-02/B - Patologia clinica"],"dc:title":["NANOPORE SEQUENCING AS A NEW TOOL TO EXPLORE COMPLEX TRANSCRIPTOMES"],"dc:type":["info:eu-repo/semantics/doctoralThesis"]},"updated_at":"2026-07-27T20:18:51Z"}