{"id":{"repo_id":"uiuc","oai_identifier":"oai:www.ideals.illinois.edu:2142/116232"},"canonical_url":"https://search.dev.ndltd.org/etd/uiuc/oai:www.ideals.illinois.edu:2142/116232","repository":{"repo_id":"uiuc","name":"University of Illinois - Urbana-Champaign","base_url":"https://www.ideals.illinois.edu/oai-pmh"},"display":{"title":"Cross-correlations in medical data: theory, algorithms, and applications in disease analytics","abstract":"Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2022-11-15 without embargo terms","abstract_html":"Submission original under an indefinite embargo labeled &#x27;Open Access&#x27;. The submission was exported from vireo on 2022-11-15 without embargo terms","abstract_has_math":false,"creators":["Vaishnavi Subramanian, -"],"institution":"University of Illinois at Urbana-Champaign","degree_name":"Ph.D.","degree_level":"Dissertation","degree_discipline":"Electrical & Computer Engr","degree_department":null,"school":null,"contributors":["Do, Minh N.","Syeda-Mahmood, Tanveer","Sinha, Saurabh","Hajek, Bruce","Bresler, Yoram"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2022,"date_issued":"2022-08","date_published":"2022-08","updated_at":"2026-07-22T22:24:55Z","subjects":["mutli-modality","canonical correlation analysis","cancer"],"languages":["en","eng"],"rights":["Copyright 2022 - Vaishnavi Subramanian"],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"https://hdl.handle.net/2142/116232","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Do, Minh N.","Syeda-Mahmood, Tanveer","Sinha, Saurabh","Hajek, Bruce","Bresler, Yoram"]},{"key":"dc:creator","label":"Author","values":["Vaishnavi Subramanian, -"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2022-08","2022-07-14"]},{"key":"dc:type","label":"Dc Type","values":["text","Thesis"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Electrical & Computer Engr"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Dissertation"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Ph.D."]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Illinois at Urbana-Champaign"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["mutli-modality","canonical correlation analysis","cancer"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["en","eng"]},{"key":"dc:rights","label":"Dc Rights","values":["Copyright 2022 - Vaishnavi Subramanian"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["https://hdl.handle.net/2142/116232"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2022-11-15 without embargo terms","The student, - Vaishnavi Subramanian, accepted the attached license on 2022-07-14 at 08:40.","The student, - Vaishnavi Subramanian, submitted this Dissertation for approval on 2022-07-14 at 08:49.","This Dissertation was approved for publication on 2022-07-14 at 13:51.","DSpace SAF Submission Ingestion Package generated from Vireo submission #18289 on 2022-11-15 at 18:20:55","Cancer, like other complex diseases, involves a multitude of interactions which make it challenging to understand, diagnose, and cure. With the advancements in data science, there is immense potential to characterize the complexities of cancer by taking advantage of datasets that provide patient information from different views or modalities. For example, the Cancer Genome Atlas (TCGA) and the Cancer Imaging Archive (TCIA) provide data from multiple modalities, including radiology, histopathology, and genomics, from the same set of patients. The relations and correlations across features from these different modalities capture the joint variation in cancer properties and enable the characterization of the underlying cancer state in patients. This dissertation explores the potential of correlations as a tool for prediction and disease understanding. Under a two-modality probabilistic graphical model, we first show mathematically that the outputs of canonical correlation analysis (CCA), a linear correlation tool, is effective in the prediction of the model's unknown latent variable. The CCA-based predictor can alternatively be viewed as an interpretable unsupervised embedding generator which captures the correlations across two sets of features using CCA, followed by a supervised predictor. To adapt CCA to high dimensional, low sample, real world medical data, penalties encouraging sparsity, group structures, or graph structures are frequently included in the CCA formulation. To solve the graph-constrained CCA problem in a computationally efficient manner, we develop an alternating algorithm utilizing proximal gradients. We introduce novel matrix deflation (update) rules which enforce orthogonality properties to generate informative embeddings for CCA with penalties. Our results on simulated data and TCGA breast cancer data highlight the potential of our proposed framework. Lastly, we work on spatially-resolved proteomics data which provide expression levels of proteins spatially across tissue. The application of CCA on a recent spatial proteomics data from breast cancer allows the discovery of cross-correlations between neighbourhoods and protein levels, and provides insights into the differences in tissue structure between normal and cancerous tissue. In summary, this dissertation presents algorithms, techniques and proofs to design approaches which utilize cross-correlations in medical data for disease analytics."]},{"key":"dc:format","label":"Dc Format","values":["application/pdf"]},{"key":"dc:title","label":"Title","values":["Cross-correlations in medical data: theory, algorithms, and applications in disease analytics"]}]}],"canonical_facts":{"dc:contributor":["Do, Minh N.","Syeda-Mahmood, Tanveer","Sinha, Saurabh","Hajek, Bruce","Bresler, Yoram"],"dc:creator":["Vaishnavi Subramanian, -"],"dc:date":["2022-08","2022-07-14"],"dc:description":["Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2022-11-15 without embargo terms","The student, - Vaishnavi Subramanian, accepted the attached license on 2022-07-14 at 08:40.","The student, - Vaishnavi Subramanian, submitted this Dissertation for approval on 2022-07-14 at 08:49.","This Dissertation was approved for publication on 2022-07-14 at 13:51.","DSpace SAF Submission Ingestion Package generated from Vireo submission #18289 on 2022-11-15 at 18:20:55","Cancer, like other complex diseases, involves a multitude of interactions which make it challenging to understand, diagnose, and cure. With the advancements in data science, there is immense potential to characterize the complexities of cancer by taking advantage of datasets that provide patient information from different views or modalities. For example, the Cancer Genome Atlas (TCGA) and the Cancer Imaging Archive (TCIA) provide data from multiple modalities, including radiology, histopathology, and genomics, from the same set of patients. The relations and correlations across features from these different modalities capture the joint variation in cancer properties and enable the characterization of the underlying cancer state in patients. This dissertation explores the potential of correlations as a tool for prediction and disease understanding. Under a two-modality probabilistic graphical model, we first show mathematically that the outputs of canonical correlation analysis (CCA), a linear correlation tool, is effective in the prediction of the model's unknown latent variable. The CCA-based predictor can alternatively be viewed as an interpretable unsupervised embedding generator which captures the correlations across two sets of features using CCA, followed by a supervised predictor. To adapt CCA to high dimensional, low sample, real world medical data, penalties encouraging sparsity, group structures, or graph structures are frequently included in the CCA formulation. To solve the graph-constrained CCA problem in a computationally efficient manner, we develop an alternating algorithm utilizing proximal gradients. We introduce novel matrix deflation (update) rules which enforce orthogonality properties to generate informative embeddings for CCA with penalties. Our results on simulated data and TCGA breast cancer data highlight the potential of our proposed framework. Lastly, we work on spatially-resolved proteomics data which provide expression levels of proteins spatially across tissue. The application of CCA on a recent spatial proteomics data from breast cancer allows the discovery of cross-correlations between neighbourhoods and protein levels, and provides insights into the differences in tissue structure between normal and cancerous tissue. In summary, this dissertation presents algorithms, techniques and proofs to design approaches which utilize cross-correlations in medical data for disease analytics."],"dc:format":["application/pdf"],"dc:identifier":["https://hdl.handle.net/2142/116232"],"dc:language":["en","eng"],"dc:rights":["Copyright 2022 - Vaishnavi Subramanian"],"dc:subject":["mutli-modality","canonical correlation analysis","cancer"],"dc:title":["Cross-correlations in medical data: theory, algorithms, and applications in disease analytics"],"dc:type":["text","Thesis"],"thesis:degree_discipline":["Electrical & Computer Engr"],"thesis:degree_level":["Dissertation"],"thesis:degree_name":["Ph.D."],"thesis:institution_name":["University of Illinois at Urbana-Champaign"]},"updated_at":"2026-07-22T22:24:55Z"}