{"id":{"repo_id":"uiuc","oai_identifier":"oai:www.ideals.illinois.edu:2142/127502"},"canonical_url":"https://search.dev.ndltd.org/etd/uiuc/oai:www.ideals.illinois.edu:2142/127502","repository":{"repo_id":"uiuc","name":"University of Illinois - Urbana-Champaign","base_url":"https://www.ideals.illinois.edu/oai-pmh"},"display":{"title":"Data-rich experimentation and computer-aided strategies in reaction discovery, selectivity optimization, and structure elucidation","abstract":"Submission published under a 24 month embargo labeled 'Closed Access', the embargo will last until 2026-12-01","abstract_html":"Submission published under a 24 month embargo labeled &#x27;Closed Access&#x27;, the embargo will last until 2026-12-01","abstract_has_math":false,"creators":["Shved, Alexander S"],"institution":"University of Illinois at Urbana-Champaign","degree_name":"Ph.D.","degree_level":"Dissertation","degree_discipline":"Chemistry","degree_department":null,"school":null,"contributors":["Denmark, Scott E","Sarlah, David","Fataftah, Majed S","Pogorelov, Taras V"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2024,"date_issued":"2024-12","date_published":"2024-12","updated_at":"2026-07-22T22:25:04Z","subjects":["High-throughput Experimentation","Reaction Discovery","Reaction Optimization"],"languages":["en","eng"],"rights":["© 2024 Alexander S. Shved"],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"https://hdl.handle.net/2142/127502","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Denmark, Scott E","Sarlah, David","Fataftah, Majed S","Pogorelov, Taras V"]},{"key":"dc:creator","label":"Author","values":["Shved, Alexander S"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2024-12","2024-12-05"]},{"key":"dc:type","label":"Dc Type","values":["text","Thesis"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Chemistry"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Dissertation"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Ph.D."]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Illinois at Urbana-Champaign"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["High-throughput Experimentation","Reaction Discovery","Reaction Optimization"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["en","eng"]},{"key":"dc:rights","label":"Dc Rights","values":["© 2024 Alexander S. Shved"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["https://hdl.handle.net/2142/127502"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["Submission published under a 24 month embargo labeled 'Closed Access', the embargo will last until 2026-12-01","The student, Alexander Shved, accepted the attached license on 2024-12-05 at 05:53.","The student, Alexander Shved, submitted this Dissertation for approval on 2024-12-05 at 06:07.","This Dissertation was approved for publication on 2024-12-05 at 18:55.","DSpace SAF Submission Ingestion Package generated from Vireo submission #21512 on 2025-03-28 at 14:56:21","Hypothesis-driven empirical research is prevalent in synthetic organic chemistry. Traditional workflows rely heavily on the intuition of experimentalists and their ability to and trends or key results in sparse and deficient data to formulate testable hypotheses. An emergent paradigm in chemistry is an approach, in which the data is collected en masse, and the computer-aided analysis then allows to reduce the data and identify meaningful trends. High-Throughput Experimentation and the approach of data rich experimentation allows to experimentally navigate large chemical spaces during reaction optimizations. Despite the success of HTE in industrial applications, it is not nearly as popular in fundamental research, as it is often more important to demonstrate novel findings, rather than a perfectly optimized process. The greatest intellectual challenge of data rich chemical science is to provide smart solutions towards the experiment design, data extraction and modeling to aid synthetic efforts. This document is a compilation of four different projects, which tell a story about the power of synergy between experimental and computational science in chemistry. Chapter 1 of this document contains the description of the data-rich platform for the reaction discovery. A structure-agnostic reaction product identification by means of Liquid Chromatography-Mass Spectrometry approach was realized through a stable isotope fingerprinting strategy. The introduction of traceable isotopic multiplets into the mass spectra by means of selective deuteration of the starting materials allowed to identify the presence of homocoupling or heterocoupling products up to 3 components. Paired with an unsupervised machine learning approach towards decomposing complicated mass spectral datasets, and an automated pattern identification, we were able to achieve a robust discovery platform. Its utility was then demonstrated in the discovery of several previously unknown reactions. Chapter 2 of this document describes the chemoinformatics-guided approach to selection of the phosphoramidite ligand universal training set. A large ligand library was enumerated from its 2D depictions using molli package, described later in this document. The library was then clustered using a k-means clustering approach on the reduced average steric occupancy + average electronic indicator field descriptors to provide a set of compounds termed the phosphoramidite univeral training set. This set was used in a high-throughput screening campaign in the borylative Heck reaction sequence, which allowed to identify highly selective reaction conditions. Chapter 3 of this document describes the development and application of the molli software package. Since the adoption of the rich in silico data oriented strategy, the necessity to handle large molecule collections programmatically increased, as did the need for a modern, consistent and fast interface to those functions. In this section, the advances towards a novel approach to parsing of ChemDraw™ .CDXML les via z-coordinate hinting are described. Benchmarking of the improved GBCA descriptor calculations, as well as algorithmic improvements are listed. Example end-to-end calculation workflows are discussed. Chapter 4 of this document describes the computational approach towards the structural investigation of an unprecedented Pd–Pd dimeric structure. It illustrates the power and the pitfalls of Density Functional Theory study of a complex, which cannot be isolated in its pure form. By means of density functional theory and multireference self-consistent field approaches we were able to establish an openshell nature of one of the key metallic intermediates in the anhydrous Suzuki–Miyaura cross-coupling reaction. So far unprecedented avoided Pd–Pd metallic bonding is described, and the likely hypotheses for the emergence of this phenomenon were formulated."]},{"key":"dc:format","label":"Dc Format","values":["application/pdf"]},{"key":"dc:title","label":"Title","values":["Data-rich experimentation and computer-aided strategies in reaction discovery, selectivity optimization, and structure elucidation"]}]}],"canonical_facts":{"dc:contributor":["Denmark, Scott E","Sarlah, David","Fataftah, Majed S","Pogorelov, Taras V"],"dc:creator":["Shved, Alexander S"],"dc:date":["2024-12","2024-12-05"],"dc:description":["Submission published under a 24 month embargo labeled 'Closed Access', the embargo will last until 2026-12-01","The student, Alexander Shved, accepted the attached license on 2024-12-05 at 05:53.","The student, Alexander Shved, submitted this Dissertation for approval on 2024-12-05 at 06:07.","This Dissertation was approved for publication on 2024-12-05 at 18:55.","DSpace SAF Submission Ingestion Package generated from Vireo submission #21512 on 2025-03-28 at 14:56:21","Hypothesis-driven empirical research is prevalent in synthetic organic chemistry. Traditional workflows rely heavily on the intuition of experimentalists and their ability to and trends or key results in sparse and deficient data to formulate testable hypotheses. An emergent paradigm in chemistry is an approach, in which the data is collected en masse, and the computer-aided analysis then allows to reduce the data and identify meaningful trends. High-Throughput Experimentation and the approach of data rich experimentation allows to experimentally navigate large chemical spaces during reaction optimizations. Despite the success of HTE in industrial applications, it is not nearly as popular in fundamental research, as it is often more important to demonstrate novel findings, rather than a perfectly optimized process. The greatest intellectual challenge of data rich chemical science is to provide smart solutions towards the experiment design, data extraction and modeling to aid synthetic efforts. This document is a compilation of four different projects, which tell a story about the power of synergy between experimental and computational science in chemistry. Chapter 1 of this document contains the description of the data-rich platform for the reaction discovery. A structure-agnostic reaction product identification by means of Liquid Chromatography-Mass Spectrometry approach was realized through a stable isotope fingerprinting strategy. The introduction of traceable isotopic multiplets into the mass spectra by means of selective deuteration of the starting materials allowed to identify the presence of homocoupling or heterocoupling products up to 3 components. Paired with an unsupervised machine learning approach towards decomposing complicated mass spectral datasets, and an automated pattern identification, we were able to achieve a robust discovery platform. Its utility was then demonstrated in the discovery of several previously unknown reactions. Chapter 2 of this document describes the chemoinformatics-guided approach to selection of the phosphoramidite ligand universal training set. A large ligand library was enumerated from its 2D depictions using molli package, described later in this document. The library was then clustered using a k-means clustering approach on the reduced average steric occupancy + average electronic indicator field descriptors to provide a set of compounds termed the phosphoramidite univeral training set. This set was used in a high-throughput screening campaign in the borylative Heck reaction sequence, which allowed to identify highly selective reaction conditions. Chapter 3 of this document describes the development and application of the molli software package. Since the adoption of the rich in silico data oriented strategy, the necessity to handle large molecule collections programmatically increased, as did the need for a modern, consistent and fast interface to those functions. In this section, the advances towards a novel approach to parsing of ChemDraw™ .CDXML les via z-coordinate hinting are described. Benchmarking of the improved GBCA descriptor calculations, as well as algorithmic improvements are listed. Example end-to-end calculation workflows are discussed. Chapter 4 of this document describes the computational approach towards the structural investigation of an unprecedented Pd–Pd dimeric structure. It illustrates the power and the pitfalls of Density Functional Theory study of a complex, which cannot be isolated in its pure form. By means of density functional theory and multireference self-consistent field approaches we were able to establish an openshell nature of one of the key metallic intermediates in the anhydrous Suzuki–Miyaura cross-coupling reaction. So far unprecedented avoided Pd–Pd metallic bonding is described, and the likely hypotheses for the emergence of this phenomenon were formulated."],"dc:format":["application/pdf"],"dc:identifier":["https://hdl.handle.net/2142/127502"],"dc:language":["en","eng"],"dc:rights":["© 2024 Alexander S. Shved"],"dc:subject":["High-throughput Experimentation","Reaction Discovery","Reaction Optimization"],"dc:title":["Data-rich experimentation and computer-aided strategies in reaction discovery, selectivity optimization, and structure elucidation"],"dc:type":["text","Thesis"],"thesis:degree_discipline":["Chemistry"],"thesis:degree_level":["Dissertation"],"thesis:degree_name":["Ph.D."],"thesis:institution_name":["University of Illinois at Urbana-Champaign"]},"updated_at":"2026-07-22T22:25:04Z"}