{"id":{"repo_id":"mit","oai_identifier":"oai:dspace.mit.edu:1721.1/156327"},"canonical_url":"https://search.dev.ndltd.org/etd/mit/oai:dspace.mit.edu:1721.1/156327","repository":{"repo_id":"mit","name":"MIT","base_url":"https://dspace.mit.edu/oai/request"},"display":{"title":"Causal Structure Learning through Double Machine Learning","abstract":"Learning the causal structure of a system solely from observational data is a fundamental yet intricate task with numerous applications across various fields, including economics, earth sciences, biology, and medicine. This task is challenging due to several reasons: i) observational data alone, as opposed to interventional data, do not characterize the properties of the system under interventions on variables; therefore, it contains information on correlation instead of cause-effect one, ii) unobserved confounders may induce biases for the algorithms, leading to false causal inferences instead of revealing the correct causal structure, like a hidden common confounder, iii) the number of potential underlying structures increases super-exponentially with the number of variables, posing significant statistical and computational challenges and iv) the identifiability problem arises because multiple causal models can yield the same observational distribution, making it impossible to conclusively determine the true structure. In this thesis, we focus on the partial identification of underlying p causal structure from observational data under minimal assumptions necessary for causal identification. To this end, inspired by Debiasd/Double machine learning machinery, we introduce efficient practical doubly robust algorithms enjoying fast √n-semiparametric convergence rate for three different tasks: (1) Finding the direct causes of the target variable under cyclic and unseen confounded high dimensional data with nonlinear structures, (2) Testing Granger causality and therefore causal structure identification from temporal data, and (3) Estimation of counterfactual prediction function in generalized nonlinear Instrumental Variables regression problem. As a natural use case, we tackle the offline policy evaluation of the confounded contextual bandit problem, when actions, contexts, and rewards have common unobserved confounding. By matching the upper bounds with the unconfounded contextual bandit settings, our algorithm is proven to achieve optimal sample complexity.","abstract_html":"Learning the causal structure of a system solely from observational data is a fundamental yet intricate task with numerous applications across various fields, including economics, earth sciences, biology, and medicine. This task is challenging due to several reasons: i) observational data alone, as opposed to interventional data, do not characterize the properties of the system under interventions on variables; therefore, it contains information on correlation instead of cause-effect one, ii) unobserved confounders may induce biases for the algorithms, leading to false causal inferences instead of revealing the correct causal structure, like a hidden common confounder, iii) the number of potential underlying structures increases super-exponentially with the number of variables, posing significant statistical and computational challenges and iv) the identifiability problem arises because multiple causal models can yield the same observational distribution, making it impossible to conclusively determine the true structure. In this thesis, we focus on the partial identification of underlying p causal structure from observational data under minimal assumptions necessary for causal identification. To this end, inspired by Debiasd/Double machine learning machinery, we introduce efficient practical doubly robust algorithms enjoying fast √n-semiparametric convergence rate for three different tasks: (1) Finding the direct causes of the target variable under cyclic and unseen confounded high dimensional data with nonlinear structures, (2) Testing Granger causality and therefore causal structure identification from temporal data, and (3) Estimation of counterfactual prediction function in generalized nonlinear Instrumental Variables regression problem. As a natural use case, we tackle the offline policy evaluation of the confounded contextual bandit problem, when actions, contexts, and rewards have common unobserved confounding. By matching the upper bounds with the unconfounded contextual bandit settings, our algorithm is proven to achieve optimal sample complexity.","abstract_has_math":false,"creators":["Soleymani, Ashkan"],"institution":"Massachusetts Institute of Technology","degree_name":"Master","degree_level":null,"degree_discipline":null,"degree_department":"Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science","school":null,"contributors":[],"advisors":["Jaillet, Patrick"],"committee_chairs":[],"committee_members":[],"year":2024,"date_issued":"2024-05","date_published":"2024-05","updated_at":"2026-07-22T22:21:27Z","subjects":[],"languages":[],"rights":["In Copyright - Educational Use Permitted","Copyright retained by author(s)"],"rights_urls":["https://rightsstatements.org/page/InC-EDU/1.0/"],"identifier_entries":[]},"links":{"outbound_url":"https://hdl.handle.net/1721.1/156327","outbound_label":"Handle","outbound_source":"dc:identifier.uri"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor.advisor","label":"Advisor","values":["Jaillet, Patrick"]},{"key":"dc:contributor.department","label":"Department","values":["Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science"]},{"key":"dc:creator","label":"Author","values":["Soleymani, Ashkan"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.accessioned","label":"Dc Date Accessioned","values":["2024-08-21T18:57:01Z"]},{"key":"dc:date.available","label":"Dc Date Available","values":["2024-08-21T18:57:01Z"]},{"key":"dc:date.issued","label":"Date","values":["2024-05"]},{"key":"dc:publisher","label":"Institution","values":["Massachusetts Institute of Technology"]},{"key":"dc:type","label":"Dc Type","values":["Thesis"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Master","Master of Science in Electrical Engineering and Computer Science"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:rights","label":"Dc Rights","values":["In Copyright - Educational Use Permitted","Copyright retained by author(s)"]},{"key":"dc:rights.uri","label":"Rights URI","values":["https://rightsstatements.org/page/InC-EDU/1.0/"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.uri","label":"Identifier URI","values":["https://hdl.handle.net/1721.1/156327"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["Learning the causal structure of a system solely from observational data is a fundamental yet intricate task with numerous applications across various fields, including economics, earth sciences, biology, and medicine. This task is challenging due to several reasons: i) observational data alone, as opposed to interventional data, do not characterize the properties of the system under interventions on variables; therefore, it contains information on correlation instead of cause-effect one, ii) unobserved confounders may induce biases for the algorithms, leading to false causal inferences instead of revealing the correct causal structure, like a hidden common confounder, iii) the number of potential underlying structures increases super-exponentially with the number of variables, posing significant statistical and computational challenges and iv) the identifiability problem arises because multiple causal models can yield the same observational distribution, making it impossible to conclusively determine the true structure. In this thesis, we focus on the partial identification of underlying p causal structure from observational data under minimal assumptions necessary for causal identification. To this end, inspired by Debiasd/Double machine learning machinery, we introduce efficient practical doubly robust algorithms enjoying fast √n-semiparametric convergence rate for three different tasks: (1) Finding the direct causes of the target variable under cyclic and unseen confounded high dimensional data with nonlinear structures, (2) Testing Granger causality and therefore causal structure identification from temporal data, and (3) Estimation of counterfactual prediction function in generalized nonlinear Instrumental Variables regression problem. As a natural use case, we tackle the offline policy evaluation of the confounded contextual bandit problem, when actions, contexts, and rewards have common unobserved confounding. By matching the upper bounds with the unconfounded contextual bandit settings, our algorithm is proven to achieve optimal sample complexity."]},{"key":"dc:description.degree","label":"Dc Description Degree","values":["S.M."]},{"key":"dc:title","label":"Title","values":["Causal Structure Learning through Double Machine Learning"]}]}],"canonical_facts":{"dc:contributor.advisor":["Jaillet, Patrick"],"dc:contributor.department":["Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science"],"dc:creator":["Soleymani, Ashkan"],"dc:date.accessioned":["2024-08-21T18:57:01Z"],"dc:date.available":["2024-08-21T18:57:01Z"],"dc:date.issued":["2024-05"],"dc:description.abstract":["Learning the causal structure of a system solely from observational data is a fundamental yet intricate task with numerous applications across various fields, including economics, earth sciences, biology, and medicine. This task is challenging due to several reasons: i) observational data alone, as opposed to interventional data, do not characterize the properties of the system under interventions on variables; therefore, it contains information on correlation instead of cause-effect one, ii) unobserved confounders may induce biases for the algorithms, leading to false causal inferences instead of revealing the correct causal structure, like a hidden common confounder, iii) the number of potential underlying structures increases super-exponentially with the number of variables, posing significant statistical and computational challenges and iv) the identifiability problem arises because multiple causal models can yield the same observational distribution, making it impossible to conclusively determine the true structure. In this thesis, we focus on the partial identification of underlying p causal structure from observational data under minimal assumptions necessary for causal identification. To this end, inspired by Debiasd/Double machine learning machinery, we introduce efficient practical doubly robust algorithms enjoying fast √n-semiparametric convergence rate for three different tasks: (1) Finding the direct causes of the target variable under cyclic and unseen confounded high dimensional data with nonlinear structures, (2) Testing Granger causality and therefore causal structure identification from temporal data, and (3) Estimation of counterfactual prediction function in generalized nonlinear Instrumental Variables regression problem. As a natural use case, we tackle the offline policy evaluation of the confounded contextual bandit problem, when actions, contexts, and rewards have common unobserved confounding. By matching the upper bounds with the unconfounded contextual bandit settings, our algorithm is proven to achieve optimal sample complexity."],"dc:description.degree":["S.M."],"dc:identifier.uri":["https://hdl.handle.net/1721.1/156327"],"dc:publisher":["Massachusetts Institute of Technology"],"dc:rights":["In Copyright - Educational Use Permitted","Copyright retained by author(s)"],"dc:rights.uri":["https://rightsstatements.org/page/InC-EDU/1.0/"],"dc:title":["Causal Structure Learning through Double Machine Learning"],"dc:type":["Thesis"],"thesis:degree_name":["Master","Master of Science in Electrical Engineering and Computer Science"]},"updated_at":"2026-07-22T22:21:27Z"}