Back to results

Massachusetts Institute of Technology

Efficient Robustness and Interpretability in Learning and Data-Driven Decision-Making

Abstract

dc:description.abstract

As machine learning algorithms are increasingly developed and deployed in high-stakes applications, ensuring their reliability has become crucial. This thesis introduces algorithmic advancements toward reliability in machine learning, emphasizing two critical dimensions: Robustness and Interpretability. The first part of this thesis focuses on robustness, which guarantees that algorithms deliver stable and predictable performance despite various data uncertainties. We study robustness when learning under diverse sources of data uncertainty, including the fundamental statistical error, as well as data noise and corruption. Our work reveals how these different sources interact and subsequently impact data-driven decisions. We introduce novel distributionally robust optimization approaches, each tailored to specific uncertainty sources. Our findings highlight that protection against one source may increase vulnerability to another. To address this, we develop distributional ambiguity sets that provide holistic robustness against all sources simultaneously. In each setting, we demonstrate that our novel approaches achieve “efficient” robustness, optimally balancing average performance with out-of-sample guarantees. Our novels algorithms are applied to various scenarios, including training robust neural networks, where they significantly outperform existing benchmarks. The second part of the thesis addresses interpretability, a critical attribute for decision-support tools in high-risk settings, which requires that algorithms provide understandable justifications for their decisions. Our work in this part was motivated by data-driven personalized patient treatment—an increasingly sought-after machine learning application. In this reinforcement learning problem, interpretability is crucial: physicians cannot rely on a black-box algorithm for prescribing treatments. We introduce theoretically the problem of learning the most concise discrete representation of a continuous state-space dynamic system. In the patient treatment setting, this corresponds to identifying treatment groups based on the evolving features of patients under treatment. Surprisingly, we prove theoretically that it is statistically possible to learn the most concise representation of a dynamic system solely from observed historic sample path data. We subsequently develop an algorithm, MRL, which learns such a concise representation, thereby enhancing interpretability and tractability.

Degree

thesis:*
Name thesis:degree_name
Doctoral
Department dc:contributor.department
Massachusetts Institute of Technology. Operations Research Center
Grantor dc:publisher
Massachusetts Institute of Technology
Year dc:date.issued
2024

Author and committee

dc:creator, dc:contributor.*
Author dc:creator
  • Bennouna, Mohammed Amine
Advisor dc:contributor.advisor
  • Van Parys, Bart P. G.

Rights

dc:rights
Statement dc:rights
  • Attribution-NonCommercial-NoDerivatives 4.0 International (CC BY-NC-ND 4.0)
  • Copyright retained by author(s)

Identifiers

dc:identifier.*
Handle dc:identifier.uri
https://hdl.handle.net/1721.1/155498
OAI identifier oai:identifier
oai:dspace.mit.edu:1721.1/155498

Chain of custody

source
Harvested from
MIT
Base URL
dspace.mit.edu/oai/request
Last updated
2026-07-22
Source record
OAI-PMH GetRecord
related terms
citation

Bennouna, Mohammed Amine. Efficient Robustness and Interpretability in Learning and Data-Driven Decision-Making. Massachusetts Institute of Technology, 2024. https://hdl.handle.net/1721.1/155498