Massachusetts Institute of Technology
Scalable Representation Learning: On Data-scarcity, Uncertainty and Symmetry
Abstract
dc:description.abstractDeep learning has experienced remarkable success in recent years, leading to significant advancements in various fields such as vision, natural language generation, complex game play, as well as solving difficult scientific problems such as predicting protein folding. Despite these successes, traditional deep learning faces fundamental challenges limiting their scalability and effectiveness. These challenges include the necessity for extensive labeled datasets, the lack of trustworthiness due to model overconfidence, and difficulties in generalizing to new, unseen data. In this thesis, our primary goal is to tackle these issues by introducing novel tools and methods that augment traditional deep learning. We explore various strategies for solving the main bottlenecks of traditional deep learning, which includes incorporating prior known symmetries and inductive biases of the problem, utilizing Bayesian and ensemble methods, and leveraging abundance of unlabeled data in a representation learning framework. We discuss and demonstrate practical applications of these novel tools in diverse domains including vision, photonics, material science and neuroscience.
Degree
thesis:*- Name thesis:degree_name
- Doctoral
- Department dc:contributor.department
- Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science
- Grantor dc:publisher
- Massachusetts Institute of Technology
- Year dc:date.issued
- 2024
Author and committee
dc:creator, dc:contributor.*- Author dc:creator
-
- Loh, Charlotte Chang Le
- Advisor dc:contributor.advisor
-
- Soljačić, Marin
Rights
dc:rights- Statement dc:rights
-
- Attribution-NonCommercial-NoDerivatives 4.0 International (CC BY-NC-ND 4.0)
- Copyright retained by author(s)
- Licence dc:rights.uri
Identifiers
dc:identifier.*- Handle dc:identifier.uri
- https://hdl.handle.net/1721.1/156614
- OAI identifier oai:identifier
- oai:dspace.mit.edu:1721.1/156614