Abstract
dc:description.abstractGenerative AI is a field that is rapidly developing and growing in scale. As research in this area shifts to building on large-scale foundation models and powerful architectures, careful thought has to go into adapting these models to new domains and tasks. The work in this thesis demonstrates novel approaches to adapting large-scale generative models and architectures to specific applications in virtual try-on, conceptual art, and domain-specific image classification. In addition to the technical contributions, this thesis explores broader open questions about domain-specific generative models; for example, how can we carefully construct our training data to mitigate bias? What do human-in-the-loop methods for creative generative AI look like in practice? To what extent are large-scale vision-language models useful for traditionally image-only tasks?
Degree
thesis:*- Name thesis:degree_name
- Doctoral
- Department dc:contributor.department
- Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science
- Grantor dc:publisher
- Massachusetts Institute of Technology
- Year dc:date.issued
- 2023
Author and committee
dc:creator, dc:contributor.*- Author dc:creator
-
- Lewis, Kathleen M.
- Advisors dc:contributor.advisor
-
- Guttag, John V.
- Durand, Frédo
Rights
dc:rights- Statement dc:rights
-
- In Copyright - Educational Use Permitted
- Copyright retained by author(s)
- Licence dc:rights.uri
Identifiers
dc:identifier.*- Handle dc:identifier.uri
- https://hdl.handle.net/1721.1/152830
- OAI identifier oai:identifier
- oai:dspace.mit.edu:1721.1/152830