Back to results

University of Illinois - Chicago

Subdominance Minimization: A Satisficing Perspective on Imitation Learning

Abstract

dc:description

Human decision-making often hinges on a delicately-balanced interplay between multiple, potentially conflicting objectives. However, prevailing imitation learning methods tend to prioritize optimizing a single imitation objective. This myopic focus on a singular objective frequently leads to unintended and undesirable behaviors in learned models. For example, an autonomous vehicle prioritizing travel time over adherence to traffic laws, or indeed a language model learning to generate increasingly creative responses at the expense of factual accuracy. We instead adopt the recently-proposed notion of subdominance which, given an arbitrary number of cost objectives to minimize simultaneously, quantifies the cost of choosing one solution relative to a reference solution. We present a novel focused satisficing approach to imitation learning, using subdominance to seek imitator policies that the demonstrator may consider acceptable, rather than optimal. We leverage both offline and online subdominance minimization to focus policy learning on parts of the demonstrator's trajectories which are hardest to imitate. In the domain of offline reinforcement learning, we present another approach for training decision transformers offline via subdominance minimization which allows us to learn autoregressive policies even in the absence of a ground truth reward function. Finally, we present an approach for language-conditioned, multitask imitation learning, where we leverage learned instruction representations to train language-conditioned subdominance minimizing policies.

Author and committee

dc:creator, dc:contributor.*
Author dc:creator
  • Rushit Shah (24400088)

Subjects

dc:subject × 6

Rights

dc:rights
Statement dc:rights
  • In Copyright
  • Open Access after 2028-05-01

Identifiers

dc:identifier.*
OAI identifier oai:identifier
oai:figshare.com:article/32995130

Chain of custody

source
Harvested from
University of Illinois - Chicago
Base URL
api.figshare.com/v2/oai
Last updated
2026-07-27
Source record
OAI-PMH GetRecord
citation

Rushit Shah (24400088). Subdominance Minimization: A Satisficing Perspective on Imitation Learning. 2026. https://doi.org/10.25417/uic.32995130.v1