Back to results

York University

Exploiting Reward Machines with Deep Reinforcement Learning in Continuous Action Domains

Abstract

dc:description.abstract

Deep reinforcement learning can solve real-world robot control problems, such as autonomous driving and robotic arm manipulation. In deep reinforcement learning, an agent does not know the problem description and learns the optimal solution through trial-and-error. This method brings two major challenges when solving real-world problems: partial observability and learning efficiency. In this thesis, we address these two challenges and extend previous work. First, we use reward machines to address the problem of partial observability. Then, we focus on finding the existing cutting-edge deep reinforcement learning algorithms and integrating them with reward machines to enhance the learning efficiency. To test the performance of all the algorithms, we proposed a series of different tasks that can be used to mimic real-world robot control problems. Finally, based on the test results, we compare the performance of all the algorithms and analyze their advantages and disadvantages.

Author and committee

dc:creator, dc:contributor.*
Author dc:creator
  • Sun, Haolin
Advisor dc:contributor.advisor
  • Lesperance, Yves

Subjects

dc:subject × 1

Rights

dc:rights
Statement dc:rights
  • Author owns copyright, except where explicitly noted. Please contact the author directly with licensing requests.
Language dc:language
en

Identifiers

dc:identifier.*
Handle dc:identifier.uri
http://hdl.handle.net/10315/41039
OAI identifier oai:identifier
oai:yorkspace.library.yorku.ca:10315/41039

Chain of custody

source
Harvested from
York University
Base URL
yorkspace.library.yorku.ca/oai/request
Last updated
2026-07-24
Source record
OAI-PMH GetRecord
related terms
citation

Sun, Haolin. Exploiting Reward Machines with Deep Reinforcement Learning in Continuous Action Domains. 2023. http://hdl.handle.net/10315/41039