Back to results

Texas Tech University

Continuous state Q-learning

Abstract

dc:description.abstract

Q-learning is a solution technique developed to solve classical Markov Decision Processes, MDPs. Markov Decision Processes are models for sequential decision making problems and address many classical control problems. In Chapter I, this paper discusses the model and some standard solution techniques used in Markov Decision Processes and its limitations [6]. Q-learning was developed by Watkins to broaden the scope of problems that dynamic programming, MDP techniques, can solve. Classical Q-learning is a model free solution technique and is therefore able to address a variety of poorly modeled decision problems which were unsolvable using standard MDP techniques. Watkins development of Q-learning is based on Markov Decision Processes with discrete action and state spaces. The model and algorithm associated with classical Q-learning are described in Chapter II. To extend the set of problems which can be addressed using Q-learning, Chapter III addresses solution techniques for poorly modeled problems with continuous state and/or action spaces. The model is slightly altered and the algorithm is adjusted to account for the continuous state and action spaces. Numerical example show that continuous Q-learning does determine the optimal policy over time. Ongoing research is being carried on to improve both the current classical Q-learning method and to prove the convergence in the continuous case.

Degree

thesis:*
Name thesis:degree_name
M.S.
Level thesis:degree_level
Masters
Grantor dc:publisher
Texas Tech University
Year dc:date.issued
1999

Author and committee

dc:creator, dc:contributor.*
Author dc:creator
  • Alcorn, Cristy Michele

Subjects

dc:subject × 3

Rights

Language dc:language.iso
eng

Identifiers

dc:identifier.*
Handle dc:identifier.uri
http://hdl.handle.net/2346/20357
OAI identifier oai:identifier
oai:ttu-ir.tdl.org:2346/20357

Chain of custody

source
Harvested from
Texas Technology University
Base URL
ttu-ir.tdl.org/server/oai/request
Last updated
2026-07-24
Source record
OAI-PMH GetRecord
citation

Alcorn, Cristy Michele. Continuous state Q-learning. Masters thesis, Texas Tech University, 1999. http://hdl.handle.net/2346/20357