Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 8 of 8 for “"Markov decision processes (MDP)"”.
-
Reinforced co-learning for semi-supervised ranking
… on the basis of treating ranking problems as Markov decision processes (MDP). We name our approach ""reinforced co-learning"" because the two modules are iteratively optimized and affect each other while training. When training the classifier module, we use the reinforcement module to give …
-
Acceleration of Iterative Methods for Markov Decision Processes
This research focuses on Markov Decision Processes (MDP). MDP is one of the most important and challenging areas of Operations Research. Every day people make many decisions: today's decisions impact tomorrow's and tomorrow's will impact the ones made the day after. Problems in Engineering, …
-
Long-term Comparative Effectiveness of Rheumatoid Arthritis Treatment Strategies
… permanent joint damage. In this thesis we use Markov decision processes (MDP) as an innovative approach to identify the optimal timing of biologics in RA. The results from this analysis have significant policy, clinical and methodological implications. This work provides important insights into …
-
Representation Learning for Agents in Non-Markovian Environments
… from observation data. Simplified models such as Markov Decision Processes (MDP), which assume a fully observable state and independence of the future from the past given current observations, are widely employed. However, such an assumption is commonly violated in practical applications as …
-
Representation Learning for Agents in Non-Markovian Environments
… from observation data. Simplified models such as Markov Decision Processes (MDP), which assume a fully observable state and independence of the future from the past given current observations, are widely employed. However, such an assumption is commonly violated in practical applications as …
-
Generic Reinforcement Learning Beyond Small MDPs
… automatically reduce a complex environment to a Markov Decision Process (MDP) by finding a map which aggregates similar histories into the states of an MDP. The primary motivation behind this thesis is to build FRL agents that work in practice, both for larger environments and larger classes of …
-
Generic Reinforcement Learning Beyond Small MDPs
… automatically reduce a complex environment to a Markov Decision Process (MDP) by finding a map which aggregates similar histories into the states of an MDP. The primary motivation behind this thesis is to build FRL agents that work in practice, both for larger environments and larger classes of …