Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 2 of 2 for “"Temporal Difference Reinforcement Learning"”.
-
Machine Learning Simulation: Torso Dynamics of Robotic Biped
… of an appropriate control scheme difficult. A temporal difference reinforcement learning method known as Q-learning develops complex control policies through environmental exploration and exploitation. As a proof of concept, Q-learning was applied through simulation to a benchmark single …
-
Competitive and Noncompetitive Credit Assignment to Reward Cues in the Lateral Orbitofrontal Cortex
… credit along a competitive-noncompetitive learning continuum, with noncompetitive learning dominating the early stages of training and competitive learning taking increasing control over training. Moreover, encoding of competitive credit emerges late in the cue epoch and gradually migrates …