Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 2 of 2 for “"Temporal Difference Reinforcement Learning"”.

  1. Machine Learning Simulation: Torso Dynamics of Robotic Biped

    … of an appropriate control scheme difficult. A temporal difference reinforcement learning method known as Q-learning develops complex control policies through environmental exploration and exploitation. As a proof of concept, Q-learning was applied through simulation to a benchmark single …

    vt Repository record for Machine Learning Simulation: Torso Dynamics of Robotic Biped (opens in a new tab)

  2. Competitive and Noncompetitive Credit Assignment to Reward Cues in the Lateral Orbitofrontal Cortex

    … credit along a competitive-noncompetitive learning continuum, with noncompetitive learning dominating the early stages of training and competitive learning taking increasing control over training. Moreover, encoding of competitive credit emerges late in the cue epoch and gradually migrates …

    cuny-grad Repository record for Competitive and Noncompetitive Credit Assignment to Reward Cues in the Lateral Orbitofrontal Cortex (opens in a new tab)