Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 4 of 4 for “"td-learning"”.

  1. Computational and Human Learning Models of Generalized Unsafety

    … markers of generalized stress impair learning of safe cues in stressful environments. Based on this model, chronic problems inhibiting physiological arousal lead to a heightened perception of threat, which involves experiencing anxiety symptoms without any obvious precipitating …

    vt Repository record for Computational and Human Learning Models of Generalized Unsafety (opens in a new tab)

  2. Dynamic Discrete Choice Estimation using Reinforcement Learning with Applications in Online Food Markets

    … datasets. This thesis develops Reinforcement Learning (RL)-based estimation methods to improve the speed and scalability of DDC estimation. The second chapter establishes a theoretical foundation for integrating RL with DDC estimation, emphasizing the shared mathematical structure of Markov …

    cambridge Repository record for Dynamic Discrete Choice Estimation using Reinforcement Learning with Applications in Online Food Markets (opens in a new tab)

  3. Sample-efficient reinforcement learning

    Reinforcement learning has been instrumental in the recent advances made by artificial intelligence agents in various domains. Most of these advances have been abetted by the availability of huge amounts of training data. But, in several practical applications such as those arising in wireless …

    uiuc Repository record for Sample-efficient reinforcement learning (opens in a new tab)

  4. Value learning through Bellman residuals and neural function approximations in deterministic systems

    … Bellman residual minimization is the problem of learning values in a Markov Decision Process (MDP) by optimizing for the Bellman residuals directly without using any heuristic surrogates. Compared to the traditional Approximate Dynamic Programming (ADP) methods, this approach can have both …

    uiuc Repository record for Value learning through Bellman residuals and neural function approximations in deterministic systems (opens in a new tab)