In this thesis, we contribute to new directions within Reinforcement Learning, which are important for many practical applications such as the control of biomechanical models. We deepen the mathematical foundations of Reinforcement Learning by deriving theoretical results inspired by classical optimal control theory. In our derivations, Deep Reinforcement Learning serves as our starting point. Based on its working principle, we derive a new type of Reinforcement Learning framework by replacing the neural network by a suitable ordinary differential equation. Coming up with profound mathematical results within this differential equation based framework turns out to be a challenging research task, which we address in this thesis. Especially the derivation of optimality conditions takes a central role in our investigation. We establish new optimality conditions tailored to our specific situation and analyze a resulting gradient based approach. Finally, we illustrate the power, working principle and versatility of this approach by performing control tasks in the context of a navigation in the two dimensional plane, robot motions, and actuations of a human arm model.


    Access

    Download


    Export, share and cite



    Title :

    Differential Equation Based Framework for Deep Reinforcement Learning


    Contributors:

    Publication date :

    2021-01-01


    Remarks:

    Fraunhofer ITWM


    Type of media :

    Theses


    Type of material :

    Electronic Resource


    Language :

    English



    Classification :

    DDC:    629



    Framework for control and deep reinforcement learning in traffic

    Wu, Cathy / Parvate, Kanaad / Kheterpal, Nishant et al. | IEEE | 2017


    Deep Reinforcement Learning

    Huang, Xiaowei / Jin, Gaojie / Ruan, Wenjie | Springer Verlag | 2012


    DEEP REINFORCEMENT LEARNING FOR A GENERAL FRAMEWORK FOR MODEL-BASED LONGITUDINAL CONTROL

    PATHAK SHASHANK / BAG SUVAM / NADKARNI VIJAY JAYANT | European Patent Office | 2020

    Free access


    DEEP REINFORCEMENT LEARNING FOR A GENERAL FRAMEWORK FOR MODEL-BASED LONGITUDINAL CONTROL

    PATHAK SHASHANK / BAG SUVAM / NADKARNI VIJAY JAYANT | European Patent Office | 2020

    Free access