SwePub
Sök i SwePub databas

  Extended search

Träfflista för sökning "WFRF:(Andersson Pontus 1995) "

Search: WFRF:(Andersson Pontus 1995)

  • Result 1-1 of 1
Sort/group result
   
EnumerationReferenceCoverFind
1.
  • Önnheim, Magnus, 1985, et al. (author)
  • Reinforcement Learning Informed by Optimal Control
  • 2019
  • In: Lecture Notes in Computer Science. - Cham : Springer International Publishing. - 0302-9743 .- 1611-3349. ; 11731, s. 403-407
  • Conference paper (peer-reviewed)abstract
    • Model-free reinforcement learning has seen tremendous advances in the last few years, however practical applications of pure reinforcement learning are still limited by sample inefficiency and the difficulty of giving robustness and stability guarantees of the proposed agents. Given access to an expert policy, one can increase sample efficiency by in addition to learning from data, and also learn from the experts actions for safer learning. In this paper we pose the question whether expert learning can be accelerated and stabilized if given access to a family of experts which are designed according to optimal control principles, and more specifically, linear quadratic regulators. In particular we consider the nominal model of a system as part of the action space of a reinforcement learning agent. Further, using the nominal controller, we design customized reward functions for training a reinforcement learning agent, and perform ablation studies on a set of simple benchmark problems.
  •  
Skapa referenser, mejla, bekava och länka
  • Result 1-1 of 1
Type of publication
conference paper (1)
Type of content
peer-reviewed (1)
Author/Editor
Önnheim, Magnus, 198 ... (1)
Jirstrand, Mats, 196 ... (1)
Gustavsson, Emil, 19 ... (1)
Andersson, Pontus, 1 ... (1)
University
Chalmers University of Technology (1)
Language
English (1)
Research subject (UKÄ/SCB)
Natural sciences (1)
Engineering and Technology (1)
Social Sciences (1)
Year

Kungliga biblioteket hanterar dina personuppgifter i enlighet med EU:s dataskyddsförordning (2018), GDPR. Läs mer om hur det funkar här.
Så här hanterar KB dina uppgifter vid användning av denna tjänst.

 
pil uppåt Close

Copy and save the link in order to return to this view