首页    期刊浏览 2024年12月04日 星期三
登录注册

文章基本信息

  • 标题:Reinforcement Learning with Temporal Logic Constraints ⁎
  • 本地全文:下载
  • 作者:Bengt Lennartson ; Qing-Shan Jia
  • 期刊名称:IFAC PapersOnLine
  • 印刷版ISSN:2405-8963
  • 出版年度:2020
  • 卷号:53
  • 期号:4
  • 页码:485-492
  • DOI:10.1016/j.ifacol.2021.04.044
  • 语种:English
  • 出版社:Elsevier
  • 摘要:AbstractReinforcement learning (RL) is an agent based AI learning method, where learning and optimization are combined. Dynamic programming is then performed iteratively, based on reward and next state observations from the system to be controlled. A brief survey of RL is given, followed by an evaluation of a recently proposed method to include temporal logic safety and liveness guarantees in RL, here combined with classical performance optimization. RL is based on Markov decision processes (MDPs), and to reduce the number of observations from the system, a modular MDP framework is proposed. In the learning process, it is then assumed that some parts of the system are represented by known MDP models, while other parts can be estimated by observations from the real system. Local information from the modular system may then be used to reduce the computational complexity, especially in the handling of safety properties.
  • 关键词:Keywordsreinforcement learningadaptiontemporal logic specificationsmodular systems
国家哲学社会科学文献中心版权所有