Skip to main content
Aggregate arXiv cs.AI 人工智能 17 Aug 2026 - 13:30

Reward Machines for Signal Temporal Logic

RSS 官方收录 · 可信分层展示

关键摘要

arXiv:2608.…

  • 13625v1 Announce Type: new Abstract: Signal temporal logic (STL) provi…
  • Control synthesis from STL specifications is of interest since manual …
  • Moreover, many modern autonomous and AI-enabled systems lack accurate …

摘要引擎:抽取

正文提要

arXiv:2608.13625v1 Announce Type: new Abstract: Signal temporal logic (STL) provides a formal language for specifying real-time properties of real-valued observations, along with a quantitative robustness score for monitoring satisfaction. Control synthesis from STL specifications is of interest since manual controller design becomes infeasible as real-world systems grow in complexity. Moreover, many modern autonomous and AI-enabled systems lack accurate and complete system models, which makes optimization-based synthesis approaches unsuitable and motivates learning-based control. Prior work uses STL robustness scores as rewards in reinforcement learning (RL) to obtain control policies satisfying given specifications; however, robustness depends on execution history, leading to intractable state space expansion for general long-horizon specifications with arbitrarily nested temporal operators. This work introduces a novel automata-based approach that provides an efficient memory mechanism and associated Markovian rewards suitable for RL frameworks. Our approach constructs a timed alternating automaton from the given STL specifications, augments the state space with automaton locations and clock valuations, and derives rewards from the automaton acceptance condition. We empirically demonstrate that our approach learns policies that achieve higher robustness scores and satisfaction rates than those learned by existing approaches using robustness-based rewards.

来源:https://arxiv.org/abs/2608.13625

打开官方原文 站点原文页 可信分区 本信源更多 今日简报 分享图 RSS 稍后再看列表