Skip to main content
Aggregate arXiv cs.AI 人工智能 28 Aug 2026 - 14:30

Predicting Consequences and Reinforcing Navigation Policies with Latent World Models

RSS 官方收录 · 可信分层展示

关键摘要

arXiv:2608.…

  • 26190v1 Announce Type: new Abstract: World models enable agents to rea…
  • In this work, we propose a compatibility prediction Latent World Model…
  • Our key insight is that spatial proximity correlates with latent featu…

摘要引擎:抽取

正文提要

arXiv:2608.26190v1 Announce Type: new Abstract: World models enable agents to reason about future outcomes and learn policies from their knowledge of state transition, but existing approaches primarily focus on reconstructing future observations or features, which introduces unnecessary complexity and limits their effectiveness for decision making. In this work, we propose a compatibility prediction Latent World Model (LWM) for robot navigation that predicts action-conditioned latent feature compatibility rather than reconstructing observations. Our key insight is that spatial proximity correlates with latent feature similarity, enabling action consequences to be evaluated directly in latent space. To support counterfactual training, our model leverages action sequences sampled across trajectories and learns to predict which sequences lead closer to the goal. Furthermore, we demonstrate how the learned world model can supervise policy learning from unlabeled video data and further improve policies through reinforcement learning entirely within the world model. This imagination-driven framework eliminates the need for action annotations and additional environment interaction. Extensive experiments on multiple real-world robot navigation datasets show that our approach significantly outperforms prior world model and imitation learning methods in prediction accuracy, policy learning, and real-world navigation performance. The code, pretrained models, and additional materials are available at https://wzm206.github.io/latent-world-model-nav.

来源:https://arxiv.org/abs/2608.26190

打开官方原文 站点原文页 可信分区 本信源更多 今日简报 分享图 RSS 稍后再看列表