cs.ROOct 8, 2026

LiteNWM: Efficient Latent World Models for Onboard Visual Navigation in the Wild

Authors: Linkai Liu, Yuntian Zhang, Zhenshan Bing, Chen Chen, Lingjuan Lyu, Shangguang Wang, Mengwei Xu, Dongqi Cai

Organizations: Nanjing University · Imperial College London · Sony AI · Beijing University of Posts and Telecommunications

Abstract

Direct visual navigation policies generate trajectories efficiently but do not explicitly evaluate their future consequences. Generative navigation world models provide this foresight through visual rollouts, which are costly when evaluating multiple candidates. We present LiteNWM, a latent navigation world model that shares visual encoding across candidates and jointly predicts their action-conditioned future representations at multiple horizons, while a learned scorer uses these predictions to select trajectories. In offline evaluations on RECON, SCAND, and SACSoN, LiteNWM reduces macro-averaged trajectory error by 17.56% relative to NoMaD+NWM-XL and achieves a 128.00-fold end-to-end speedup on an RTX 5090. The same evaluator transfers from NoMaD to MBRA without proposer-specific retraining, reducing MBRA's macro-averaged trajectory error by 16.2%. In real-robot experiments in unseen indoor and outdoor environments, LiteNWM improves navigation success from 43.3% to 83.3% relative to NoMaD. These results demonstrate that LiteNWM can be deployed for future-aware planning and closed-loop navigation on a physical robot.

Figures & tables

Appendix figures & tables3 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. NavWAM: A Navigation World Action Model for Goal-Conditioned Visual Navigation

    Jun 11, 2026Daichi Azuma, Taiki Miyanishi, Koya Sakamoto +6World Models for RoboticsRobot Navigation

  2. NavWM: A Unified Navigation World Model for Foresight-Driven Planning

    Jun 23, 2026Yanghong Mei, Longteng Guo, Ming-Ming Yu +3Visual NavigationWorld Models for Robotics