cs.ROSep 30, 2026

CF-JEPA: Improving Robustness of JEPA World Models via Controllability Factorization

Authors: Morgan Byrd, Robert Wright, Sehoon Ha

Organizations: Georgia Institute of Technology, Atlanta, GA, 30308, USA · Georgia Tech Research Institute, Atlanta, GA, 30308, USA

Abstract

Controlling an agent with vision requires being able to separate useful information from irrelevant background information. JEPA-style latent world models seem like a natural approach for this, as they do not perform pixel-level reconstruction; however, they are still sensitive to these distractor signals and experience latent collapse. In this work, we introduce Controllability Factorized JEPA (CF-JEPA), a JEPA-style world model which splits the latent space into controllable and uncontrollable subspaces. This factorization allows us to capture all the distractor information into the uncontrollable region, while we use the control-relevant latent information for our task. With this, we show comparable performance across 2D and 3D control tasks under nominal conditions and improved performance under distracted conditions, where CF-JEPA is the only model that does not experience latent collapse. We also validate our model under distracted conditions for a simulated robot task, highlighting the practical application of such a scheme.

Figures & tables

Explore similar work

CardsList
  1. PhyLatent: Learning Dynamics-Relevant Representations for JEPA World Models

    Aug 6, 2026Xi Zeng, Haojie Ren, Ziying Song +2Latent World ModelsJoint-Embedding Predictive Architectures

  2. D-JEPA: A Decision-Aligned Latent World Model

    Sep 21, 2026Shuaijun Liu, Chengyu Wu, Qifu Wen +5Latent World ModelsWorld Models

  3. Subspace-Decomposed JEPAs: Disentangling Progression and Content in Latent World Models

    May 29, 2026Lucas Thil, Jesse Read, Rim Kaddah +1Joint-Embedding Predictive ArchitecturesLatent Prediction