cs.ROSep 29, 2026

World4Scorer: Outcome-Grounded World Modeling for Autonomous Driving

Authors: Jieyuan Pei, Meiyi Lu, Sining Ang, Yubo Zhao, Zhangyi Hu, Mingwei Xu, Haokai Ding, Wei Li, +7 more

Organizations: Institute for AI Industry Research (AIR), Tsinghua University · HiThink Research · The Hong Kong University of Science and Technology (Guangzhou) · Zhejiang University · University of Science and Technology of China · SMBU · University of Washington · Mohamed bin Zayed University of Artificial Intelligence · Zhejiang University of Technology · Southeast University · Changan Automobile

Abstract

Autonomous driving requires choosing a safe and efficient plan as surrounding traffic evolves. Generate-and-select planners propose multiple trajectories and score them for execution, and they have outperformed representative direct-prediction baselines on NAVSIM. Their scorer must compare plans that were never executed. Driving logs record the future of only the executed trajectory, so matching the logged future can leave predictions for the alternatives unconstrained; a simulator, in contrast, can label the outcome of every candidate. We introduce World4Scorer, which builds the scorer as a trajectory-conditioned JEPA-style predictor: it predicts a state for each candidate and reads the candidate's scores from that state. Simulator outcome labels supervise the states of all candidates, and the observed future of the executed trajectory anchors the predictor to real scene evolution. Because one predictor produces every candidate's state, the anchor can constrain shared parameters used to score unexecuted plans, while the future itself is needed only during training. Generated candidates mostly score well, so a scene-matched bank adds low-scoring plans to the outcome supervision; framewise choices can conflict, so inertial re-ranking keeps consecutive selections consistent. World4Scorer achieves state-of-the-art NAVSIM-v2 performance and a strong adapted-system result on closed-loop Bench2Drive. With the LeWM world model and planning budget fixed, outcome-based scoring also improves manipulation planning on the OGBench-Cube benchmark.

Figures & tables

Explore similar work

CardsList
  1. Test-Time Trajectory Optimization for Autonomous Driving

    Jun 5, 2026Yihong Xu, Eloi Zablocki, Yuan Yin +4Autonomous DrivingRobust Trajectory

  2. DriveVA: Video Action Models are Zero-Shot Drivers

    Apr 5, 2026Mengmeng Liu, Diankun Zhang, Jiuming Liu +7Autonomous DrivingDrives

  3. Driving on Memory

    Aug 31, 2026Christian Löwens, Thorben Funke, Alexandru Paul ConduracheCausality-Aware End-To-End Autonomous Driving