cs.LGSep 30, 2026

TRACE: Trajectory Selection for Parallel Scaling of Search Agents

Authors: Qisheng Zhou, Zhen Xiong, Qiaoyu Tan

Organizations: New York University Shanghai · New York University

Abstract

Parallel search may generate a correct answer that final-answer voting fails to select. We formulate this consolidation stage as trajectory selection and introduce TRACE (Trajectory Ranking with Aggregated Cross-Rollout Evidence), a lightweight learned selector that ranks completed trajectories using the search evidence behind their answers. TRACE preserves individual query and evidence occurrences, connects rollouts through shared content or document identity, and propagates information across these relations. Each candidate answer then reads the updated states of its own trajectory, preserving retrieval provenance while incorporating evidence from related rollouts. Trained with answer-level supervision over frozen text embeddings, TRACE returns an existing answer without additional search or autoregressive aggregation. One selector per search setting transfers across rollout policies and agent backbones without agent-specific fine-tuning, improving over voting across six WebQA policies and six long-horizon dataset-backbone combinations at K=16K=16. On Qwen2.5-14B Base/SFT WebQA pools, TRACE achieves 45.2/49.2% EM, compared with 43.9/48.0% for the strongest Qwen3-32B generative aggregators. On long-horizon FRAMES, GAIA, and BrowseComp, it reaches 78.6% average accuracy, exceeding majority voting by 3.1 percentage points. On Base WebQA pools, TRACE with only 8 rollouts comes within 0.4 points of majority voting over 64. TRACE also achieves at least 10×10\times higher processing throughput than SolAgg, SummAgg, and AggAgent across all seven WebQA benchmarks. These results show that reusing cross-rollout search evidence provides an effective and efficient alternative to heavyweight generative aggregation for parallel search. Code is available at https://github.com/Jaasssoooonnnnn/TRACE.

Figures & tables

Appendix figures & tables7 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Beyond Parallel Sampling: Diverse Query Initialization for Agentic Search

    Jun 15, 2026Sidhaarth Murali, João Coelho, Jingjie Ning +3Agentic SearchQueries

  2. Argus: Evidence Assembly for Scalable Deep Research Agents

    May 15, 2026Zhen Zhang, Liangcai Su, Zhuo Chen +7Deep Research

  3. SearchAtlas: Analyzing Agentic Search Strategies via Evidential Query Graphs

    Sep 11, 2026Jiacheng Sang, Mengyuan Li, Sanxing Chen +3Agentic SearchQueries