cs.AIOct 7, 2026

RECAST: Learning to Compute the Right Context through Adaptive Evidence Routing

Authors: Yilun Hao, Krishna Sayana, Isabella Ye, James S Ren, Sukhdeep Sodhi, Craig Boutilier, Chuchu Fan

Organizations: MIT · Google Research

Abstract

Large language models are increasingly applied to tasks grounded in long, heterogeneous information sources. Conventional Retrieval-Augmented Generation (RAG) relies on fixed similarity-based retrieval, while agentic variants adapt queries and tool use but remain largely retrieval-centric. However, in many tasks, the evidence required for a solution is not explicitly present in any single source item. Instead, it must be derived through filtering, aggregation, or computation across multiple source items. In this work, we introduce RECAST (Routing Evidence through Computation, Access, and Synthesized Tools), a learned framework that formulates evidence construction as a sequential decision process over heterogeneous retrieval and computation operations, allowing evidence to be actively derived rather than merely retrieved. A lightweight RouterLM iteratively selects and formulates primitive operations or specifies customized operations for a frozen CompilerLM to translate into executable code. Once it judges the evidence sufficient, RouterLM passes the accepted evidence to a frozen AnswerLM to produce the final solution. We train RouterLM with supervised fine-tuning (SFT) followed by group relative policy optimization (GRPO). Across six heterogeneous benchmark families, RECAST achieves a mean success rate of 75.6%, outperforming the strongest large-model baseline by 15.9%. Moreover, training enables the Qwen3.5-9B RouterLM to outperform a training-free Gemini 3.5 Flash RouterLM by 5.0%. On three held-out benchmarks, RECAST improves over the strongest baseline by 15.0% on average, demonstrating strong zero-shot generalization across tasks and heterogeneous source representations.

Figures & tables

Appendix figures & tables9 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. MASS-RAG: Multi-Agent Synthesis Retrieval-Augmented Generation

    Apr 20, 2026Xingchen Xiao, Heyan Huang, Runheng Liu +1Agentic Retrieval-Augmented Generation SystemsHievi-Rag

  2. When to Retrieve During Reasoning: Adaptive Retrieval for Large Reasoning Models

    Apr 29, 2026Dongxin Guo, Jikun Wu, Siu Ming YiuLarge Reasoning ModelsDeepseek

  3. REVA: Reusable Evidence View Aggregation for Context-Efficient RAG Serving

    Sep 11, 2026Tuan Nguyen, Qiran Hu, Banruo Liu +3Retrieval-Augmented Generation PipelinesInteraction History