cs.CLSep 24, 2026

No More Free Lunch: Corpus Task Complexity Matters as Corpora Grow

Authors: Prasann Singhal, Amanda Bertsch, Jacob Steinhardt, Sewon Min

Organizations: UC Berkeley · Carnegie Mellon University · Allen Institute for AI

Abstract

Given a large corpus, the questions one might ask can vary -- from "When was the first human heart transplant?" to "What are all the contradictory claims in this literature?" -- but what makes some questions more challenging than others? In this work, we define a notion of Corpus Task Complexity (CTC) that characterizes tasks by how their difficulty grows with corpus size; for instance, a retrieval query only requires a single linear pass over a corpus, while finding contradictions requires checking a quadratically growing set of claim pairs. Observing that prior work has largely only studied tasks whose difficulty grows linearly with corpus size, which we call low CTC tasks, we introduce 10 new tasks belonging to a class of high CTC whose difficulty grows quadratically or more in corpus size. We find that high-CTC tasks not only grow much more challenging on average at longer contexts for LCLMs, they reverse many modeling conclusions drawn solely from low-CTC evaluations. For instance, efficient block-sparse and hybrid attention approaches consistently match full attention performance on low-CTC tasks, but degrade much more on high-CTC tasks. Large-corpus high-CTC reasoning thus remains an open challenge as full attention is too costly to scale, motivating future research on these tasks. We release our code, data, and 22-task suite (CTC-Bench), to facilitate future research in this area.

Figures & tables

Appendix figures & tables4 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. MixRea: Benchmarking Explicit-Implicit Reasoning in Large Language Models

    May 19, 2026Yuanqing Cai, Ziyi Huang, Minhao Liu +3LLM Reasoning StrategiesProgressive Reasoning

  2. Oolong: Evaluating Long Context Reasoning and Aggregation Capabilities

    Nov 4, 2025Amanda Bertsch, Adithya Pratapa, Teruko Mitamura +2Question-Answering Benchmarks

  3. Scaling Multi-Hop Training Data via Graph-Constrained Path Selection

    May 29, 2026Pengyu Chen, Yonggang Zhang, Mingming Chen +3Multi-Hop ReasoningReasoning Paths