cs.AI · 2609.25852 Copy arXiv ID · Sep 22, 2026 Save Prediction Is Not Detection: Evaluating Pre-Recognition Claims in Longitudinal Clinical AI Authors: Jing Yang , Long R. Jiao , Xiujun Cai , Zongjiu Zhang
Abstract Clinically useful early detection requires validated pre-recognition lead time. Yet event-based evaluations of longitudinal clinical AI can treat recognition-mediated care-process signals as shortcuts and recognition-dependent endpoints as reference standards, inflating apparent performance and lead time while undermining cross-center transport. Such results may serve prognosis without establishing detection before recognition. We define an interval-censored pre-recognition transition, an independent as-of reference standard, and a prespecified recognition proxy to make the claim testable.
Explore similar work Apr 17, 2026 · Jianyou Wang, Youze Zheng, Longtian Bao +11 Clinical Prediction Outcome
May 12, 2026 · Benjamin Turtel, Paul Wilczewski, Kris Skotheim Clinical Prediction Longitudinal
May 16, 2026 · Pujun Feng, Xiaoyu Guo, Seyed Ehsan Saffari +10 Clinical Prediction Clinical Decision Support
Apr 17, 2026 · cs.AI J/K move · Enter open · S save
Jianyou Wang, Youze Zheng, Longtian Bao, Hanyuan Zhang +10
Laboratory for Emerging Intelligence, University of California, San Diego · Elsevier · Department of Dermatology, University of California, San Diego
Scientists have long sought to accurately predict outcomes of real-world events before they happen. Can AI systems do so more reliably? We study this question through clinical trial outcome prediction, a high-stakes open challenge even for domain experts. We introduce CT Open, an open-access, live platform that will run four challenge every year. Anyone can submit predictions for each challenge. CT Open evaluates those submissions on trials whose outcomes were not yet public at the time of submission but were made public afterwards. Determining if a trial's outcome is public on the internet before a certain date is surprisingly difficult. Outcomes posted on official registries may lag behind by years, while the first mention may appear in obscure articles. To address this, we propose a novel, fully automated decontamination pipeline that uses iterative LLM-powered web search to identify the earliest mention of trial outcomes. We validate the pipeline's quality and accuracy by human expert's annotations. Since CT Open's pipeline ensures that every evaluated trial had no publicly reported outcome when the prediction was made, it allows participants to use any methodology and any data source. In this paper, we release a training set and two time-stamped test benchmarks, Winter 2025 and Summer 2025. We believe CT Open can serve as a central hub for advancing AI research on forecasting real-world outcomes before they occur, while also informing biomedical research and improving clinical trial design. CT Open Platform is hosted at
\href \href{https://ct-open.net/}{https://ct-open.net/} \href