Aligning Clinical Needs and AI Capabilities: A Survey on LLMs for Medical Reasoning
Authors: Qi Peng, Jiatong Li, Sirui Huang, Yiyang Jiang, Kaisong Gong, Ronger Ding, Shijie Ye, Changmeng Zheng, +5 more
Organizations: 1The Hong Kong Polytechnic University, Hong Kong, China. · 2Hong Kong University, Hong Kong, China. · 3South China University of Technology, Guangzhou, China. · University of Toronto, Toronto, Canada. · 5Peking Union Medical College Hospital, Beijing, China. · 6West China Hospital, Sichuan University, Sichuan, China.
Large language models (LLMs) have emerged as important tools in healthcare, showing growing potential for clinical reasoning and patient care. This survey examines recent progress in medical LLMs, focusing on reasoning applications and requirements. We present a dual-view approach that connects clinical practice with computational methods. On the clinical side, we establish a five-level competency scheme following Miller's Pyramid, progressing from knowledge recall to dynamic case management. On the computational side, we link deductive, inductive, and abductive reasoning patterns to common medical goals and tasks. We also introduce a benchmark dataset spanning five levels of medical reasoning capability and report results on 18 state-of-the-art models, revealing that medical specialist models excel in diagnosis-centric tasks while general models lead in decision support and dialogue. We conclude by discussing current progress and open challenges, including data limitations, hallucination, and grounding issues, and outline directions toward safer, more reliable, and workflow-ready systems.