cs.LGOct 14, 2025

Demystifying Hybrid Thinking: Can LLMs Truly Switch Between Think and No-Think?

Authors: Shouren Wang, Wang Yang, Xianxuan Long, Debargha Ganguly, Qifan Wang, Vipin Chaudhary, Xiaotian Han

Organizations: Case Western Reserve University · Meta AI

Abstract

Hybrid thinking enables LLMs to switch between reasoning and direct answering, offering a balance between efficiency and reasoning capability. Yet our experiments reveal that current hybrid thinking LLMs only achieve partial mode separation: reasoning behaviors often leak into the no-think mode. To understand and mitigate this, we analyze the factors influencing controllability and identify four that matter most: (1) larger data scale, (2) using think and no-think answers from different questions rather than the same question, (3) a moderate increase in no-think data number, and (4) a two-phase strategy that first trains reasoning ability and then applies hybrid think training. Building on these findings, we propose a practical recipe that, compared to standard training, can maintain accuracy in both modes while significantly reducing no-think output length (from 1085 to 585 on MATH500) and occurrences of reasoning-supportive tokens such as "wait" (from 5917 to 522 on MATH500). Our findings highlight the limitations of current hybrid thinking and offer directions for strengthening its controllability. The code is available at: https://github.com/SR-A-W/demystifying-hybrid-thinking

Figures & tables

Appendix figures & tables41 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. HRBench: Benchmarking and Understanding Thinking-Mode Switch Strategies in Hybrid-Reasoning LLMs

    May 27, 2026Yansong Ning, Mianpeng Liu, Jingwen Ye +2Reasoning BenchmarkLarge Language Model Benchmarks

  2. Path-Lock Expert: Separating Reasoning Mode in Hybrid Thinking via Architecture-Level Separation

    Apr 29, 2026Shouren Wang, Wang Yang, Chuang Ma +7ThoughtsLock-In

  3. Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders

    Aug 8, 2026Bo Cheng, Qiaolin Lu, Yi Chang +1Metacognition