LLM Mathematical Reasoning

LLM: Large Language Model

Momentum

19 papers in the last four weeks, up 375% on the four weeks before. 0.2% of all new papers.

Jul 13Week of Sep 28

Latest papers 194

All topics
CardsList
  1. TUDUM: A Turkish-Thinking Reasoning Pipeline for Qwen3.5-27B

    Jul 2, 2026Baran Bingol, Bahaeddin TurkogluLLM Fine-TuningLLM Mathematical Reasoning

  2. Geometric Signatures of Reasoning: A Spectral Perspective on Task Hardness

    Jul 2, 2026Aria Masoomi, Mahsa Bazzaz, Adel Javanmard +1LLM EvaluationRepresentation Geometry in Language Models

  3. ISM:Self-Improving Strategy Memory for Continual Mathematical Reasoning

    Jun 30, 2026Prakhar Dixit, Tim OatesContinual Learning for LLMsMemory-Augmented Language Models

  4. The Ramanujan Challenge For AI

    Jun 27, 2026Michael Shalyt, Rotem Kalisch, Carsten Schneider +7Mathematical Reasoning BenchmarksLLM Mathematical Reasoning

  5. Self-Supervised Theorem Discovery in a Formal Axiomatic System

    Jun 27, 2026Kazuki Ota, Takayuki Osa, Tatsuya HaradaAutomated Theorem ProvingKnowledge Augmentation for Language Models

  6. Riazi-8B: An Urdu Large Language Model for Mathematical Reasoning

    Jun 24, 2026Azher Ali, Ibtsam Haider, Raja Khurram Shahzad +2Numerical Reasoning in Language ModelsLLM Fine-Tuning

  7. Cliff Tokens: Analyzing Failure Trigger Tokens in LLM Mathematical Reasoning

    Jun 24, 2026Jaeyong Ko, Jinu Lee, Pilsung Kang +1Numerical Reasoning in Language ModelsLLM Reliability

  8. Hard or Just Unreached? Diagnosing the Sampling Blind Spot in Math-Reasoning Difficulty Estimation

    Jun 17, 2026Luca Zhou, Sajel Shah, Emanuele Rodolà +1Mathematical Reasoning BenchmarksLLM Evaluation

  9. First Proof Second Batch

    Jun 16, 2026Mohammed Abouzaid, Nikhil Srivastava, Rachel Ward +1Mathematical Reasoning BenchmarksLLM Mathematical Reasoning

  10. MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning

    Jun 16, 2026Wanshi Xu, Haokun Zhao, Haidong Yuan +2Multimodal ReasoningMultimodal Reward Modeling

  11. Visored: A Controlled-Natural-Language Prover for LLM-Generated Mathematics

    Jun 16, 2026Xiyu Zhai, Xinyi Chen, Yiping Wang +3Automated Theorem ProvingLean Theorem Proving

  12. Quantifying Consistency in LLM Logical Reasoning via Structural Uncertainty

    Jun 15, 2026Baishali Chaudhury, Mengdie Flora Wang, Hyunji Hayley Park +3Reasoning Consistency in Language ModelsLLM Reliability

  13. The Quality-Utility Paradox: Why High-Reward Data Impairs Small Model Mathematical Reasoning

    Jun 15, 2026Haolong Qian, Xianliang Yang, Yinuo ma +6CoT DistillationLanguage Model Distillation

  14. Theorem-Grounded Execution Ontologies for Interpretable Machine Reasoning

    Jun 14, 2026Raghu AnantharangacharInterpretable MLFormal Verification

  15. Mask-Proof: An LLM-based Automated Data Curation Pipeline on Mathematical Proofs

    Jun 13, 2026Jierui Zhang, Siyuan Tan, Xinhang Li +8Mathematical Reasoning BenchmarksLLM Evaluation

  16. Failure Modes of Large Language Models on Research-Level Mathematics: A Taxonomy and an Empirical Characterisation

    Jun 12, 2026Arnesh Banerjee, Ayushi BhattacharjeeMathematical Reasoning BenchmarksLLM Auditing

  17. MaxProof: Scaling Mathematical Proof with Generative-Verifier RL and Population-Level Test-Time Scaling

    Jun 11, 2026Jiacheng Chen, Xinyu Zhang, Shunkai Zhang +20Automated Theorem ProvingInference-Time Scaling

  18. Architecture-Aware Reinforcement Learning Makes Sliding-Window Attention Competitive in Math Reasoning

    Jun 10, 2026Kai Liu, Peijie Dong, Xinchen Xie +5RL for Language Model ReasoningLLM Inference Acceleration

  19. ComBench: A Benchmark for Rigorous Proof Reasoning and Constructive Realization in Olympiad-Level Combinatorics

    Jun 9, 2026Shunkai Zhang, Haoran Zhang, Yun Luo +15Mathematical Reasoning BenchmarksBenchmark Design

  20. Diverse Thinking Schemata Elicit Better Reasoning in Large Language Models

    Jun 8, 2026Xinyue Liang, Yizhe Yang, Yu Bai +3Reasoning DiversityRL for Language Model Reasoning

  21. Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery

    Jun 7, 2026Syed Rifat Raiyan, Mohsinul Kabir, Hasan Mahmud +2Mathematical Reasoning BenchmarksAutomated Theorem Proving

  22. A Comprehensive Anatomy of Human and DeepSeek-R1 LLM Mathematical Reasoning

    Jun 5, 2026Yuxiang Chen, Jun WangCoT ReasoningStructured Reasoning

  23. From Correctness to Utility: Gain-Based Prefix Evaluation for LLM Reasoning

    Jun 5, 2026Yuhang Zhou, Yixin Cao, Guangnan YeProcess Reward ModelsRL for Language Model Reasoning