LLM Mathematical Reasoning

LLM: Large Language Model

Momentum

19 papers in the last four weeks, up 375% on the four weeks before. 0.2% of all new papers.

Jul 13Week of Sep 28

Latest papers 194

All topics
CardsList
  1. Pessimistic Verification for Open Ended Math Questions

    Nov 26, 2025Yanxing Huang, Zihan Tang, Zejin Lin +2LLM Answer VerificationLLM Mathematical Reasoning

  2. OpenSIR: Open-Ended Self-Improving Reasoner

    Nov 1, 2025Wai-Chung Kwan, Joshua Ong Jun Leang, Pavlos Vougiouklis +3Self-Play RLRL for Language Model Reasoning

  3. TopoAlign: A Framework for Aligning Code to Math via Topological Decomposition

    Oct 13, 2025Yupei Li, Philipp Borchert, Gerasimos LampourasAutomated Theorem ProvingAutoformalization

  4. Aria: An Agent For Retrieval and Iterative Auto-Formalization via Dependency Graph

    Oct 6, 2025Hanyu Wang, Ruohan Xie, Yutong Wang +3Lean Theorem ProvingAutoformalization

  5. Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning

    Sep 27, 2025Ningning Xu, Yuxuan Jiang, Shubhashis Roy Dipta +1LLM Tool UseAlgorithmic Reasoning

  6. Generative AI performance in core undergraduate mathematics: a curriculum-level case study

    Sep 15, 2025Benjamin J. Walker, Nikoleta Kalaydzhieva, Beatriz Navarro Lameda +1Educational AssessmentGenerative AI in Education

  7. Formally Solving Answer-Construction Problems in Lean

    May 24, 2025Jialiang Sun, Yuzhi Tang, Ao Li +2Lean Theorem ProvingLLM Mathematical Reasoning

  8. EquivPruner: Boosting Efficiency and Quality in LLM-Based Search via Action Pruning

    May 22, 2025Jiawei Liu, Qisi Chen, Jianshu Zhang +2LLM Inference EfficiencySemantic Textual Similarity

  9. The relationship between reasoning and performance in large language models--o3 (mini) thinks harder, not longer

    Feb 21, 2025Marthe Ballon, Andres Algaba, Vincent GinisCoT ReasoningEfficient Language Model Reasoning

  10. Benchmarking LLMs' Mathematical Reasoning with Unseen Random Variables Questions

    Jan 20, 2025Zijin Hong, Hao Wu, Su Dong +8Mathematical Reasoning BenchmarksSynthetic Benchmark Generation

  11. Measuring Progress in Reasoning Toward Mathematical Discovery with Automatic Verification

    Date pendingErik Y. Wang, Sumeet R. Motwani, James V. Roggeveen +9Mathematical Reasoning BenchmarksLLM Mathematical Reasoning

  12. TREAT: Evaluating Access to Formal Knowledge across Equivalent Mathematical Representations

    Date pendingFateme Mazdarani, Carlos ToxtliMathematical Reasoning BenchmarksLanguage Model Robustness