Research-Level Mathematics

Momentum

1 paper in the last four weeks, with none the four weeks before. 0.0% of all new papers.

Jul 6Week of Sep 21

Latest papers 26

All topics
CardsList
  1. AI and Human Approaches to Mathematical Problem Solving

    Sep 15, 2026Yang DingResearch-Level MathematicsArtificial Intelligence Research

  2. Albilich: Steerable Proof-State Orchestration for LLM-Based Mathematical Research with CAS Integration

    Jul 30, 2026Ting Gong, Michael Ruofan Zeng, Yong YangResearch-Level MathematicsTheorem Proving

  3. AdvancedMathBench: A Benchmark Suite for Advanced Mathematical Proof Generation and Verification

    Jul 13, 2026Lingkai Kong, Zijian Wu, Yuzhe Gu +11Research-Level MathematicsMathematical Reasoning Benchmarks

  4. From Solvers to Research: Large Language Model-Driven Formal Mathematics at the Research Frontier

    Jul 8, 2026Eric Jiang, Xiao Liang, Yikai Zhang +16Research-Level MathematicsMathematics

  5. First Proof Second Batch

    Jun 16, 2026Mohammed Abouzaid, Nikhil Srivastava, Rachel Ward +1Research-Level MathematicsProof

  6. Mask-Proof: An LLM-based Automated Data Curation Pipeline on Mathematical Proofs

    Jun 13, 2026Jierui Zhang, Siyuan Tan, Xinhang Li +8Research-Level MathematicsData-Curation

  7. Failure Modes of Large Language Models on Research-Level Mathematics: A Taxonomy and an Empirical Characterisation

    Jun 12, 2026Arnesh Banerjee, Ayushi BhattacharjeeResearch-Level MathematicsLLM Reasoning Strategies

  8. MA-ProofBench: A Two-Tiered Evaluation of LLMs for Theorem Proving in Mathematical Analysis

    Jun 11, 2026Lushi Pu, Weiming Zhang, Xinheng Xie +6Research-Level MathematicsTheorem Proving

  9. Benchmarks in Leipzig

    Jun 4, 2026Andrei Balakin, Miklós Bóna, Marie-Charlotte Brandenburg +45Research-Level Mathematics

  10. GTBench: A Curriculum-Grounded Benchmark for Evaluating LLMs as Mathematical Research Assistants in Graph Theory

    Jun 2, 2026Noujoud Nader, Ibrahem Aljabea, Patrick Diehl +1Research-Level MathematicsTheorem Proving

  11. ResearchMath-14K: Scaling Research-Level Mathematics via Agents

    May 27, 2026Guijin Son, Seungyeop Yi, Minju Gwak +3Research-Level MathematicsMathematics

  12. RMA: Context-Orchestrated Research Math Agents

    May 20, 2026Zelin Zhao, Bo Yuan, Yuchen Zhu +2Research-Level MathematicsTheorem Proving

  13. MathAtlas: A Benchmark for Autoformalization in the Wild

    May 13, 2026Nilay Patel, Noah Arias, Davit Babayan +7AutoformalizationResearch-Level Mathematics

  14. Formal Conjectures: An Open and Evolving Benchmark for Verified Discovery in Mathematics

    May 13, 2026Moritz Firsching, Paul Lezeau, Salvatore Mercuri +8Research-Level MathematicsOpen Problems

  15. Not All Proofs Are Equal: Evaluating LLM Proof Quality Beyond Correctness

    May 11, 2026Ivo Petrov, Jasper Dekoninck, Dimitar I. Dimitrov +1Research-Level MathematicsProof

  16. Re2^2Math: Benchmarking Theorem Retrieval in Research-Level Mathematics

    May 9, 2026Zicheng Lyu, Wenjie Yang, Shengzhong Zhang +1Research-Level MathematicsTheorem Proving

  17. Matlas: A Semantic Search Engine for Mathematics

    Apr 19, 2026Haocheng Ju, Leheng Chen, Peihao Wu +2Research-Level MathematicsScholar Information Database

  18. Bolzano: Case Studies in LLM-Assisted Mathematical Research

    Apr 18, 2026Martin Balko, Jan Grebík, Pavel Hubáček +5Research-Level MathematicsTheorem Proving

  19. LiveMathematicianBench: A Live Benchmark for Research-Level Mathematical Reasoning with Proof Sketches

    Apr 2, 2026Linyang He, Qiyao Yu, Hanze Dong +5Research-Level MathematicsLarge Language Model Benchmarks

  20. LemmaBench: A Live, Research-Level Benchmark to Evaluate LLM Capabilities in Mathematics

    Feb 27, 2026Antoine Peyronnet, Fabian Gloeckle, Amaury HayatResearch-Level MathematicsTheorem Proving