cs.AIOct 7, 2026

From Expert-Guided Proof Search to Automated Open-Problem Solving

Authors: Adrián Zámečník, Matěj Kripner, Martin Koutecký, Martin Balko, Jan Grebík, Pavel Hubáček, Robert Šámal, Václav Rozhoň

Organizations: Computer Science Institute, Charles University · Institute of Formal and Applied Linguistics, Charles University · Department of Applied Mathematics, Charles University · Institute of Mathematics, Czech Academy of Sciences

Abstract

Large language models are increasingly contributing to mathematical research, where progress often depends on efficient proof search, incremental improvements and careful verification. We describe Bolzano, a multi-agent open-source system that uses parallel prover agents with a verifier agent and maintains a human-readable research state. Initial manual use on expert-selected problems yielded 8 results whose proofs were checked by domain experts. Motivated by these case studies, we ran Bolzano without problem-specific human guidance on about 3,800 open problems extracted from four sets of papers, solving about 200 open problems. One experiment used papers accepted to STOC 2026, a top conference in theoretical computer science. There, we answered four questions raised in the papers, as confirmed by their authors.

Figures & tables

Explore similar work

CardsList
  1. Bolzano: Case Studies in LLM-Assisted Mathematical Research

    Apr 18, 2026Martin Balko, Jan Grebík, Pavel Hubáček +5Research-Level MathematicsTheorem Proving

  2. Advancing Mathematics Research with AI-Driven Formal Proof Search

    May 21, 2026George Tsoukalas, Anton Kovsharov, Sergey Shirobokov +18Theorem ProvingOpen Problems

  3. ProofCouncil: An LLM Agent for Solving Open Mathematical Problems

    Jul 10, 2026Johannes Schmitt, Tim Gehrunger, Jasper Dekoninck +4Open ProblemsLarge Language Model Agents