cs.CLMay 26, 2026

LLMs Are Already Good Tutors: Training-Free Prompt Optimization for Pedagogical Math Tutoring

Authors: Unggi LeeMinchul ShinYeil JeongSookbun LeeJeongsu MoonKyungtae JooEunjoo LeeHoilym Kwon

Organizations: Korea University Sejong Campus · 1Korea University Sejong Campus · Gyeonggi Institute of Education · 2Gyeonggi Institute of Education · Indiana University Bloomington · 3Indiana University Bloomington · Opentutorials · 4Opentutorials · Chosun University · 5Chosun University · Korea University Korean Studies Center · 6Korea University Korean Studies Center

Abstract

Aligning LLMs for math tutoring typically requires RL-based training with multi-GPU infrastructure. We investigate whether training-free prompt optimization-evolving only the system prompt via API calls-can serve as a practical alternative. We adapt 7 published methods and propose 5 education-specialized methods, evaluating these 12 methods under 5 conditions on 2 OOD benchmark suites. All 12 best-per-method configurations surpass the strongest RL-trained baseline (R_total = 0.633), and our ParetoGrad achieves the best Pareto balance across post-test solve rate, leak control, and helpfulness, rather than dominating any single component. Behavioral analysis with an 82-code educational codebook reveals that training-free methods rely on teaching-knowledge patterns at 2-3x the rate of RL-trained models, with a compensating ~10 percentage-point reduction in intent-level scaffolding. We also find a task-dependent reasoning mode effect consistent across training-free and RL-based paradigms. Our approach enables efficient development of pedagogically aligned LLM tutors with prompts alone and minimal compute.

Explore similar work

CardsList
  1. Towards Pedagogically Aligned LLM Tutors for Math Mistake Remediation

    Jun 19, 2026Kseniia Petukhova, Tien Dat Nguyen, Ekaterina KochmarTutorsText Corpora