cs.CLSep 28, 2026

How Well Can LLMs Simulate Real Learner Evaluations of Educational Feedback?

Authors: Momoka Furuhashi, Kouta Nakayama, Takashi Kodama, Saku Sugawara, Kyosuke Takami

Organizations: Tohoku University · Research and Development Center for Large Language Models, National Institute of Informatics · National Institute of Informatics · University of Tokyo · Osaka Kyoiku University

Abstract

While recent studies have explored human behavior and preference simulation using large language models (LLMs), it remains unclear how well LLMs can simulate subjective evaluations from real learners in educational settings. We investigate this question using real learner evaluation data on feedback for high-school biology questions at both the group and individual levels. We compare performance with and without learner-specific information, such as personality traits and evaluation examples, across six models. Our results show that LLMs still have a limited ability to simulate learner evaluations. Providing learner profiles and examples improves score calibration and individual-level simulation, but more often fails to improve group-level consistency. These findings highlight the need to investigate which learner information and adaptation strategies are effective for learner preference simulation.

Figures & tables

Appendix figures & tables11 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Investigating Learner-Aware Design of LLM-Generated Educational Feedback

    Feb 12, 2026Momoka Furuhashi, Kouta Nakayama, Noboru Kawai +3FeedbackMultiple-Choice Questions

  2. Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators

    May 12, 2026Heejin Do, Shashank Sonkar, Mrinmaya SachanUser SimulationStudent Misconceptions