cs.CLOct 5, 2026

CLARA: Can AI Assess Developmental Appropriateness in Children's Stories?

Authors: Sijing Yin, Zirui Wang, Qian Liu, Jiamou Liu

Organizations: University of Auckland

Abstract

Assessing the developmental suitability of children's narratives is important for educational recommendation and developmental literacy research, yet such assessment typically relies on subjective and difficult-to-scale human judgment. This raises an important question: Can AI systems approximate human developmental judgments of children's stories? To study this problem, we introduce CLARA, a cognitively grounded framework for developmental narrative understanding through structured annotation across cognitive (COG), language (LAN), and social-emotional (SEL) dimensions, together with a bilingual benchmark resource containing 1107 Chinese--English children's stories with normalized silver developmental references and structured developmental annotations. We evaluate CLARA through benchmark comparison, component analysis, translated bilingual consistency analysis, and blinded human evaluation with educators. Experimental results show that structured developmental annotation achieves substantially stronger alignment with developmental references and human judgments than readability-based methods and direct prompting baselines. Overall, our findings suggest that AI systems can approximate certain aspects of human developmental judgment when guided by structured developmental annotation, while also highlighting the importance of interpretability and human oversight in educational NLP.

Figures & tables

Appendix figures & tables6 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Evaluating Developmental Cognition Capabilities of LLMs

    May 8, 2026Xiao Xiao, Hayoun Noh, Mar Gonzalez-FrancoConversational Artificial IntelligenceDevelopmental Trajectories

  2. BIASEDTALES-ML: A Multilingual Dataset for Analyzing Narrative Attribute Distributions in LLM-Generated Stories

    Apr 18, 2026Yuxuan Ouyang, yingfeng luo, JingBo Zhu +1Narrative GenerationNarratives

  3. CraftAlign: Feature-Grounded Evaluation and Revision Guidance for AI Stories

    Aug 2, 2026Yang Yang, Boyun Xu, Shaofeng Liang +5Value AlignmentMulti-Dimensional Evaluation