Automated Grading

Momentum

3 papers in the last four weeks, against 1 the four weeks before. 0.0% of all new papers.

Jul 13Week of Sep 28

Latest papers 57

All topics
CardsList
  1. Leveraging BART to Assess CS1 C++ Programming Assignments using Rubric-based Criteria

    Jun 2, 2026Kelsey Rainey, Jesse RobertsComputer Science EducationMulti-Task Learning

  2. LLMs as Teaching Assistants for Mathematics Exam Grading: Reliability, and Practical Usability

    Jun 1, 2026Aastha Sapkota, M. G. Sarwar MurshedLLM EvaluationLLM-as-a-Judge

  3. Cost-Effective Automated Judging of Natural-Language Mathematical Proofs

    May 29, 2026Benjamin GrayzelLLM-as-a-JudgeAutomated Grading

  4. From Kellgren-Lawrence to Calcium Pyrophosphate Crystal Deposition: A Soft-Labelling Framework for Knee Osteoarthritis Assessmen

    May 27, 2026Francisco Bérchez-Moreno, Riccardo Rosati, Maria Chiara Fiorentino +6Label Distribution LearningMedical Image Classification

  5. Machine learning applied to emerald gemstone grading: framework proposal and creation of a public dataset

    May 22, 2026FB Pena, D Crabi, Sandro C Izidoro +2Automated GradingImage Classification

  6. Exploring the Effectiveness of Using LLMs for Automated Assessment of Student Self Explanations in Programming Education

    May 20, 2026Arun-Balajiee Lekshmi-Narayanan, Mohammad Hassany, Peter BrusilovskyLLM-as-a-JudgeProgramming Education

  7. GradeLegal: Automated Grading for German Legal Cases

    May 20, 2026Abdullah Al Zubaer, Lorenz Wendlinger, Simon Alexander Nonn +2LLM EvaluationAutomated Grading

  8. Generative-Evaluative Agreement: A Necessary Validity Criterion for LLM-Enabled Adaptive Assessment

    May 19, 2026Grandee Lee, Yue Wang, Che Yee Lye +1LLM EvaluationEducational Assessment

  9. Automated Grading of Handwritten Mathematics Using Vision-Capable LLMs

    May 18, 2026Jacob Levine, Miguel Aenlle, Craig Zilles +2VLM EvaluationRubric-Based Evaluation

  10. RETUYT-INCO at BEA 2026 Shared Task 2: Meta-prompting in Rubric-based Scoring for German

    May 11, 2026Ignacio Sastre, Ignacio Remersaro, Facundo Díaz +4LLM PromptingRubric-Based Evaluation

  11. High Precision Hydraulic Excavator Control for Heavy-Duty Grading

    May 10, 2026Lennart Werner, Pol Eyschen, Sean Costello +2Robotic ControlRobotics

  12. Creating and Evaluating K-12 GenAI Assessment Graders Through Context Engineering

    May 8, 2026Zewei Tian, Alex Liu, Lief Esbenshade +6Educational AssessmentGenerative AI in Education

  13. Quality-Conditioned Agreement in Automated Short Answer Scoring: Mid-Range Degradation and the Impact of Task-Specific Adaptation

    May 8, 2026Abigail Victoria Gurin Schleifer, Moriah Ariely, Beata Beigman Klebanov +2LLM EvaluationAI in Education

  14. LaTA: A Drop-in, FERPA-Compliant Local-LLM Autograder for Upper-Division STEM Coursework

    May 6, 2026Jesse A. RodríguezLLM-as-a-JudgeGenerative AI in Education

  15. AISSA: Implementation and Deployment of an AI-based Student Slides Analysis tool for Academic Presentations

    May 6, 2026Alvaro Becerra, Diego Gomez, Ruth CobosGenerative AI in EducationAutomated Grading

  16. Estimating LLM Grading Ability and Response Difficulty in Automatic Short Answer Grading via Item Response Theory

    Apr 30, 2026Longwei Cong, Sonja Hahn, Sebastian Gombert +3LLM EvaluationAutomated Grading

  17. Confidence Estimation in Automatic Short Answer Grading with LLMs

    Apr 30, 2026Longwei Cong, Sonja Hahn, Sebastian Gombert +3Uncertainty QuantificationSelective Prediction

  18. Human-in-the-Loop Benchmarking of Heterogeneous LLMs for Automated Competency Assessment in Secondary Level Mathematics

    Apr 29, 2026Jatin Bhusal, Nancy Mahatha, Aayush Acharya +1LLM EvaluationEducational Assessment

  19. Knee-xRAI: An Explainable AI Framework for Automatic Kellgren-Lawrence Grading of Knee Osteoarthritis

    Apr 25, 2026Azmul A. Irfan, Nur Ahmad Khatim, Alfan Alfian Irfan +3Explainable Medical Image AnalysisAutomated Grading

  20. REC-CBM: Rubric-Aware Error-Correction Concept Bottleneck Models for Trustworthy Open-Ended Grading

    Apr 24, 2026Chengshuai Zhao, Fan Zhang, Kumar Satvik Chaudhary +4Concept Bottleneck ModelsAutomated Essay Scoring

  21. Do Agents Dream of Root Shells? Partial-Credit Evaluation of LLM Agents in Capture the Flag Challenges

    Apr 21, 2026Ali Al-Kaswan, Maksim Plotnikov, Maxim Hájek +3LLM Agent EvaluationAI Agent Security Benchmarks

  22. From Scoring to Explanations: Evaluating SHAP and LLM Rationales for Rubric-based Teaching Quality Assessment

    Apr 18, 2026Ivo Bueno, Babette Bühler, Philipp Stark +5Educational AssessmentShapley Value Attribution

  23. Persona Matters: Effects of Activation Steering on Short Answer Generation and Scoring

    Apr 8, 2026Yongchao Wu, Aron HenrikssonOpen-Ended GenerationLanguage Model Steering

  24. LLM-as-a-judge validity in physics assessment depends more on the task than the model

    Mar 16, 2026Will Yeadon, Tom Hardy, Paul Mackay +1LLM-as-a-JudgeEducational Assessment