Human Preference Evaluation

Latest papers 41

All topics
CardsList
  1. Why Expert Alignment Is Hard: Evidence from Subjective Evaluation

    May 6, 2026Tzu-Mi Lin, Wataru Hirota, Tatsuya Ishigaki +2Human Preference EvaluationLLM Alignment

  2. Aggregate vs. Personalized Judges in Business Idea Evaluation: Evidence from Expert Disagreement

    Apr 24, 2026Wataru Hirota, Tomoki Taniguchi, Tomoko Ohkuma +6Human Preference EvaluationAnnotator Disagreement

  3. Preferences of a Voice-First Nation: Large-Scale Pairwise Evaluation and Preference Analysis for TTS in Indian Languages

    Apr 23, 2026Srija Anand, Ashwin Sankar, Ishvinder Sethi +10Pairwise Preference LearningHuman Preference Evaluation

  4. Rater State Bias in RLHF Preference Data: An Audit Framework

    Apr 14, 2026Elena Kopteva, Vitaliy Hlynianyi-ZhukHuman Preference EvaluationHuman-in-the-Loop Annotation

  5. Why That Robot? A Qualitative Analysis of Justification Strategies for Robot Color Selection Across Occupational Contexts

    Mar 30, 2026Jiangen He, Wanqi Zhang, Jessica K. BarfieldHuman Preference EvaluationHuman-Robot Interaction

  6. Mediocrity is the key for LLM as a Judge Anchor Selection

    Mar 17, 2026Shachar Don-Yehiya, Asaf Yehudai, Leshem Choshen +1LLM-as-a-JudgePairwise Preference Evaluation

  7. Evaluating Alignment of Behavioral Dispositions in LLMs

    Feb 11, 2026Amir Taubenfeld, Zorik Gekhman, Lior Nezry +8Human Preference EvaluationLLM Alignment

  8. Expert Preference-based Evaluation of Automated Related Work Generation

    Aug 11, 2025Furkan Şahinuç, Subhabrata Dutta, Iryna GurevychHuman Preference EvaluationLLM-as-a-Judge

  9. Harmonious Color Pairings: Insights from Human Preference and Natural Hue Statistics

    Aug 3, 2025Ortensia Forni, Alexandre Darmon, Michael BenzaquenHuman Preference EvaluationHuman Preference Modeling