Human Preference Modeling

Latest papers 44

All topics
CardsList
  1. In-Context Reward Adaptation for Robust Preference Modeling

    May 28, 2026Zhenyu Sun, Zheng Xu, Ermin WeiHuman Preference ModelingIn-Context Learning

  2. PRISM-X: Experiments on Personalised Fine-Tuning with Human and Simulated Users

    May 13, 2026Hannah Rose Kirk, Liu Leqi, Fanzhi Zeng +4LLM EvaluationLLM Sycophancy

  3. Response Time Enhances Alignment with Heterogeneous Preferences

    May 7, 2026Federico Echenique, Alireza Fallah, Baihe Huang +1Pairwise Preference LearningLLM Alignment

  4. Mitigating Cognitive Bias in RLHF by Altering Rationality

    May 7, 2026Tiffany Horter, Andrew Markham, Niki Trigoni +1Human Preference ModelingRL from Human Feedback

  5. StoryAlign: Evaluating and Training Reward Models for Story Generation

    May 6, 2026Haotian Xia, Hao Peng, Yunjia Qi +4Reward ModelingLanguage Model Generation Evaluation

  6. Putting HUMANS first: Efficient LAM Evaluation with Human Preference Alignment

    Apr 20, 2026Woody Haosheng Gan, William Held, Diyi YangAudio-Language Model EvaluationHuman Preference Modeling

  7. Efficient Personalization of Generative User Interfaces

    Apr 10, 2026Yi-Hao Peng, Jeffrey P. Bigham, Jason WuPairwise Preference LearningHuman Preference Modeling

  8. Who Laughs with Whom? Disentangling Influential Factors in Humor Preferences across User Clusters and LLMs

    Jan 6, 2026Soichiro Murakami, Hidetaka Kamigaito, Hiroya Takamura +1LLM EvaluationHuman Preference Modeling

  9. HAL: Inducing Human-likeness in LLMs with Alignment

    Jan 6, 2026Masum Hasan, Junjie Zhao, Ehsan HoqueReward ModelingLLM Alignment

  10. Harmonious Color Pairings: Insights from Human Preference and Natural Hue Statistics

    Aug 3, 2025Ortensia Forni, Alexandre Darmon, Michael BenzaquenHuman Preference EvaluationHuman Preference Modeling

  11. DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable)

    Jul 10, 2025Wenxuan Zhou, Shujian Zhang, Brice Magdalou +4Direct Preference OptimizationHuman Preference Modeling

  12. A Probabilistic Approach for Model Alignment with Human Comparisons

    Mar 16, 2024Junyu Cao, Mohsen BayatiHuman Preference ModelingStatistical Learning Theory

  13. KTO: Model Alignment as Prospect Theoretic Optimization

    Feb 2, 2024Kawin Ethayarajh, Winnie Xu, Niklas Muennighoff +2LLM AlignmentHuman Preference Modeling