Human Preference Modeling

Latest papers 43

All topics
CardsList
  1. Perceptually Aligned Evaluation of Style Transfer

    Oct 7, 2026Yang Deng, Eleftherios Ioannou, David Mould +3Image Style TransferAutomated Evaluation

  2. Verifiable, Articulable, and Tacit Components of Preference

    Oct 2, 2026Alexander Spangher, Sheldon S. Huang, Andreas Haupt +4Human Preference ModelingPreference Learning

  3. SpeechCritic: Learning a Diagnostic Speech Judge from Limited Human Preferences

    Sep 28, 2026Mingyue Huo, Shivam Mehta, Bhavin Jawade +2Audio-Language Model EvaluationHuman Preference Modeling

  4. Learning Heterogeneous Preferences

    Sep 15, 2026Shiwali Mohan, Matt Hong, Dule Shu +3Pairwise Preference LearningHuman Preference Modeling

  5. SVG-Score: Human-Aligned Evaluation of Text-to-SVG Generation

    Sep 3, 2026Marco Cipriano, Leonardo Zini, Alexandra Schild +5VLM EvaluationHuman Preference Modeling

  6. VA-Judger: Reward Modeling from Human Preference Feedback for Joint Video-Audio Generation

    Aug 19, 2026Yinming Huang, Shuyuan Tu, Xi Yan +6Reward ModelingAudio-Video Generation

  7. PALMs: Using Multi Construct-Grounded Rationales for Modeling Population Preferences in LLMs

    Aug 2, 2026Priyanka Dey, Brihi Joshi, Preyashi Poddar +2LLM AlignmentHuman Preference Modeling

  8. What do Reward Models Memorize?

    Jul 27, 2026Ivo Verhoeven, Pushkar Mishra, Ekaterina ShutovaReward ModelingPairwise Preference Evaluation

  9. StARS: Socially Appropriate Robot Actions via a Recommender System-Driven Approach

    Jul 23, 2026Erencem Ozbey, Fethiye Irmak Dogan, Jin Huang +1Human Preference ModelingHuman-Robot Interaction

  10. Step-Level Preference Learning for Generative Agents in Social Simulations

    Jul 16, 2026Wenchang Gao, Pingyue Sheng, Lanlan Qiu +7Human Preference ModelingHuman Behavior Simulation

  11. Consensus vs. Dissent: Dynamic LLM Modeling of Subjective Preferences in Group Recommenders

    Jul 11, 2026Cedric Waterschoot, Nava Tintarev, Francesco BarileHuman Preference ModelingRecommender Systems

  12. A Bayesian framework for the uncanny valley in humanoid robot design

    Jul 7, 2026Shimon Honda, Rin Shibano, Hideyoshi YanagisawaHuman Preference ModelingHierarchical Bayesian Modeling

  13. Human-Centric Reflective Architecture for Human-AI Collaborative Decision-Making

    Jul 3, 2026Andreas Kouridakis, Dimitrios Patiniotis Spyropoulos, George VourosHuman Preference ModelingHuman-in-the-Loop AI

  14. LLM Consumer Behavior Theory: Foundations of a Novel Research Field

    Jun 16, 2026Manon Reusens, Sofie Goethals, David MartensHuman Preference ModelingLLM Decision-Making

  15. City landscape in sight: A crowdsourced framework for unlocking urban-scale window view perceptions from real estate imagery

    Jun 13, 2026Chucai Peng, Sijie Yang, Ang Liu +3Human Preference ModelingUrban Planning

  16. LLMs Can Better Capture Human Judgments--With the Right Prompts

    Jun 10, 2026Danica Dillion, Chen Cecilia Liu, Baihui Wang +5Human Preference EvaluationLLM Alignment

  17. Nonslop: A Gamified Experiment in Human-AI Collaborative Writing

    Jun 10, 2026Maria Edwards, Julian TogeliusHuman Preference ModelingHuman-AI Interaction

  18. Hidden Consensus:Preference-Validity Compression in Human Feedback

    Jun 9, 2026Dorcas Chia Ern Chua, Karen Myn Hui Lee, Jia Yue Tan +9Human Preference ModelingPreference Alignment

  19. Whose Norms? Disentangling Cultural and Personal Alignment in Large Language Models

    Jun 5, 2026Angana Borah, Isabelle Augenstein, Rada MihalceaLLM AlignmentNormative Conformity in Language Models

  20. What Do People Actually Want From AI? Mapping Preference Plurality

    Jun 4, 2026Julia Sepúlveda Coelho, Scott A. HaleLLM AlignmentHuman Preference Modeling

  21. Coherence Maximization Improves Pluralistic Alignment

    Jun 2, 2026Taslim Mahbub, Yiding Pei, Shi FengLLM AlignmentHuman Preference Modeling

  22. Large Language Models Should Learn Personalized Rather Than Aggregated Human Preferences

    May 30, 2026Cristina GarbaceaLLM AlignmentSocial Choice Theory

  23. In-Context Reward Adaptation for Robust Preference Modeling

    May 28, 2026Zhenyu Sun, Zheng Xu, Ermin WeiHuman Preference ModelingIn-Context Learning

  24. PRISM-X: Experiments on Personalised Fine-Tuning with Human and Simulated Users

    May 13, 2026Hannah Rose Kirk, Liu Leqi, Fanzhi Zeng +4LLM EvaluationLLM Sycophancy

  25. Response Time Enhances Alignment with Heterogeneous Preferences

    May 7, 2026Federico Echenique, Alireza Fallah, Baihe Huang +1Pairwise Preference LearningLLM Alignment

  26. Mitigating Cognitive Bias in RLHF by Altering Rationality

    May 7, 2026Tiffany Horter, Andrew Markham, Niki Trigoni +1Human Preference ModelingRL from Human Feedback

  27. StoryAlign: Evaluating and Training Reward Models for Story Generation

    May 6, 2026Haotian Xia, Hao Peng, Yunjia Qi +4Reward ModelingLanguage Model Generation Evaluation

  28. Putting HUMANS first: Efficient LAM Evaluation with Human Preference Alignment

    Apr 20, 2026Woody Haosheng Gan, William Held, Diyi YangAudio-Language Model EvaluationHuman Preference Modeling

  29. Efficient Personalization of Generative User Interfaces

    Apr 10, 2026Yi-Hao Peng, Jeffrey P. Bigham, Jason WuPairwise Preference LearningHuman Preference Modeling

  30. Who Laughs with Whom? Disentangling Influential Factors in Humor Preferences across User Clusters and LLMs

    Jan 6, 2026Soichiro Murakami, Hidetaka Kamigaito, Hiroya Takamura +1LLM EvaluationHuman Preference Modeling

  31. HAL: Inducing Human-likeness in LLMs with Alignment

    Jan 6, 2026Masum Hasan, Junjie Zhao, Ehsan HoqueReward ModelingLLM Alignment

  32. Harmonious Color Pairings: Insights from Human Preference and Natural Hue Statistics

    Aug 3, 2025Ortensia Forni, Alexandre Darmon, Michael BenzaquenHuman Preference EvaluationHuman Preference Modeling

  33. DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable)

    Jul 10, 2025Wenxuan Zhou, Shujian Zhang, Brice Magdalou +4Direct Preference OptimizationHuman Preference Modeling

  34. A Probabilistic Approach for Model Alignment with Human Comparisons

    Mar 16, 2024Junyu Cao, Mohsen BayatiHuman Preference ModelingStatistical Learning Theory

  35. KTO: Model Alignment as Prospect Theoretic Optimization

    Feb 2, 2024Kawin Ethayarajh, Winnie Xu, Niklas Muennighoff +2LLM AlignmentHuman Preference Modeling