Annotator Disagreement

Latest papers 56

All topics
CardsList
  1. Language-model ratings of depression reflect the rater more than the patient

    Oct 6, 2026Baihan LinInter-Rater ReliabilityAnnotator Disagreement

  2. QuanReview: Offline, Auditable Reconciliation of Human and LLM Span Annotations

    Sep 28, 2026Matteo Musacchio, Juan Cruz Giner Pulero, Isabel Castañeda +4Annotator DisagreementLLM-Assisted Annotation

  3. How Many Humans Are 32 LLM Judges Worth?

    Sep 18, 2026Chao Li, Yingying Yu, Yunfeng LiAnnotator DisagreementLLM-as-a-Judge

  4. Predicting Human Disagreement for Calibrated Dynamic Facial Expression Recognition

    Sep 15, 2026Yiming Wang, Frederick W. B. Li, Jingyun WangAnnotator DisagreementUncertainty Quantification

  5. Uncertainty-Aware Sea-Ice Type Mapping with Multiple Ice Charts

    Sep 8, 2026Samira Alkaee Taleghan, Younghyun Koo, Andrew P. Barrett +1Remote Sensing Image UnderstandingAnnotator Disagreement

  6. Aggregate Disambiguation Systems

    Aug 31, 2026José María Lago, Albert Castellana, Edgars NemšeAnnotator DisagreementConfidence Region Estimation

  7. Definitional Sensitivity in Media Bias Detection: A Multi-Definition Dataset and Benchmark

    Aug 24, 2026Martin Wessel, Timo Spinde, Jürgen Pfeffer +1Annotator Disagreement

  8. Learning Sexism Detection Using Multi-Agent Perspectivist Preference Optimization

    Aug 4, 2026Hadi Mohammadi, Tina Shahedi, Robert A. Bagheri +2Annotator DisagreementPreference Optimization

  9. Ensemble Diversity Optimization for Subjective Supervision

    Jul 9, 2026Xia Cui, Ziyi Huang, N. R. AbeynayakeAnnotator DisagreementEnsemble Learning

  10. A Hybrid Framework for Song Lyric Annotation Based on Human-LLM Alignment

    Jun 28, 2026Rashini Liyanarachchi, Frank Tran, Md Mahmudul Hasan +2Annotator DisagreementLLM-Assisted Annotation

  11. Learning from Annotation Uncertainty: Entropy-Aware Curriculum for Speech Emotion Recognition

    Jun 25, 2026Zahra Omidi, John H. L. HansenAnnotator DisagreementLabel Distribution Learning

  12. Introducing corpora Hlava Cor and Hlava AD: Human Label Variation in Coreference and Discourse Relations

    Jun 24, 2026Anna Nedoluzhko, Šárka Zikánová, Jiří Mírovský +2Inter-Rater ReliabilityCoreference Resolution

  13. Learning Moral Diversity: Modelling Individual Perspectives in Moral Classification of Texts

    Jun 22, 2026Yi Ren, Lewis Mitchell, Matthew RoughanAnnotator DisagreementText Classification

  14. Quality and Agreement in Multilabel Emotion Annotation: A Case Study and Evaluation Framework

    Jun 19, 2026Emily Öhman, Anna KoufakouSoft-Label LearningAnnotator Disagreement

  15. Event-Aligned Analysis of Multi-Rater Pain Assessments Using Continuous Wearable Physiology

    Jun 11, 2026Saba A. Farahani, Elahe Khatibi, Thomas D. Hughes +3Inter-Rater ReliabilityAnnotator Disagreement

  16. A Resource for Enthymeme Detection in Controversial Political Discourse

    Jun 10, 2026Martial Pastor, Nelleke OostdijkAnnotator DisagreementHuman-in-the-Loop Annotation

  17. The Ghost Annotator: a Framework to Explore Human Label Variation in Content Moderation through Conformal Prediction

    Jun 1, 2026Mirko Lai, Alessandra Urbinati, Simona Frenda +2Annotator DisagreementConformal Prediction

  18. Bayesian Spectral Emotion Transition Discovery from Multi-Annotator Disagreement

    Jun 1, 2026Keito Inoshita, Takato UenoAnnotator DisagreementEmotion Recognition in Conversations

  19. Disagreeing Rationales: Rethinking Classification and Explainability Evaluation in Hate Speech Detection

    May 29, 2026Benedetta Muscato, Beiduo Chen, Gizem Gezici +2Annotator DisagreementHate Speech Detection

  20. When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis

    May 27, 2026Aisha Najera, Alvin Moon, Vedant Srinivasan +1LLM EvaluationAnnotator Disagreement

  21. Temporal Simultaneity Predicts Annotation Quality in Sentiment Corpora

    May 26, 2026Idris Abdulmumin, Mokgadi Penelope Matloga, Tadesse Destaw Belay +5Annotator DisagreementSentiment Analysis

  22. A Two-Phase Stability Study of LLM Judges and Bar Council Examiners on Thai Bar-Exam Free-Form Essays

    May 25, 2026Pawitsapak Akarajaradwong, Wuttikrai Lertprasertphakorn, Chompakorn Chaksangchaichot +1Inter-Rater ReliabilityAnnotator Disagreement

  23. Uncertainty Decomposition via Cyclical SG-MCMC and Soft-label Learning for Subjective NLP

    May 23, 2026Keito Inoshita, Takato UenoSoft-Label LearningBayesian Neural Networks