Annotator Disagreement

Latest papers 56

All topics
CardsList
  1. Calibrating Probabilistic Object Detectors with Annotator Disagreement

    May 23, 2026Zhi Qin Tan, Owen Addison, Yunpeng LiAnnotator DisagreementObject Detection

  2. Aligning LLM Uncertainty with Human Disagreement in Subjectivity Analysis

    May 11, 2026Junyu Lu, Deyi Ji, Xuanyi Liu +5Annotator DisagreementLLM Alignment

  3. Beyond Majority Voting: Agreement-Based Clustering to Model Annotator Perspectives in Subjective NLP Tasks

    May 11, 2026Tadesse Destaw Belay, Ibrahim Said Ahmad, Idris Abdulmumin +6Annotator DisagreementClustering

  4. Parser agreement and disagreement in L2 Korean UD: Implications for human-in-the-loop annotation

    May 7, 2026Hakyung Sung, Gyu-Ho ShinAnnotator DisagreementMorphosyntactic Tagging

  5. Who and What? Using Linguistic Features and Annotator Characteristics to Analyze Annotation Variation

    May 7, 2026Maximilian Maurer, Maximilian Linde, Gabriella LapesaAnnotator DisagreementHuman-in-the-Loop Annotation

  6. Understanding Annotator Safety Policy with Interpretability

    May 6, 2026Alex Oesterling, Donghao Ren, Yannick Assogba +4Annotator DisagreementHuman-in-the-Loop Annotation

  7. Annotation Quality in Aspect-Based Sentiment Analysis: A Case Study Comparing Experts, Students, Crowdworkers, and Large Language Model

    May 5, 2026Niklas Donhauser, Jakob Fehle, Nils Constantin Hellwig +3Annotator DisagreementLLM-Assisted Annotation

  8. STABLEVAL: Disagreement-Aware and Stable Evaluation of AI Systems

    May 4, 2026Akash Bonagiri, Gerard Janno Anderias, Saee Patil +6Annotator DisagreementAI Agent Evaluation

  9. Quantifying and Predicting Disagreement in Graded Human Ratings

    May 1, 2026Leixin Zhang, Çağrı ÇöltekinAnnotator Disagreement

  10. Are You the A-hole? A Fair, Multi-Perspective Ethical Reasoning Framework

    Apr 30, 2026Sheza Munir, Ahanaf Rodoshi, Sumin Lee +3Annotator DisagreementNeuro-Symbolic Reasoning

  11. LLMs Capture Emotion Labels, Not Emotion Uncertainty: Distributional Analysis and Calibration of Human-LLM Judgment Gaps

    Apr 30, 2026Keito Inoshita, Xiaokang Zhou, Akira Kawai +1LLM EvaluationAnnotator Disagreement

  12. Aggregate vs. Personalized Judges in Business Idea Evaluation: Evidence from Expert Disagreement

    Apr 24, 2026Wataru Hirota, Tomoki Taniguchi, Tomoko Ohkuma +6Human Preference EvaluationAnnotator Disagreement

  13. Fine-Grained Perspectives: Modeling Explanations with Annotator-Specific Rationales

    Apr 23, 2026Olufunke O. Sarumi, Charles Welch, Daniel BraunAnnotator DisagreementNatural Language Inference

  14. Who Watches the Watchmen? Humans Disagree With Translation Metrics on Unseen Domains

    Apr 19, 2026Finn Schmidt, Jan Philip Wahle, Terry Ruas +1Annotator DisagreementAutomated Evaluation

  15. Beyond Black-Box Labels: Interpretable Criteria for Diagnosing Subjective NLP Tasks

    Apr 18, 2026Nisrine Rair, Alban Goupil, Valeriu Vrabie +1Annotator Disagreement

  16. IYKYK (But AI Doesn't): Automated Content Moderation Does Not Capture Communities' Heterogeneous Attitudes Towards Reclaimed Language

    Apr 17, 2026Christina Chance, Rebecca Pattichis, Arjun Subramonian +4Social Media AnalysisAnnotator Disagreement

  17. Where Experts Disagree, Models Fail: Detecting Implicit Legal Citations in French Court Decisions

    Mar 24, 2026Avrile Floro, Tamara Dhorasoo, Soline Pellez +1Annotator DisagreementLegal Citation Verification

  18. Multi-Perspective LLM Annotations for Valid Analyses in Subjective Tasks

    Mar 22, 2026Navya Mehrotra, Adam Visokay, Kristina GligorićAnnotator DisagreementLLM-Assisted Annotation

  19. Labels have Human Values: Value Calibration of Subjective Tasks

    Jan 10, 2026Mohammed Fayiz Parappan, Ricardo HenaoAnnotator DisagreementAI Alignment