LLM-Assisted Annotation

Latest papers 124

All topics
CardsList
  1. BEATS: Bootstrapping E-commerce Attribute Taxonomies for Search through Iterative Human-AI Collaboration

    Jun 3, 2026Yung-Yu Shih, Shang-Yu Su, Tzu-I Ho +2Semantic SearchLLM-Assisted Annotation

  2. Framing Migration News with LLMs: Structured CoT as a Support for Human Interpretation

    Jun 2, 2026David Alonso del Barrio, Jing Wen, Daniel Gatica-PerezLLM InterpretabilityCoT Reasoning

  3. TalkTag: Fine-Grained Morphosyntactic Error Annotation for Transcribed Speech

    Jun 1, 2026Shamira Venturini, Oliver Hennhöfer, Steffen Kinkel +1LLM Fine-TuningLLM-Assisted Annotation

  4. EvoPool: Evolutionary Programmatic Annotation for Label-Efficient Specialized Supervision

    Jun 1, 2026Tianyi Xu, Yaolun Zhang, Xuan Ouyang +1Multi-Agent LLM SystemsLLM-Assisted Annotation

  5. On the Limits of LLM Adaptability: Impact of Model-Internalized Priors on Annotation Task Performance

    May 30, 2026Etienne Casanova, Rafal Kocielnik, R. Michael AlvarezLLM EvaluationZero-Shot Text Classification

  6. Wind Turbine Maintenance Log Labelling Framework: LLM-Driven Data Correction and Enrichment via Semantic Extraction of Reliability Intelligence

    May 29, 2026Max Malyi, Jonathan Shek, Alasdair McDonald +1LLM-Assisted AnnotationDocument Information Extraction

  7. Industrializing Prediction-Powered Inference: The GLIDE Library for Reliable GenAI and Agentic Systems Evaluation

    May 29, 2026Grégoire Martinon, Ibrahim Merad, Mohammed RakiLLM-Assisted AnnotationLLM Agent Evaluation

  8. Traceable by Design: An LLM Pipeline and Dashboard for EU Regulatory Consultation Analysis

    May 29, 2026Thales Bertaglia, Haoyang Gui, Catalina Goanta +1LLM-Assisted AnnotationDocument Information Extraction

  9. Beyond Agreement: Scoring Panel-Surfaced Biomedical Entity Candidates for Curator Triage

    May 29, 2026Shuheng Cao, Ruiqi Chen, Renjie Cao +3Named Entity RecognitionLLM-Assisted Annotation

  10. When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis

    May 27, 2026Aisha Najera, Alvin Moon, Vedant Srinivasan +1LLM EvaluationAnnotator Disagreement

  11. Frontier LLM-based agents can overcome the ontology curation bottleneck for natural phenotypes

    May 27, 2026James P. Balhoff, Hilmar LappLLM-Assisted AnnotationLLM Agent Evaluation

  12. Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization

    May 27, 2026Beiduo Chen, Pingjun Hong, Ziyun Zhang +3LLM-Assisted AnnotationPreference Optimization

  13. From Learning Resources to Competencies: LLM-Based Tagging with Evidence and Graph Constraints

    May 27, 2026Ngoc Luyen Le, Marie-Hélène Abel, Bertrand LaforgeLLM-Assisted AnnotationLLM Information Extraction

  14. Where LLM Annotators Fail: Label-Free Learning on Graphs with LLMs

    May 27, 2026Safal Thapaliya, Jiatan Huang, Chuxu ZhangLLM-Assisted AnnotationText-Attributed Graph Learning

  15. Attribute-Based Diagnosis of LLM Alignment with Hate Speech Annotations

    May 26, 2026Mohammad Amine Jradi, Faeze Ghorbanpour, Alexander FraserLLM EvaluationLLM Alignment

  16. Agent-as-Peer-Debriefer: A Multi-Agent Framework with Perspective-Based Refinement for Qualitative Analysis

    May 23, 2026Zhimin Lin, Kun Cheng, Zhiyao Shu +4Multi-Agent LLM SystemsLLM-Assisted Annotation

  17. Refining and Reusing Annotation Guidelines for LLM Annotation

    May 20, 2026Kon Woo Kim, Jin-Dong Kim, Akiko AizawaLLM AlignmentLLM-Assisted Annotation

  18. iPOE: Interpretable Prompt Optimization via Explanations

    May 18, 2026Jiahui Li, Yarik Menchaca Resendiz, Sean Papay +1LLM InterpretabilityLLM-Assisted Annotation

  19. LLMs for automatic annotation of Mandarin narrative transcripts

    May 17, 2026Qingwen Zhao, Hongao Zhu, Yunqi He +3Multilingual Language Model EvaluationLLM-Assisted Annotation

  20. How to Instruct Your Robot: Dense Language Annotations Power Robot Policy Learning

    May 16, 2026Bosung Kim, Ruiyi Wang, David Acuna +5Robot Policy LearningRobot Skill Learning

  21. LLMs as annotators of credibility assessment in Danish asylum decisions: evaluating classification performance and errors beyond aggregated metrics

    May 13, 2026Galadrielle Humblot-Renaux, Mohammad N. S. Jahromi, Rohat Bakuri-Jørgensen +7Legal NLPLLM Evaluation

  22. Do Benchmarks Underestimate LLM Performance? Evaluating Hallucination Detection With LLM-First Human-Adjudicated Assessment

    May 8, 2026I. F. Atasoy, B. Mutlu, E. A. Sezer +1LLM EvaluationLLM-Assisted Annotation

  23. Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study

    May 8, 2026Moaath Alshaikh, Tasneem Alshaher, Ricardo Vieira +7Software EngineeringLLM Evaluation

  24. MultiSoc-4D: A Benchmark for Diagnosing Instruction-Induced Label Collapse in Closed-Set LLM Annotation of Bengali Social Media

    May 7, 2026Souvik Pramanik, S. M. Riaz Rahman Antu, Shak Mohammad Abyad +2LLM EvaluationLLM-Assisted Annotation