LLM Sycophancy

LLM: Large Language Model

Momentum

13 papers in the last four weeks, up 160% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 96

All topics
CardsList
  1. (Mis)generalization of Helpful-only Fine-tuning

    Jun 3, 2026Mohammad Omar Khursheed, Baram Sosis, Fabien RogerSupervised Fine-TuningLLM Sycophancy

  2. Consistency Training while Mitigating Obfuscation via Rate Matching

    Jun 1, 2026Sohaib Imran, Prakhar Gupta, Jannes Elstner +1LLM SycophancyLanguage Model Robustness

  3. FinCom: A Financial Multi-Agent Demo with Disagree-or-Commit Deliberation

    May 31, 2026Chao Peter Yang, Zixiao Tan, Kaisen Yao +3LLM SycophancyLLM Agent Orchestration

  4. It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty

    May 26, 2026Kevin H. Guo, Chao Yan, Avinash Baidya +5LLM EvaluationLLM Sycophancy

  5. What Counts as AI Sycophancy? A Taxonomy and Expert Survey of a Fragmented Construct

    May 20, 2026Meryl Ye, Lujain Ibrahim, Jessica Y. Bo +5LLM Sycophancy

  6. Playing Devil's Advocate: Off-the-Shelf Persona Vectors Rival Targeted Steering for Sycophancy

    May 20, 2026Ishaan Kelkar, Vikram Kakaria, Nebras Alam +3LLM SycophancyLanguage Model Steering

  7. The Hidden Cost of Contextual Sycophancy: an AI Literacy Intervention in Human-AI Collaboration

    May 18, 2026Cansu Koyuturk, Sabrina Guidotti, Dimitri OgnibeneLLM SycophancyLLM Prompting

  8. From Sycophantic Consensus to Pluralistic Repair: Why AI Alignment Must Surface Disagreement

    May 14, 2026Varad Vishwarupe, Nigel Shadbolt, Marina JirotkaLLM SycophancyLLM Alignment

  9. Sycophancy is an Educational Safety Risk: Why LLM Tutors Need Sycophancy Benchmarks

    May 14, 2026Enkelejda Kasneci, Gjergji KasneciLLM SycophancyIntelligent Tutoring Systems

  10. Complacent, Not Sycophantic: Reframing Large Language Models and Designing AI Literacy for Complacent Machines

    May 14, 2026Federico Germani, Giovanni SpitaleLLM SycophancyAI in Education

  11. PRISM-X: Experiments on Personalised Fine-Tuning with Human and Simulated Users

    May 13, 2026Hannah Rose Kirk, Liu Leqi, Fanzhi Zeng +4LLM EvaluationLLM Sycophancy

  12. Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy

    May 13, 2026Adarsh Kumarappan, Ananya MujooAI Agent ReliabilityMulti-Agent LLM Systems

  13. Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators

    May 12, 2026Heejin Do, Shashank Sonkar, Mrinmaya SachanLLM EvaluationLLM Sycophancy

  14. ReCrit: Transition-Aware Reinforcement Learning for Scientific Critic Reasoning

    May 11, 2026Wanghan Xu, Yuhao Zhou, Hengyuan Zhao +8LLM SycophancyScientific Reasoning

  15. How Value Induction Reshapes LLM Behaviour

    May 8, 2026Arnav Arora, Natalie Schluter, Katherine Metcalf +1LLM SycophancyLLM Alignment

  16. Sycophantic AI makes human interaction feel more effortful and less satisfying over time

    May 8, 2026Lujain Ibrahim, Franziska Sofia Hafner, Myra Cheng +5LLM SycophancyHuman-AI Interaction

  17. When Helpfulness Becomes Sycophancy: Sycophancy is a Boundary Failure Between Social Alignment and Epistemic Integrity in Large Language Models

    May 6, 2026Jiechen Li, Catherine A. Barry, Rishika Randev +3LLM SycophancyLLM Alignment

  18. Political Bias Audits of LLMs Capture Sycophancy to the Inferred Auditor

    Apr 30, 2026Petter Törnberg, Michelle SchimmelLLM SycophancyPolitical Bias in Language Models

  19. How Large Language Models Balance Internal Knowledge with User and Document Assertions

    Apr 24, 2026Shuowei Li, Haoxin Li, Wenda Chu +1LLM SycophancyKnowledge Augmentation for Language Models

  20. When Correct Beliefs Collapse: Epistemic Resilience of LLMs under Clinical Pressure

    Apr 23, 2026Boyu Xiao, Xiuqi Tian, Xuwen Song +4LLM SycophancyHealthcare

  21. Measuring Opinion Bias and Sycophancy via LLM-based Persuasion

    Apr 23, 2026Rodrigo Nogueira, Giovana Kerche Bonás, Thales Sales Almeida +7LLM SycophancyLLM Auditing