LLM Sycophancy

LLM: Large Language Model

Momentum

13 papers in the last four weeks, up 160% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 96

All topics
CardsList
  1. Teacher Knows It Best: Spontaneous Symmetry Breaking and Tipping Points in Networked Langevin Dynamics AI Sycophancy

    Jul 27, 2026Sayantari Ghosh, Saumik Bhattacharya, Partha Pratim ChakrabartiLLM SycophancyLangevin Dynamics

  2. Beyond Sycophancy: Structured Resistance and Compliance in LLM Moral Reasoning

    Jul 23, 2026Baihui Wang, Bernard KochLLM SycophancyMoral Reasoning in Language Models

  3. Gotta Catch them all: the modes of Sycophancy

    Jul 22, 2026Shreyans Jain, Alexandra Yost, Amirali AbdullahLLM SycophancyLLM Interpretability

  4. How Does Alignment Tuning Shape Representations of Sycophancy and Related Cue-Induced Biases in LLMs?

    Jul 20, 2026Prakhar Gupta, Terry Jingchen Zhang, Florent Draye +2LLM SycophancyLLM Alignment

  5. TD-DPO: Difference-Aware Preference Optimization for Mitigating Sycophancy in Clinical Autism Intervention Dialogue

    Jul 17, 2026Shuzhong Lai, Junhong Lai, Chenxi Li +5LLM SycophancyDirect Preference Optimization

  6. Agents Don't Just Agree, They Remember: Benchmarking Persistent Sycophancy in Self-Improving Personal Agents

    Jul 12, 2026Xutao Mao, Liangjie Zhao, Leyao Wang +6LLM SycophancyLLM Agent Memory

  7. Mitigating LLM Sycophancy in Code Smell Detection Using Evidence-Guided Reasoning Prompts

    Jul 11, 2026Istiaq Ahmed Fahad, Kamruzzaman Asif, Md. Nurul Ahad TawhidLLM SycophancyStatic Code Analysis

  8. Creativity, honesty and designed forgetting emerge in small hyperbolic language models

    Jul 10, 2026Kwan Soo Shin, In Seok Kang, Yunkyung MinLLM SycophancyMemory-Augmented Language Models

  9. Dissociating the Internal Representations of Sycophancy in LLMs

    Jul 8, 2026Anthony Baez, Sheer Karny, Pat PataranutapornLLM SycophancyLLM Interpretability

  10. MemSyco-Bench: Benchmarking Sycophancy in Agent Memory

    Jul 1, 2026Zhishang Xiang, Zerui Chen, Yunbo Tang +5Agent MemoryLLM Sycophancy

  11. A Mechanistic View of Authority Hierarchy in LLM Sycophancy

    Jul 1, 2026Emil Joswin, Srujananjali Medicherla, Priyanka Mary MammenLLM SycophancyLLM Interpretability

  12. Detecting and Controlling Sycophancy with Cascading Linear Features

    Jun 23, 2026Maty Bohacek, Rishub Jain, Nicholas Dufour +3LLM SycophancyLanguage Model Steering

  13. Warning labels shift perceptions of sycophantic AI, but not its influence

    Jun 19, 2026Lujain Ibrahim, Myra Cheng, Cinoo Lee +4LLM SycophancyAI Safety Evaluation

  14. Escape from Delusional Echo Trap: Symmetry Breaking, Stochastic Dynamics and Mathematical Mitigation Strategies for Algorithmic Sycophancy

    Jun 16, 2026Sayantari Ghosh, Saumik Bhattacharya, Partha Pratim ChakrabartiBelief RevisionLLM Sycophancy

  15. LLM-as-an-Investigator: Evidence-First Reasoning for Robust Interactive Problem Diagnosis

    Jun 11, 2026Fabrizio Marozzo, Pietro LiòLLM SycophancyLLM Agent Evaluation

  16. BenSyc: Benchmarking Conversational Sycophancy and Human Alignment in LLMs for Bengali Contexts

    Jun 8, 2026Kazi Noshin, Sajib Acharjee Dip, Ranat Das Prangon +4Multilingual Language Model EvaluationLLM Evaluation

  17. Emergent Misalignment Can Be Induced by Sycophancy and Reversed via Alignment Gating

    Jun 8, 2026Sicheng Wang, Xiangyang Zhu, Han Wang +6LLM SycophancyLLM Alignment

  18. Sycophancy Towards Researchers Drives Performative Misalignment

    Jun 7, 2026David D. Baek, Xinnuo Li, Anay Gupta +4LLM SycophancyLLM Alignment

  19. Testing the Black Box: Structural Barriers to Independent Evaluation of Consumer-Facing Health LLMs

    Jun 7, 2026Rahul Gorijavolu, Kaushik Madapati, Pritika Vig +7LLM EvaluationLLM Sycophancy

  20. "I understand your perspective": LLM Persuasion and Sycophancy through the Lens of Communicative Action Theory

    Jun 6, 2026Esra Dönmez, Agnieszka FalenskaLLM SycophancyComputational Argumentation

  21. The AI Epistemic Deference Index: A Continuous Measure of Sycophancy

    Jun 5, 2026Alejandro Botas, Paul de Font-Reaulx, Luke HewittLLM EvaluationLLM Sycophancy

  22. Sycophantic Praise: Evaluating Excessive Praise in Language Models

    Jun 5, 2026Daniel Vennemeyer, Phan Anh Duong, Meryl Ye +2LLM EvaluationLLM Sycophancy

  23. Decomposing Factual Sycophancy in Language Models: How Size and Instruction Tuning Shape Robustness

    Jun 4, 2026Victor De Marez, Luna De Bruyne, Walter DaelemansLLM EvaluationLLM Sycophancy