Instruction Following

Momentum

11 papers in the last four weeks, down 8% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 85

All topics
CardsList
  1. Dimension-Level Intent Fidelity Evaluation for Large Language Models: Evidence from Structured Prompt Ablation

    May 14, 2026GAng PengMultilingual Language Model EvaluationLLM Evaluation

  2. When Attention Closes: How LLMs Lose the Thread in Multi-Turn Interaction

    May 13, 2026Vardhan Dongre, Joseph Hsieh, Viet Dac Lai +3Self-AttentionLLM Interpretability

  3. Premover: Fast Vision-Language-Action Control via Early Execution During Instruction Delivery

    May 12, 2026Joonha Park, Jiseung Jeong, Taesik GongVisuomotor Policy LearningAI Control

  4. Instructions Shape Production of Language, not Processing

    May 11, 2026Andreas Waldis, Leshem Choshen, Yufang Hou +1LLM InterpretabilityInstruction Following

  5. SEIF: Self-Evolving Reinforcement Learning for Instruction Following

    May 8, 2026Qingyu Ren, Qianyu He, Jiajie Zhu +7RL for Language ModelsReinforcement Learning

  6. Think-with-Rubrics: From External Evaluator to Internal Reasoning Guidance

    May 8, 2026Jiachen Yu, Zhihao Xu, Junjie Wang +1Instruction FollowingRubric-Based Evaluation

  7. SEQUOR: A Multi-Turn Benchmark for Realistic Constraint Following

    May 7, 2026Beatriz Canaverde, Duarte M. Alves, José Pombal +2Instruction FollowingGenerative AI Evaluation

  8. Decomposing the Basic Abilities of Large Language Models: Mitigating Cross-Task Interference in Multi-Task Instruct-Tuning

    May 7, 2026Bing Wang, Ximing Li, Changchun Li +3LLM Fine-TuningGradient Interference

  9. MCJudgeBench: A Benchmark for Constraint-Level Judge Evaluation in Multi-Constraint Instruction Following

    May 5, 2026Jaeyun Lee, Junyoung Koh, Zeynel Tok +2LLM EvaluationLLM-as-a-Judge

  10. When LLMs Stop Following Steps: A Diagnostic Study of Procedural Execution in Language Models

    May 1, 2026Sailesh Panda, Pritam Kadasi, Abhishek Upperwal +1LLM EvaluationInstruction Following

  11. From Coarse to Fine: Benchmarking and Reward Modeling for Writing-Centric Generation Tasks

    Apr 30, 2026Qingyu Ren, Tianjun Pan, Xingzhou Chen +1Reward ModelingRL for Language Models

  12. Compliance versus Sensibility: On the Reasoning Controllability in Large Language Models

    Apr 29, 2026Xingwei Tan, Marco Valentino, Mahmud Elahi Akhter +3LLM InterpretabilityInstruction Following

  13. Omission Constraints Decay While Commission Constraints Persist in Long-Context LLM Agents

    Apr 22, 2026Yeran GamageAI Agent MonitoringLLM Security

  14. Measuring Distribution Shift in User Prompts and Its Effects on LLM Performance

    Apr 19, 2026Parker Seegmiller, Sarah Masud PreumLLM EvaluationInstruction Following

  15. TinyJudge: Unverifiable Constraint Alignment via Lightweight Specialist Ensembles

    Apr 19, 2026Yirong Zeng, Yufei Liu, Xiao Ding +9Reward ModelingLLM Alignment

  16. Mini-BEHAVIOR-Gran: Revealing U-Shaped Effects of Instruction Granularity on Language-Guided Embodied Agents

    Apr 18, 2026Sukai Huang, Chenyuan Zhang, Fucai Ke +4AI Agent EvaluationAI Agent Benchmarks

  17. Reasoning Up the Instruction Ladder for Controllable Language Models

    Oct 30, 2025Zishuo Zheng, Vidhisha Balachandran, Chan Young Park +2LLM Safety AlignmentInstruction Following

  18. COCORELI: Enforcing Execution Preconditions for Reliable Collaborative Instruction Following

    Aug 29, 2025Swarnadeep Bhar, Omar Naim, Eleni Metheniti +4Instruction FollowingLLM Agent Planning

  19. RECAST: Expanding the Boundaries of LLMs' Complex Instruction Following with Multi-Constraint Data

    May 25, 2025Zhengkang Guo, Wenhao Liu, Mingchen Xie +13LLM Fine-TuningSynthetic Data Generation

  20. Measuring Pragmatic Influence in Large Language Model Instructions

    Date pendingYilin Geng, Omri Abend, Eduard Hovy +1LLM EvaluationLLM Prompting

  21. How LLMs Follow Instructions: Skillful Coordination, Not a Universal Mechanism

    Date pendingElisabetta Rocchetti, Alfio FerraraLLM InterpretabilityInstruction Following