AI Coding Agents

Momentum

28 papers in the last four weeks, up 115% on the four weeks before. 0.3% of all new papers.

Jul 13Week of Sep 28

Latest papers 212

All topics
CardsList
  1. Self Improvement via Fast Tree-search

    Sep 17, 2026Xinghong Fu, Aravinth Kulanthaivelu, Yutaro YamadaInference-Time SearchAI Coding Agents

  2. Reflections on Trusting Trust, Revisited: Contaminating Self-Modifying AI Coding Agents with Poisoned Benchmarks

    Sep 15, 2026Franziska Roesner, Tadayoshi KohnoAI Coding AgentsAI Agent Security

  3. MTAC-IFBench: Benchmarking Instruction-Following in Multi-Turn Agentic Coding

    Sep 14, 2026Bosi Wen, Cunxiang Wang, Jiayi Gui +6Software Engineering AgentsAI Coding Agents

  4. RepoNav: From Snippet Retrieval to File-Centered Repository Navigation for Code Agents

    Sep 8, 2026Hongzheng Chai, Jiakun Li, Hongyue Yu +1AI Coding AgentsAgentic Retrieval

  5. Do AI Coding Assistants Check Before They Install? A Pre-Registered Demand-Side Audit of Trust Signals in the Research Software Supply Chain

    Sep 7, 2026Pengyin ShanAlgorithmic AuditingAI Coding Agents

  6. ττ\tau^\tau-Bench: An Environment for End-To-End, Realistic Agent Construction

    Sep 7, 2026Quan Shi, Keshav Dhandhania, Karthik Narasimhan +1AI Coding AgentsLLM Agent Evaluation

  7. Measure Before You Manage: Evaluating Agent Working Memory in Coding Agents

    Aug 31, 2026Le Chen, Zishen Wan, Baixi Sun +6Agent MemoryAI Coding Agents

  8. Beyond the Payload: How User Invocation Shapes Coding Agent Vulnerability to Repository Poisoning

    Aug 31, 2026Fukang Zhu, Binbin Zhao, Ruixiao Lin +3Prompt SensitivityAI Coding Agents

  9. VibeJam: A User Study Platform for Web Development with Agents

    Aug 30, 2026Nishant Balepur, Connor Baumler, Valerie Chen +3AI Coding AgentsHuman-AI Collaboration

  10. Cost-Effective Repository Exploration for Agentic Issue Localization

    Aug 30, 2026Mohammad Nour Al Awad, Sergey IvanovAI Coding AgentsAgentic Search

  11. A Few Pages of Markdown: Committed AI Configuration and Lower Quality Cost after Coding-Agent Adoption

    Aug 26, 2026Yegor Denisov-Blanch, Shyam Agarwal, Pavel Azaletskiy +5Software EngineeringAI Coding Agents

  12. AutoDesign: Meta-Harness Optimization for Long-Horizon Agentic Design

    Aug 13, 2026Yaxin Luo, Haobin Jiang, Jialv Zou +11AI Coding AgentsAgent Harness Optimization

  13. Harness-IF: Evaluating Instruction Following Across Instruction Surfaces in Coding Agents

    Aug 12, 2026Zining Huang, Haoran Que, Hong Zeng +8AI Coding AgentsLLM Agent Evaluation

  14. SWE-Bench ProMax: Benchmarking Agents on Large-Scale Multilingual Code Refactoring

    Aug 10, 2026Yuling Shi, Jinghan Xu, Kelin Fu +12AI Coding AgentsSoftware Engineering Benchmarks

  15. Can Coding Agents Solve Repository-Level Issues with Rendered Code? An Exploratory Study of Visual Representations

    Aug 10, 2026Weijie Liang, Yuanfeng Song, Xing Chen +3Automated Program RepairAI Coding Agents

  16. A Unified Issue Resolution Benchmark for Requirement Clarification, Planning, and Code Generation for Coding Agents

    Aug 10, 2026Xin Zhou, Chun Yong Chong, Kisub Kim +11Requirements EngineeringAI Coding Agents

  17. The Scaffolding Matters More Than the Interface: A Controlled Comparison of MCP and CLI Tool Use Across Seven Agent Scaffoldings, Five Language Models, and One Software Task

    Aug 9, 2026Marc Alier Forment, María José Casañ Guerrero, Francisco José García-Peñalvo +1AI Coding AgentsLLM Agent Evaluation

  18. Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution

    Aug 8, 2026Anton Razzhigaev, Andrei Gritsaev, Andrei Kaznacheev +3AI Coding AgentsAgent Harness Optimization

  19. CodeGrep: An RL-Trained Retrieval Agent for LLM Coding Agents

    Aug 6, 2026Wuya Chen, Yihao yang, Yang Cao +1AI Coding AgentsLLM Agent Workflow Optimization

  20. Active-SWE: Benchmarking Coding Agents for Proactive Bug Fixing without Issue Reports

    Aug 5, 2026Haobin Li, Ping Deng, Weizhong Qian +4Automated Program RepairAI Coding Agents

  21. From Social Coding to Agentic Coding: Productivity and Relational Reconfiguration in Open-Source Communities

    Aug 4, 2026Mengying Zhou, Yongjie Yin, Yang ChenSoftware EngineeringAI Coding Agents

  22. Coding Agents as Test-Suite Auditors: Finding What Official Suites Miss While Approaching What They Catch

    Aug 3, 2026Shuyang Xie, Shuxiao Xie, Feng Zhu +2Automated Software TestingBenchmark Auditing

  23. LoopsBench: From Harness Engineering to Loop Engineering in Coding Agent Evaluation

    Jul 31, 2026Han Li, Zhemin Fang, Rili Feng +8AI Coding AgentsSoftware Engineering Benchmarks