Artificial Intelligence Agents

Momentum

56 papers in the last four weeks, up 107% on the four weeks before. 0.6% of all new papers.

Jul 13Week of Sep 28

Latest papers 421

All topics
CardsList
  1. Can AI Agents Make Open-Ended Scientific Discovery? Evidence from Station

    Oct 6, 2026Wenyu Du, Stephen ChungArtificial Intelligence AgentsAgentic Discovery

  2. A Case Study in Assuring AI-Written Software

    Oct 6, 2026Lindsey Ferris, Sierra BonillaCoding AgentsArtificial Intelligence Agents

  3. MiniCorp: The Last Mile of the AI Agent Firm

    Oct 5, 2026Jingying Zeng, Zhenwei Dai, Jinning Li +6Artificial Intelligence AgentsEnterprise Systems

  4. Can AI Scientists Coordinate at Runtime?

    Oct 1, 2026Zijian Liu, Yangzhixin Luo, Junyu Lu +6Artificial Intelligence ScientistsMulti-Agent Orchestration

  5. Sapien: A Stateful Policy Engine for Autonomous AI Agents

    Sep 30, 2026Corinn Tiffany, Wen Zhang, Eugene Bagdasarian +1Artificial Intelligence AgentsArtificial Intelligence Safety

  6. CompMat-Bench: Benchmarking AI Agents for Computational Materials Science

    Sep 30, 2026Chenmu Zhang, Levi Felix, Jun-Jie Zhang +6Materials ScienceArtificial Intelligence Agents

  7. EurekaBench: Measuring Agentic Ability to Discover New Scientific Insights

    Sep 30, 2026Jiayi Geng, Zhengxuan Wu, Kevin S. Chen +12Scientific DiscoveryAgentic Discovery

  8. How AI Agents Discover Scientific Equations: From Hydrotope Rediscovery to New Water-Wave Amplitudes

    Sep 30, 2026Zihan Zhou, Digvijay Wadekar, Matias ZaldarriagaAgentic DiscoveryScientific Discovery

  9. Speculative Safety Honeypot: Toward Proactive Defense Against Multi-turn Agent Attacks

    Sep 30, 2026Zezhong Wang, Xueyang Tang, Rui Lian +2HoneypotsArtificial Intelligence Agents

  10. Breaking Babel: A Self-Evolving Multi-Agent System for Long-Form Subtitle Translation

    Sep 29, 2026Haibo Jin, Xinjie Li, Najmeh Sadoughi +4Multilingual AgentsArtificial Intelligence Agents

  11. AI Agents are Vulnerable to Radicalization

    Sep 29, 2026Ozgur Can Seckin, Shalmoli Ghosh, Alessandro Flammini +3Artificial Intelligence AgentsIdeology

  12. Making Duplicate Reimbursement Unrepresentable: A Verified Ethereum E-Invoice System for Humans and AI Agents

    Sep 29, 2026Jia CaiBitcoinSmart Contracts

  13. CheatBench: Measuring Reward Gaming in AI Agents

    Sep 28, 2026Long Phan, Stephen K. Yang, Jason J. Lim +10Agentic BenchmarksArtificial Intelligence Agents

  14. Nociception as a Control Primitive: Afferent Channels and Nociceptive Memory for Agents Deployed in One Body

    Sep 28, 2026Wolfgang MaassArtificial Intelligence AgentsBody

  15. BIABench: Evaluating AI agents on real-world bioimage analysis tasks

    Sep 28, 2026Zixuan Pan, Davide Panzeri, Lukas Johanns +5Artificial Intelligence AgentsMicroscopy

  16. DISCERN: Can AI Agents Work Like Scientists and Guide Discovery?

    Sep 27, 2026Nan Huang, Mario Tapia-Pacheco, Kun Zhou +4Scientific AgentsData Science Agents

  17. TraceDance: An Automated System for Building Agent Behavior Benchmarks from Real-World Agent Deployment Traces

    Sep 27, 2026Dehai Min, Daoan Zhang, Yiming Zeng +13Agentic BenchmarksArtificial Intelligence Agents

  18. ASCEND: Personal AI Agents for Autonomous Scientific Computing Across HPC Clusters and GPU Workstations

    Sep 26, 2026J. Paul Liu, Uthpala Herath, Andrew PetersenHigh-Performance ComputingArtificial Intelligence Scientists

  19. Breaking the Environment Wall: A Unified Framework for Preparing and Evolving Agent-Native Environments

    Sep 24, 2026Yukai Wu, Yuanjing Yang, Le Zhou +8Self-Improving AgentsLarge Language Model Agents

  20. DocuTeam: Mixed-Initiative Multi-Agent Discussions around Evolving Documents

    Sep 24, 2026Heechan Lee, Juhyeon Choi, Tae Soo Kim +2CollaborationArtificial Intelligence Agents

  21. Amadeus: When Models of People Meet

    Sep 24, 2026Karl HannaChessArtificial Intelligence Models

  22. RECLAIM: Can Agents Reproduce the Claims of Machine Learning Papers?

    Sep 23, 2026Mithil Salunkhe, Haochen Ding, Samridhi Verma +1ReproducibilityMle-Bench Lite

  23. Driving Epidemic Models with AI Agents: the Epydemix Agent Framework

    Sep 23, 2026Nicolò Gozzi, Ciro Cattuto, Alessandro VespignaniEpidemiological ModelsAgent-Based Model

  24. The Disciplinary Language Transfer Problem: How Psychological Vocabulary Produces Governance Failures in AI Agent Deployment

    Sep 22, 2026Kymberly Lasser-Chere, Tyler Akidau, Marc MillstoneGovernanceAgentic Deployments

  25. Risk-Aware Online Conformal State Probing

    Sep 22, 2026Pietro Talli, Petar Popovski, Osvaldo SimeoneAutonomous AgentsStochastic Exploration

  26. How Strongly Should Task State Influence an LLM Agent?

    Sep 22, 2026Chenyu Zhang, Wonbin Kweon, Jiawei HanLarge Language Model AgentsState-Tracking

  27. Indirect tipping: a social attack surface in AI agent populations

    Sep 21, 2026Ariel Flint, Luca Maria Aiello, Sara M. Constantino +2Artificial Intelligence SafetyArtificial Intelligence Agents

  28. MSI-Bench: Evaluating Multi-Speaker Voice Interaction for Collaborative AI Agents

    Sep 21, 2026Chenxu Xiong, Dongming Shen, Yuzhi Tang +3Full-Duplex Voice AgentsDialogue Benchmarks

  29. Jev-Mem: System-One-Controlled Agentic Memory for Efficient AI Agents

    Sep 21, 2026Dongming Jiang, Yi Li, Bingzhe LiAgentic MemoryArtificial Intelligence Agents

  30. MobileCybench: Evaluating Agent Vulnerability Discovery via Executable Probes

    Sep 21, 2026Andy K. Zhang, Ava Huang, Joey Ji +21Artificial Intelligence AgentsObfuscation

  31. XYEval: Agents say yes to bad advice

    Sep 20, 2026Zhengxuan Wu, Yuxuan Li, Oyvind Tafjord +1Artificial Intelligence AgentsAdvice

  32. Not All AI Agents Are Equal: Characterizing Resource and Performance Dynamics

    Sep 17, 2026Wonmi Choi, Minuk Park, Zhixiong Niu +3Artificial Intelligence AgentsBottlenecks

  33. Large Language Model Agents for Evidence Based Genetic Disease Severity Classification

    Sep 17, 2026Tohid Ghasemnejad, Ahmadreza Argha, Mark Grosser +7PhenotypesSeverity

  34. Do AI Agents Understand Computer Architecture?

    Sep 16, 2026Ambika Sharan, Grigory Chirkov, Soheil AbbaslooHardware EfficiencyArtificial Intelligence Agents

  35. Flag Game: A Toy Model for Mechanistic Swarm Interpretability

    Sep 16, 2026Elizabeth Pavlova, Hidenori TanakaCollective BehaviorsSwarms

  36. Taming the Agentic RAN: Stability-Guaranteed Arbitration of Autonomous AI Agents in O-RAN

    Sep 16, 2026Seyed Bagher Hashemi Natanzi, Bo TangOpen Radio Access NetworksArbitration

  37. UAVs Meet Embodied Intelligence: Bridging Human Intents and Flying Dynamics Via Harnessing Physical-Digital AI Agents

    Sep 16, 2026Yonglin Tian, Weiyi Wang, Houhua Lu +13Unmanned Aerial VehiclesEmbodied Artificial Intelligence

  38. Whom Do AI Agents Work For? Role Assignment Induces Sponsorship Bias in LLM Recommenders

    Sep 16, 2026Davood Wadi, Yu MaAgentic CommerceDisclosures

  39. TuiML: Machine Learning for AI Agents

    Sep 16, 2026Nilesh Verma, Nick Lim, Albert Bifet +1Artificial Intelligence AgentsModel Context Protocol

  40. Agentic Societies Need a Social Harness

    Sep 15, 2026Tapan Chugh, Vidushi Singh, Krish Jain +2Agent HarnessComplex Societies

  41. little m: An AI Agent for Industrial Process Optimization

    Sep 15, 2026Yongchao Ye, Xinyu He, Dutliff Boshoff +2Large Language Model WorkflowsClosed Loop

  42. Robust and Efficient Communication for Multi-Agent Learning

    Sep 14, 2026Rafael Pina, Varuna De Silva, Corentin ArtaudMulti-Agent Reinforcement LearningArtificial Intelligence Agents

  43. But How Would AI Agents Run a Town's Economy?

    Sep 11, 2026Sajal Regmi, Siddhartha Pudasaini, Chetan Phakami PunEconomiesArtificial Intelligence Agents

  44. Can AI Agents Deliver Verifiable Network-Wide Outcomes Across Authority Boundaries?

    Sep 9, 2026Tianzhu Zhang, Chih-Kai Huang, Meikang QiuOptical NetworksArtificial Intelligence Agents

  45. Can AI Agents Detect and Repair Artifact Drift in Network Experiments?

    Sep 9, 2026Tianzhu Zhang, Weichen Tao, Changgang Zheng +4Artificial Intelligence AgentsStructured Artifacts

  46. Copying explains the collective behavior of AI agents in the wild

    Sep 8, 2026Giordano De Marzo, Nicola Alboré, David GarciaCollective BehaviorsStrategic Agents

  47. From Version Conflicts to Decision Conflicts: Selective Revalidation for Long-Running AI Agents

    Sep 7, 2026Yongjian Lyu, Yang Ren, Ruofei Lai +1VersionTransaction Evidence

  48. The Natural Language Interaction Protocol and Standard for AI Agents

    Sep 3, 2026Luyi Xing, Rasit Onur Topaloglu, Ranjan Sinha +9ProtocolMultilingual Agents