Agentic Deployments

Momentum

11 papers in the last four weeks, up 83% on the four weeks before. 0.1% of all new papers.

Jul 6Week of Sep 21

Latest papers 62

All topics
CardsList
  1. From Migration to Calibration: Preserving Agent Capabilities across Models, Jurisdictions, and Scale

    Sep 28, 2026Yaxiao Liu, Pengbo Liu, Yiwen Liu +2Agentic DeploymentsHandoff

  2. When Valid Tool Calls Change Meaning: Formation-Consistent Dispatch for LLM Agents

    Sep 28, 2026Geonwoo Kim, Brent ByungHoon KangAgentic DeploymentsCall

  3. Before Acting, Change the State: Prospective State Intervention for Web Agents under Deceptive Interfaces

    Sep 28, 2026Ruozhao Yang, Mingfei Cheng, Xiaofei XieWeb AgentsAgentic Deployments

  4. AgentWare: Automating the Lifecycle of Agentic Applications across the Edge-to-Cloud Continuum

    Sep 28, 2026Michalis Kasioulis, Moysis Symeonides, George Pallis +1Agentic DeploymentsEdge Platforms

  5. The Disciplinary Language Transfer Problem: How Psychological Vocabulary Produces Governance Failures in AI Agent Deployment

    Sep 22, 2026Kymberly Lasser-Chere, Tyler Akidau, Marc MillstoneGovernanceAgentic Deployments

  6. Trains but Doesn't Learn: A Post-Training Delivery Benchmark for LLM Agents as Forward-Deployed Engineers

    Sep 21, 2026Weihang Ding, Junfei ZhanAgentic DeploymentsLarge Language Model Agents

  7. ActGov: Governing LLM Agent Actions via Policy-Constrained Validation

    Sep 21, 2026Kaiyuan Zhang, Yuke Peng, Ke Jiang +1AuthorizationRuntime Enforcement

  8. K-Bench: A Benchmark for LLM Unlearning in Agentic Deployments

    Sep 11, 2026Guangsheng Yu, Yanna Jiang, Qin Wang +2Large Language Model UnlearningAgentic Deployments

  9. From Prior-Guided Heuristics to Deployable Agents: Accelerating Demonstration-Driven Reinforcement Learning for Deadline-Constrained Network Control

    Sep 3, 2026Vincenzo Norman Vitale, Mohammad Solki, Antonia Maria Tulino +2Bringing Network CodingLearning-Based Control

  10. Capability-Gated Language Models: Security Composes, Utility Does Not

    Aug 31, 2026Patrikas Vanagas, Augustas Mačijauskas, Laurynas LopataLarge Language Model SafetyAccess Control

  11. The Irreversibility Budget: Fleet-Level Risk Accounting and Admission Control for Agent Operating Systems

    Aug 31, 2026Bardia Mohammadi, Laurent BindschaedlerRuntime Safety FilteringAgentic Deployments

  12. Agent Safety Should Be a Runtime Contract

    Aug 11, 2026Albus W. Ng, Yi Han, Jusheng Zhang +1Artificial Intelligence SafetyAgentic Deployments

  13. When Do Prompt-Side Agent Playbooks Transfer? Accuracy, Cost, and Runtime Shift in Agent Deployment

    Aug 6, 2026Weihong Lin, Lin Sun, Xiangzheng ZhangAgentic DeploymentsFrozen

  14. RouteGuard: Certifying Routing Gain in LLM Multi-Agent Systems When Complementarity Is Not Enough

    Aug 5, 2026Anchen Sun, Kaiqi YangLarge Language Model RoutingMulti-Agent Large Language Model Systems

  15. Formal Verification of Agentic Systems over Operational Data

    Aug 4, 2026Alejandro J. Mercado, Alessio LomuscioAgentic DeploymentsFormal Verification

  16. Securing Agentic AI: From Per-Action Checks to Trajectory Assurance

    Aug 3, 2026Alireza Lotfi, Subangkar Karmaker Shanto, Imtiaz Karim +1Agentic DeploymentsDelegation

  17. Where Is the Cost of Third-Party API Routers in Agentic Software Development?

    Jul 26, 2026Donghao Fu, Jingxin Li, Xue Jiang +1Coding AgentsAgentic Deployments

  18. Are You Still the Agent I Authorized? Earned Authority under a Fixed Ceiling for Evolving Agents

    Jul 26, 2026Zhaoxi Zhang, Xiaomei ZhangAuthorizationAuthority

  19. ToolGuardian: Declarative Security for AI Agent-Tool Interactions

    Jul 23, 2026Arun Ravindran, Saurabh DeochakeAuthorizationAgentic Deployments

  20. Operational Hallucination and Safety Drift in AI Agents

    Jul 20, 2026Shasha Yu, Fiona Carroll, Barry L. BentleyArtificial Intelligence SafetyAgentic Deployments

  21. Democratizing Agent Deployment Safety: A Structural Monitoring Approach

    Jul 16, 2026Preeti Ravindra, Rahul Tiwari, Vincent WolowskiAgentic DeploymentsSecurity Evaluation

  22. ToolAlignBench: Investigating Alignment Conflicts in Tool-Calling Enabled LLMs

    Jul 15, 2026Aryan Keluskar, Amrita Bhattacharjee, Huan LiuLarge Language Model SafetyAgentic Deployments

  23. ANCHOR: Automated Alignment Auditing for CLI Agents on Real-World Harm

    Jul 11, 2026Kefan Song, Yanjun QiAgentic DeploymentsMalicious Agents

  24. Token-Flow Firewall: Semantic Runtime Auditing for Persistent AI Agents

    Jul 9, 2026Puji Wang, Yingchen Zhang, Ruqing Zhang +2Agentic DeploymentsSemantic Firewall

  25. Budgeted Act-or-Defer Multi-Agent LLM Deliberation with Local Reliability Bounds

    Jun 28, 2026Mengdie Flora Wang, Haochen Xie, Guanghui Wang +2DeliberationAgentic Deployments

  26. Narration-of-Thought: Inference-Time Scaffolding for Defeasible Ethical Reasoning in Large Language Models

    Jun 24, 2026Patrick Cooper, Alvaro VelasquezEthicsMoral Reasoning

  27. EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments

    Jun 11, 2026Jundong Xu, Qingchuan Li, Jiaying Wu +11Large Language Model AgentsAgentic Deployments

  28. Token Budgets: An Empirical Catalog of 63 LLM-Agent Budget-Overrun Incidents, with an Affine-Typed Rust Mitigation as a Case Study

    Jun 2, 2026Sajjad KhanToken Budget AllocationAgentic Deployments

  29. Toward Pre-Deployment Assurance for Enterprise AI Agents: Ontology-Grounded Simulation and Trust Certification

    Jun 2, 2026Thanh Luong Tuan, Abhijit SanyalAgentic DeploymentsPre-Deployment Safety Assessments

  30. Adaptive Auto-Harness: Sustained Self-Improvement for Agentic System Deployment on Open-Ended Task Streams

    Jun 1, 2026Zewen Liu, Zhan Shi, Yisi Sang +7Evolving HarnessMulti-Agent Evolution

  31. ICAN-Deploy: Identity-Stable Canary Deployment for Safety-Critical Embodied Agents

    May 27, 2026Xue Qin, Simin Luan, John See +3Agentic DeploymentsMicroservices

  32. Your Agents Are Aging Too: Agent Lifespan Engineering for Deployed Systems

    May 25, 2026Jianing Zhu, Yeonju Ro, John Robertson +5Agentic DeploymentsLong-Term Agent Memory

  33. E3E^3-Agent: An Executable and Evolving Agent for Resource Management of Edge Generative Inference

    May 21, 2026Rui Bao, Yaping Sun, Zhiyong Chen +4Agentic InferenceAgentic Deployments

  34. PocketAgents: A Manifest-Driven Library of Autonomous Defense Agents

    May 20, 2026Sidnei Barbieri, Ágney Lopes Roth Ferraz, Lourenço Alves Pereira JúniorAgentic DeploymentsDeception

  35. Agent Meltdowns: The Road to Hell Is Paved with Helpful Agents

    May 18, 2026Rishi Jha, Harold Triedman, Arkaprabha Bhattacharya +1Agentic DeploymentsLatent Failure Patterns

  36. Position: A Three-Layer Probabilistic Assume-Guarantee Architecture Is Structurally Required for Safe LLM Agent Deployment

    May 18, 2026S. Bensalem, Y. Dong, M. Franzle +6Agentic DeploymentsLarge Language Model Agents

  37. MATRA: Modeling the Attack Surface of Agentic AI Systems -- OpenClaw Case Study

    May 11, 2026Tim Van hamme, Thomas Vissers, Javier Carnerero-Cano +4Realistic Threat ModelArtificial Intelligence Risk

  38. The Granularity Mismatch in Agent Security: Argument-Level Provenance Solves Enforcement and Isolates the LLM Reasoning Bottleneck

    May 11, 2026Linfeng Fan, Ziwei Li, Yuan Tian +3Untrusted ContentAgentic Deployments

  39. Securing Computer-Use Agents: A Unified Architecture-Lifecycle Framework for Deployment-Grounded Reliability

    May 8, 2026Zejian Chen, Zhanyuan Liu, Chaozhuo Li +6Computer-Use AgentsAgentic Deployments

  40. Hybrid Inspection and Task-Based Access Control in Zero-Trust Agentic AI

    May 4, 2026Majed El Helou, Benjamin Ryder, Chiara Troiani +3Access ControlAgentic Deployments

  41. Ambient Persuasion in a Deployed AI Agent: Unauthorized Escalation Following Routine Non-Adversarial Content Exposure

    Apr 29, 2026Diego F. Cuadros, Abdoul-Aziz MaigaAgentic DeploymentsEscalation