AI Agent Monitoring

Momentum

20 papers in the last four weeks, up 186% on the four weeks before. 0.2% of all new papers.

Jul 13Week of Sep 28

Latest papers 170

All topics
CardsList
  1. Token-Flow Firewall: Semantic Runtime Auditing for Persistent AI Agents

    Jul 9, 2026Puji Wang, Yingchen Zhang, Ruqing Zhang +2AI Agent SecurityAI Agent Monitoring

  2. Persuasion Attacks Can Decrease Effectiveness of CoT Monitoring

    Jul 9, 2026Jennifer Za, Julija Bainiaksina, Nikita Ostrovsky +2Language Model Safety EvaluationAdversarial Attacks on LLMs

  3. Agent Delivery Engineering Predictive Reliability Framework

    Jul 8, 2026Dexing LiuAI Agent ReliabilityMulti-Agent LLM Systems

  4. Multi-Agent AI Control: Distributed Attacks Hamper Per-Instance Monitors

    Jul 8, 2026Oliver Makins, Orazio Angelini, Zohreh Shams +1Multi-Agent ControlAI Agent Security

  5. When Agents Go Rogue: Activation-Based Detection of Malicious Behaviors in Multi-Agent Systems

    Jul 7, 2026Haowen Xu, Xue Tan, Lei Ma +6LLM Agent SecurityAI Agent Monitoring

  6. Distributed Attacks in Persistent-State AI Control

    Jul 2, 2026Josh Hills, Ida Caspary, Asa Cooper SticklandAI Agent SecurityAI Agent Security Benchmarks

  7. Behind the Refusal: Determining Guardrail Activation via Behavioral Monitoring

    Jul 2, 2026William Hackett, Peter GarraghanLLM AuditingLLM Guardrails

  8. LabGuard: Grounding Natural-Language Laboratory Rules into Runtime Guards for Embodied Laboratory Agents

    Jun 30, 2026Jingpu Yang, Fengxian Ji, Zhengzhao Lai +8AI Agent SafetyAI Agent Monitoring

  9. MESA: Prioritizing Vulnerable Communication Channels for Securing Multi-Agent Systems

    Jun 29, 2026Kunyang Li, Kyle Domico, Jonathan Gregory +1AI Agent MonitoringMulti-Agent System Security

  10. Training Observable Control Policies to Expose Agent State Through Actions

    Jun 25, 2026Andres Enriquez Fernandez, John J. BirdAI Agent MonitoringState Tracking

  11. Litmus: Zero-Label, Code-Driven Metric Specification for Evaluating AI Systems

    Jun 22, 2026Prajjwal Gupta, Prasang Gupta, Vishal Bhutani +4AI Agent EvaluationAI Agent Monitoring

  12. Efficient and Sound Probabilistic Verification for AI Agents

    Jun 18, 2026Alaia Solko-Breslin, Pramod Kaushik Mudrakarta, Mihai Christodorescu +2AI Agent SecurityAI Agent Monitoring

  13. DeXposure-Claw: An Agentic System for DeFi Risk Supervision

    Jun 17, 2026Aijie Shu, Bowei Chen, Wenbin Wu +2Financial ServicesAI Agent Monitoring

  14. Execution-bound advisory automation for agentic AI: a reproducible AIBOM-driven CSAF-VEX framework

    Jun 16, 2026Petar Radanliev, Omar Santos, Carsten Maple +1CybersecuritySoftware Security

  15. Agent Behavior Mining: Generative AI Agent Governance in Business Processes

    Jun 12, 2026Hoang Vu, Maximilian Körner, Adrian Rebmann +4Process MiningBusiness Process Automation

  16. The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment

    Jun 9, 2026Filippo Tonini, Federico Torrielli, Anton Danholt Lautrup +3AI Agent AuditingLLM Auditing

  17. MIRAGE: A Polarity-Flipping Encoding Subspace in LLM Agents

    Jun 9, 2026Pratibha Revankar, Kargi Chauhan, Jihye Kim +3LLM InterpretabilityAI Agent Security

  18. Projecting the Emerging Mindset of SWE Agent by Launching a Wild Code Understanding Journey

    Jun 7, 2026Zhengyi Zhuo, Yan LiuSoftware EngineeringSoftware Engineering Agents