AI Agent Monitoring

Momentum

20 papers in the last four weeks, up 186% on the four weeks before. 0.2% of all new papers.

Jul 13Week of Sep 28

Latest papers 170

All topics
CardsList
  1. TRACE: Trajectory Reasoning through Adaptive Cross-Step Evidence Aggregation for LLM Agents

    Jun 5, 2026Vijitha Mittapalli, Shreyaa Jayant Dani, Satya Srujana Pilli +7Long-Horizon Agent EvaluationAI Agent Security

  2. From Reward-Hack Activations to Agentic Risk States: Context-Calibrated Mechanistic Monitoring in LLM Agents

    Jun 4, 2026Patrick Wilhelm, Odej KaoReward HackingAI Agent Safety

  3. Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage?

    Jun 4, 2026Jingheng Ye, Huiqi Zou, Simon Yu +1AI Coding AgentsAI Agent Monitoring

  4. From Agent Traces to Trust: A Survey of Evidence Tracing and Execution Provenance in LLM Agents

    Jun 3, 2026Yiqi Wang, Jiaqi Zhang, Taotao Cai +8AI Agent ReliabilityData Provenance

  5. Tracking the Behavioral Trajectories of Adapting Agents

    Jun 1, 2026Jonah Leshin, Manish Shah, Ian TimmisAgent EvaluationAI Agent Monitoring

  6. Monitoring Agentic Systems Before They're Reliable

    Jun 1, 2026Marisa Ferrara Boston, Glen Hanson, Effi Georgala +2Agent Failure AnalysisAgent Reliability

  7. POIROT: Interrogating Agents for Failure Detection in Multi-Agent Systems

    Jun 1, 2026Iñaki Dellibarda Varela, R. Sendra-Arranz, Pablo Romero-Sorozabal +5Agent Failure AnalysisAI Agent Reliability

  8. Agent System Operations: Categorization, Challenges, and Future Directions

    Jun 1, 2026Zexin Wang, Changhua Pei, Yuanhao Liu +10Agent Failure AnalysisAI Agent Monitoring

  9. Early Diagnosis of Wasted Computation in Multi-Agent LLM Systems via Failure-Aware Observability

    May 31, 2026Xianyou Li, Weiran Yan, Yichao Wu +4LLM Inference EfficiencyMulti-Agent LLM Systems

  10. BraveGuard: From Open-World Threats to Safer Computer-Use Agents

    May 31, 2026Yunhao Feng, Xiaohu Du, Xinhao Deng +13LLM GuardrailsAI Agent Safety

  11. Stateful Online Monitoring Catches Distributed Agent Attacks

    May 29, 2026Davis Brown, Samarth Bhargav, Arav Santhanam +7LLM Agent SecurityAI Agent Monitoring

  12. Syll: Open-Source Personal Automation with Cross-Surface Execution

    May 28, 2026Bo Zhang, Borui Zhang, Chenghao Jiang +5LLM Agent Skill LearningHuman-in-the-Loop AI

  13. Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms

    May 28, 2026Botao Amber Hu, Helena Rong, Max Van KleekAI Agent MonitoringAI Agent Governance

  14. Training Deliberative Monitors for Black-Box Scheming Detection

    May 28, 2026Aditya Sinha, Akshat Naik, Victor Gillioz +5AI Agent SafetyAI Agent Monitoring

  15. OpenClawBench: Benchmarking Process-side Anomalies in Real-world Agent Execution Trajectories

    May 28, 2026Yibing Liu, Yangze Liu, Xiaolong Yin +4Agent Failure AnalysisAnomaly Localization

  16. The Importance of Out-of-Band Metadata for Safe Autonomous Agents: The Redpanda Agentic Data Plane

    May 27, 2026Tyler Akidau, Tyler Rockwood, Johannes Brüderl +1AI Agent SecurityAI Agent Safety

  17. FinHarness: An Inline Lifecycle Safety Harness for Finance LLM Agents

    May 26, 2026Haoxuan Jia, Yang Liu, Bin Chong +10AI Agent ReliabilityLLM Guardrails

  18. Benchmarks are Not Enough: RAMP for Runtime Assessing of Agentic Models in Production Systems

    May 26, 2026Yipeng Ouyang, Xin Huang, Bingjie Liu +3Software Engineering AgentsLLM Agent Evaluation

  19. Retrying vs Resampling in AI Control

    May 25, 2026James Lucassen, Adam KaufmanAI Agent SafetyAI Agent Monitoring

  20. Causal Past Logic for Runtime Verification of Distributed LLM Agent Workflows

    May 20, 2026Benedikt BolligLLM Agent OrchestrationAgentic Workflows