Sandbox

Momentum

3 papers in the last four weeks, level with the four weeks before. 0.0% of all new papers.

Jul 13Week of Sep 28

Latest papers 24

All topics
CardsList
  1. Memory Compression for High-Fanout Agent Sandboxes

    Sep 12, 2026Mengming Li, Ceyu XU, Qijun Zhang +4SandboxTask-Aware Compression

  2. CityPlanner: A Sandbox Agent for Executable Urban Planning

    Sep 10, 2026Wentao Zhang, Jingyuan Wang, Zetong Zhou +2Classical PlanningUrban Environments

  3. Science sandboxes measure the scientific capability of AI agents

    Aug 31, 2026Arya S. Rao, Rodrigo I. Castro, Sager J. Gosai +8Scientific DiscoverySandbox

  4. One Success Isn't Reliability: Thinkingbox, a Sandbox and Benchmark for Agents in Stateful Business Workflows

    Aug 20, 2026Zhuochun Li, Youngmin Ko, Ali Keramati +11Agentic Workflow DesignSandbox

  5. OBLIVION: Workflow-Level Operational Skill Unlearning for Deployed Agents

    Aug 8, 2026Zhengyang Shan, Xu Qian, Jiayun Xin +3SkillsRemediation

  6. WeClawArena: An Auditable Sandbox and Benchmark for Cross-User Agents Collaboration and Security in Human-Centered Agent Networks

    Aug 4, 2026Prince Zizhuang Wang, Aojie Yuan, Haiyue Zhang +3OpenclawSandbox

  7. AI Sandbox: Technical Report

    Aug 2, 2026Muhammad Waseem, Md Aidul Islam, Md Nasir Uddin Shuvo +8SandboxArtificial Intelligence Governance

  8. SpecBox: Speculative Sandbox Scheduling for Efficient LLM Agent Serving

    Jul 27, 2026Yihui Zhang, Tianyu Wo, Jinghao Wang +7SandboxLarge Language Model Agents

  9. Branching Policy Optimization: Sandbox-Native Language Agent Reinforcement Learning

    Jul 15, 2026Bowei He, Yankai Chen, Xiaokun Zhang +1Frictive Policy OptimizationOffline Reinforcement Learning

  10. Seduced by the Narrative: Assessing Rule Adherence in Semi-Open Textual Sandboxes

    Jul 2, 2026Weiying Chen, Junlong Shen, Zhanyuan Guo +1RhetoricSandbox

  11. Code Isn't Memory: A Structural Codebase Index Inside a Coding Agent

    Jun 21, 2026Ishaan Bhola, Adithyan Krishnan, Sravanth Kurmala +1CodebasesCoding Agents

  12. AI Sandboxes: A Threat Model, Taxonomy, and Measurement Framework

    Jun 16, 2026Inderjeet Singh, Haitham Mahmoud, Andrés MurilloSandboxArtificial Intelligence Risk

  13. DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/Rollback

    May 21, 2026Yunpeng Dong, Jingkai He, Shiqi Liu +7SandboxRollback

  14. Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation

    May 11, 2026Zhixuan Shen, Jiawei Du, Ziyu Guo +5Open-WorldGrounding

  15. Crab: A Semantics-Aware Checkpoint/Restore Runtime for Agent Sandboxes

    Apr 30, 2026Tianyuan Wu, Chaokun Chang, Lunxi Cao +2SandboxIntermediate Checkpoints

  16. ClawGym: A Scalable Framework for Building Effective Claw Agents

    Apr 29, 2026Fei Bai, Huatong Song, Shuang Sun +11Claw-Like AgentAgentic Learning

  17. Mythos and the Unverified Cage: Z3-Based Pre-Deployment Verification for Frontier-Model Sandbox Infrastructure

    Apr 22, 2026Dominik BlainPre-Deployment Safety AssessmentsSandbox

  18. Context: Proactive Goal-Directed Intelligence via Composable Sandboxed Programs, Declarative Wiring, and Structured Interaction

    Apr 21, 2026Gregory MagarshakChatbotsInteraction Data

  19. Operationalising AI Regulatory Sandboxes: Activities, Requirements, and Technical Assessment under the EU AI Act

    Sep 27, 2025Alessio Buscemi, Thibault Simonetto, Daniele Pagani +3Eu Artificial Intelligence ActSandbox