Action Selection

Momentum

9 papers in the last four weeks, up 200% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 43

All topics
CardsList
  1. Mimir: Physics-Grounded LLM Agents for Long-Horizon Irrigation Control

    Oct 1, 2026Yimeng Liu, Mi Zhang, Younsuk Dong +1Action SelectionWater

  2. Code Owns the Simulation, Jev Owns the Evaluation

    Oct 1, 2026Yaodong Yang, Hongyao Tang, Yi Ma +4Action Selection

  3. In CEM, a World Model Is Also a Proposal Mechanism

    Oct 1, 2026Oliver Obst, Frieder StolzenburgModel SelectionWorld Models

  4. WorldAuditBench: Interactive 3D World Auditing with Multimodal Agents

    Sep 30, 2026Ziyan Jiang, Jingbo Yang, Jiabao Ji +53D WorldMultimodal Agents

  5. Code to Control: Synthesizing Parameterized Reactive Controllers

    Sep 30, 2026Zergham Ahmed, Joshua B. Tenenbaum, Chris Bates +1Llm-Driven Code SynthesisReactive

  6. Learning to Explore Hidden Kinematics for Articulated Object Manipulation

    Sep 29, 2026Ruiyao Liu, Boshu Lei, Zhuoyang Pan +1Articulated KinematicsReal-World Manipulation Tasks

  7. D-JEPA: A Decision-Aligned Latent World Model

    Sep 21, 2026Shuaijun Liu, Chengyu Wu, Qifu Wen +5Latent World ModelsWorld Models

  8. Beyond Visual Quality: A Study of Test-Time Planning with World Action Models

    Sep 21, 2026Jianhao Yuan, Yu Yuan, Benjamin Ramtoula +5Action SelectionWorld Model Planning

  9. From Prediction to Decision: World-Model-Guided Action Selection for Continuous Pile Excavation

    Sep 14, 2026Ailing Zhang, Fan Gao, Song Zhang +3World ModelsAction Selection

  10. STAIR: Effective Incident Response Using an End-to-End Agentic Planning Framework

    Aug 10, 2026Hanlin Jiang, Jionghao Huang, Shaofei Li +6Incident ResponsePlanning

  11. ChronoState: Hidden Elapsed-Time Conditioning for Temporal-State Action Selection in Frozen-Backbone Language Models

    Aug 10, 2026Sam Siavoshian, Omar Ramadan, Amir K. Saeed +3ChronosFrozen Language Model

  12. StARS: Socially Appropriate Robot Actions via a Recommender System-Driven Approach

    Jul 23, 2026Erencem Ozbey, Fethiye Irmak Dogan, Jin Huang +1Human-Robot InteractionAction Selection

  13. How to Guide LLM Generation: Dual-Surrogate Guided Search for Automated Heuristic Design

    Jul 15, 2026Yuhan Wang, Chaoda Peng, Xingyu Wu +2Automated Heuristic DesignArtificial Intelligence Search

  14. COLMAR: Cooperative View Policy Learning for Multi-Agent Active 3D Reconstruction

    Jul 15, 2026Phu Pham, Damon Conover, Aniket Bera3D ReconstructionMulti-Objective Reinforcement Learning

  15. Strategic Buying Agents

    Jul 6, 2026Mingyang Fu, Ming HuAgentic CommerceDynamic Pricing

  16. EgoGapBench: Benchmarking Egocentric Action Selection in Multi-Agent Scenes

    Jul 1, 2026Jihyeok Jung, Jeewu Lee, Sanghyeop Kim +2Egocentric DatasetEgocentric Vision

  17. Ask the World Before Acting: Environment Probing for Calibrated Agent World Models

    Jun 30, 2026Xinyuan Song, Zekun CaiWorld ModelsAction Selection

  18. Constraint Tax in Open-Weight LLMs: An Empirical Study of Tool Calling Suppression Under Structured Output Constraints

    Jun 24, 2026Fangzheng Li, Aimin Zhang, Chen LvTool InvocationCall

  19. One-to-Two Acting: A Novel Framework for Single-arm Agent Action Expansion to Dual Arms

    Jun 18, 2026Youbin Yao, Nieqin Cao, Mingyan Li +3Bimanual ManipulationMultimodal Action Distributions

  20. Environment-Grounded Automated Prompt Optimization for LLM Game Agents

    Jun 16, 2026Rean Clive Fernandes, Lukas Fehring, Theresa Eimer +2Automatic Prompt OptimizationAgentic Optimization

  21. CCKS: Consensus-based Communication and Knowledge Sharing

    Jun 10, 2026Jinyuan Zu, Xiaowei Lv, Yongcai Wang +5Multi-Agent Reinforcement LearningDecentralized Learning

  22. From Reward-Hack Activations to Agentic Risk States: Context-Calibrated Mechanistic Monitoring in LLM Agents

    Jun 4, 2026Patrick Wilhelm, Odej KaoAgentic Reinforcement LearningLarge Language Model Agents

  23. Cross-Environment Neural Reranking for Sample-Efficient Action Selection in Text-Based Agents

    Jun 1, 2026Kan ShaoAction SelectionBert-Based Models

  24. TapSampling: Inference-Time Sampling with a Task-Progress-Understanding Verifier for Robotic Manipulation

    May 25, 2026Sizhe Zhao, Shengping Zhang, Shuo Yang +3Robotic ManipulationInference-Time Steering

  25. Know You Before You Speak: User-State Modeling for LLM Personalization in Multi-Turn Conversation

    May 23, 2026Jiani Luo, Xiaoyan Zhao, Yang Zhang +4Large Language Model PersonalizationPersonalization

  26. Offline Contextual Bandits in the Presence of New Actions

    May 18, 2026Ren Kishimoto, Tatsuhiro Shimizu, Kazuki Kawamura +6Contextual Bandit FrameworkOff-Policy Learning

  27. Think Twice, Act Once: Verifier-Guided Action Selection For Embodied Agents

    May 12, 2026Nishad Singhi, Christian Bialas, Snehal Jauhri +4Mllm-BasedAction Selection

  28. OLIVIA: Online Learning via Inference-time Action Adaptation for Decision Making in LLM ReAct Agents

    May 11, 2026Sheldon Yu, Junda Wu, Xintong Li +6Large Language Model AgentsPersonalized Large Language Model Agents

  29. Efficient Multi-Robot Motion Planning with Precomputed Translation-Invariant Edge Bundles

    May 10, 2026Himanshu Gupta, Paul Motter, Aritra Chakrabarty +5Multi-Robot Motion PlanningAction Selection

  30. Sequential Strategic Classification with Multi-Stage Selective Classifiers

    May 5, 2026Ziyuan Huang, Lina Alkarmi, Mingyan LiuSequential Decision MakingStrategic Agents

  31. cotomi Act: Learning to Automate Work by Watching You

    May 4, 2026Masafumi Oyamada, Kunihiro Takeoka, Kosuke Akimoto +5Browser AgentAutomated

  32. DRACULA: Hunting for the Actions Users Want Deep Research Agents to Execute

    Apr 26, 2026Nishant Balepur, Malachi Hamada, Varsha Kishore +9Action SelectionLong-Horizon Agents

  33. What Capable Agents Must Know: Selection Theorems for Robust Decision-Making under Uncertainty

    Mar 3, 2026Aran NayebiPartially Observable Markov Decision ProcessNext-State Prediction

  34. ProAct: A Benchmark and Multimodal Framework for Structure-Aware Proactive Response

    Feb 3, 2026Xiaomeng Zhu, Fengming Zhu, Weijie Zhou +8Multimodal AgentsAction Selection

  35. Deep Active Inference with Diffusion Policy and Multiple Timescale World Model for Real-World Exploration and Navigation

    Oct 27, 2025Riko Yokozawa, Kentaro Fujii, Yuta Nomura +1Robot NavigationDiffusion Policies

  36. Verifier-free Test-Time Sampling for Vision-Language-Action Models

    Oct 7, 2025Suhyeok Jang, Dongyoung Kim, Changyeon Kim +2Diffusion-Based Vision-Language-ActionsAction Selection

  37. Learning to Visually Connect Actions and their Effects

    Jan 19, 2024Paritosh Parmar, Eric Peh, Basura FernandoFine-Grained Video UnderstandingSelf-Supervised Learning