Fine-Grained Actions

Momentum

7 papers in the last four weeks, against 2 the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 30

All topics
CardsList
  1. EgoTools: Towards Tool-Centric Reasoning in Real-World Egocentric Videos

    Sep 30, 2026Shulin Tian, Junsu Kim, Shuai Liu +17Egocentric VideoEgocentric Vision

  2. EgoHumanoid-V2: Human-to-Humanoid Transfer of Coordinated Whole-Body Skills for Loco-Manipulation

    Sep 29, 2026Jin Chen, Yiming Jiang, Chongyang Xu +8Human-To-Robot TransferHumanoid Loco-Manipulation

  3. HEIR: Learning Human-Entity Interactions with Functional Roles

    Sep 28, 2026Di Wen, Wenhao Guo, Yuedong Tan +15Interaction DataFine-Grained Actions

  4. Probing Large Audio-Language Models for Compositional Understanding of Sounding Actions

    Sep 28, 2026Michel Olvera, Paraskevas Stamatiadis, Changhong Wang +1Large Audio Language ModelsFine-Grained Actions

  5. Understanding Clinical Cognitive Dialogues Using Large Language Models

    Sep 28, 2026Vishalakshi Arumugam, Dan Schumacher, Veronica Rammouz +3Cognitive DiagnosisClinical Reasoning Training

  6. KathDB-FAO: Synthesized Query Plans in a Multimodal DBMS

    Sep 23, 2026Guorui Xiao, Douglas Brown, Artur Borycki +1QueriesNatural Language

  7. The Fellowship of the Query: Learning Retrieval Actions

    Sep 23, 2026Mohammed Al-Maamari, Saber Zerhoudi, Michael Granitzer +1Fine-Grained ActionsModel Fine-Tuning

  8. AgenticRag-R1: Agentic Reinforcement Learning with Stack Memory for Multi-Step Reasoning, Retrieval and Memorizing

    Aug 30, 2026Xinke Jiang, Yue Fang, Zhibang Yang +12Agentic Reinforcement LearningAgentic Benchmarks

  9. Per-Shipment Multi-Agent Reinforcement Learning for Intermodal Freight Routing Under Hurricane Disruption

    Aug 7, 2026Aliza Sharmin, Xudong Wang, Mustafa Can Camur +1LogisticsMulti-Agent Reinforcement Learning

  10. TARL: Transaction-Aware Reliable Ledgers for Executable Memory Management in Long-Term Agents

    Aug 4, 2026Han Xiao, Hongjun Xu, Xin Zhang +2Long-Term Agent MemoryPeak Memory

  11. MoHallBench: A Benchmark for Motion Hallucination in Video Large Language Models

    Jul 1, 2026Sihan Chen, Jiale Li, Jianghang Lin +1Large Language Model HallucinationObject Hallucination

  12. Gold Points Sniper: Self-guided Visual Reasoning in VLM for Fine-grained Action Understanding

    Jun 21, 2026Haodi Liu, Xinhang Yang, Kunda Yan +3Fine-Grained ActionsOpen-Vocabulary Action Recognition

  13. PearlVLA: Progressive Embodied Action-Plan Refinement in Latent Space

    Jun 16, 2026Bochen Yang, Lianlei ShanFine-Grained ActionsLatent Space

  14. A New Multi-Domain Benchmark for Micro-Action Recognition and Detection

    Jun 12, 2026Yanbin Hao, Pengyu Liu, Xing Wei +3Fine-Grained ActionsEmotion Recognition

  15. OR-Action: Multi-Role Video Understanding with Fine-Grained Actions

    Jun 11, 2026Felix Tristram, Ege Özsoy, Christian Benz +3Open-Vocabulary Action RecognitionFine-Grained Actions

  16. GEAR-VLA: Learning Geometry-Aware Action Representations for Generalizable Robotic Manipulation

    Jun 7, 2026Yuan Zhang, Shiqi Zhang, Yedong Shen +11Robotic ManipulationFine-Grained Actions

  17. Robotic Policy Adaptation via Weight-Space Meta-Learning

    Jun 5, 2026Christian Bianchi, Siamak Yousefi, Alessio Sampieri +4Robot PoliciesRobotic Manipulation

  18. Understanding-Enhanced Model Collaboration for Long-Tailed Egocentric Mistake Detection

    Jun 1, 2026Boyu Han, Qianqian Xu, Shilong Bao +3Egocentric VideoFine-Grained Video Understanding

  19. HiERO-StepG @ Ego4D Step Grounding Challenge: hierarchical activity understanding enables zero-shot step grounding

    May 29, 2026Andrea Zenotto, Simone Alberto Peirone, Francesca Pistilli +1Fine-Grained ActionsGrounding

  20. Zero-Shot Temporal Action Localization Through Textual Guidance

    May 21, 2026Benedetta Liberatori, Alessandro Conti, Lorenzo Vaquero +3Temporal Action LocalizationVideo Temporal Grounding

  21. TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning

    May 18, 2026ZhiYuan Feng, Yu Deng, Ruichuan An +11Fine-Grained Actions

  22. Beyond Binary: Reframing GUI Critique as Continuous Semantic Alignment

    May 14, 2026Yuchen Sun, Pei Fu, Shaojie Zhang +6Graphical User InterfaceFine-Grained Actions

  23. Pro2^2Assist: Continuous Step-Aware Proactive Assistance with Multimodal Egocentric Perception for Long-Horizon Procedural Tasks

    May 5, 2026Lilin Xu, Bufang Yang, Siyang Jiang +6Multimodal AgentsArtificial Intelligence Assistants

  24. FineState-Bench: Benchmarking State-Conditioned Grounding for Fine-grained GUI State Setting

    Apr 30, 2026Fengxian Ji, Jingpu Yang, Zirui Song +5Graphical User InterfaceWeak Visual Grounding

  25. Mini-BEHAVIOR-Gran: Revealing U-Shaped Effects of Instruction Granularity on Language-Guided Embodied Agents

    Apr 18, 2026Sukai Huang, Chenyuan Zhang, Fucai Ke +4Fine-Grained ActionsEmbodied Agents

  26. Beyond Description: Cognitively Benchmarking Fine-Grained Action for Embodied Agents

    Nov 24, 2025Dayong Liu, Chao Xu, Weihong Chen +5Embodied AgentsFine-Grained Actions