Neural Policies

Latest papers 30

All topics
CardsList
  1. Reward as Observation: Learning Reward-Based Policies for Rapid Adaptation

    Sep 30, 2026Morgan Byrd, Maks Sorokin, Robert Wright +1Neural PoliciesRapid Adaptation

  2. Depot-Closed Multi-Component Construction for Neural Vehicle Routing

    Sep 28, 2026Shinichiro Hamada, Hisashi KashimaVehicle Routing ProblemNeural Combinatorial Optimization

  3. Mixed-Integer Nonlinear Differentiable Predictive Control for Underground Pumped Hydro Energy Storage Systems

    Sep 16, 2026Honghui Zheng, Ján Boldocký, Yury Dvorkin +1Model Predictive ControlNeural Policies

  4. Feedback-Modulated Harmonic Policies for Quadruped Locomotion

    Sep 16, 2026Yixuan Jia, Steven Roche, Jonathan P. HowLocomotionRobust Control

  5. SpikingNav: Robust Embodied Navigation with Spiking Neural Policies

    Aug 5, 2026Jiahong Zhang, Sijun Shen, Dehua Wu +5Spiking Neural NetworksNeuromorphic Computing

  6. Belief-Guided Decision Making with Uncertainty Gating in the Game of Go

    Jul 29, 2026Mehrad Yaghoubi, Azam Bastanfard, Abbas Jalilvand +1Monte Carlo Tree SearchChess

  7. PRISM: Polynomial Representations for Interaction-Structured Motor Control

    Jul 26, 2026Seung Hyun Lee, Stella X. YuNeural PoliciesMotion Control

  8. NSMA: Neuro-Symbolic Manifold Alignment for Generalizable Adaptive Bitrate Streaming under Texture Shift

    Jul 21, 2026Zhiqiang He, Zhi LiuNeural PoliciesManifolds

  9. Two Black Boxes, One Solver: Encoder Probing and Decoder Attribution for Neural Multi-Attribute VRP under Hard-Mask and Recourse Decoders

    Jul 5, 2026Sohaib AfifiVehicle Routing ProblemNeural Solvers

  10. ShardNet: Training Neural Controllers with Hard, Non-Convex Constraints

    Jun 29, 2026Long Kiu Chung, Shreyas KousikSafety ConstraintsLearning-Based Control

  11. Inference-Time Robot Behavior Steering through Physically-Aware Reconfiguration of Task-Structure

    Jun 25, 2026Yiyuan Pan, Hanjiang Hu, Shangtao Li +2Robot PoliciesInference-Time Steering

  12. Learning to Place Guards by Reinforcement: A Geo-Free Neural Policy for the Vertex-Guard Art Gallery Problem

    Jun 19, 2026Domagoj Ševerdija, Jurica Maltar, Nathan Chappel +1Neural Combinatorial OptimizationNeural Policies

  13. Heterogeneous Policy Networks for Composite Robot Team Communication and Coordination

    Jun 18, 2026Esmaeil Seraj, Rohan Paleja, Luis Pimentel +7Multi-Agent Reinforcement LearningHeterogeneous Robot Teams

  14. MirrorDuo: Reflection-Consistent Visuomotor Learning from Mirrored Demonstration Pairs

    Jun 18, 2026Zheyu Zhuang, Ruiyu Wang, Giovanni Luca Marchetti +2Behavior CloningVisuomotor Control

  15. Implicit Neural Representations of Individual Behavior

    Jun 10, 2026Andrew Kang, Priya NarasimhanNeural PoliciesNeural Representations

  16. DLO-Lab: Benchmarking Deformable Linear Object Manipulations with Differentiable Physics

    Jun 2, 2026Junyi Cao, Yian Wang, Ziyan Xiong +3Deformable Linear ObjectsRobot Systems

  17. Learning with Foresight: Enhancing Neural Routing Policy via Multi-Node Lookahead Prediction

    May 19, 2026Xia Jiang, Yaoxin Wu, Yew-Soon Ong +1Neural PoliciesLong-Horizon Planning

  18. Learning Bilevel Policies over Symbolic World Models for Long-Horizon Planning

    May 15, 2026Dillon Z. Chen, Till Hofmann, Toryn Q. Klassen +1Long-Horizon PlanningImitation Learning

  19. Distributed Zeroth-Order Policy Gradient for Networked Multi-agent Reinforcement Learning from Human Feedback

    May 15, 2026Pengcheng Dai, He Wang, Dongming Wang +2Multi-Agent Reinforcement LearningReinforcement Learning From Human Feedback

  20. Probabilistic Verification of Recurrent Neural Networks for Single and Multi-Agent Reinforcement Learning

    May 14, 2026Luca Marzari, Enrico MarchesiniReinforcement Learning With Verifiable RewardNeural Policies

  21. Causal Explanations from the Geometric Properties of ReLU Neural Networks

    May 11, 2026Hector Woods, Philippa Ryan, Rob AlexanderRectified Linear Unit NetworksCausal

  22. PMCTS: Principled Parallelized Inference Time Scaling with Particle Monte Carlo Tree Search

    May 9, 2026Yaniv Oren, Viliam Vadocz, Joery A. de Vries +3Monte Carlo Tree SearchMonte Carlo

  23. RN-D: Discretized Categorical Actors for On-Policy Reinforcement Learning

    Jan 30, 2026Yuexin Bian, Jie Feng, Tao Wang +3Soft Actor-CriticNeural Policies

  24. Neural Particle Automata: Learning Self-Organizing Particle Dynamics

    Jan 22, 2026Hyunsoo Kim, Ehsan Pajouheshgar, Sabine Süsstrunk +2Neural Cellular AutomataParticle Dynamics

  25. Survival Dynamics of Neural and Programmatic Policies in Evolutionary Reinforcement Learning

    Jan 7, 2026Anton Roupassov-Ruiz, Yiyang ZuoNeural PoliciesOn-Policy Self-Evolution

  26. Policy Gradient with Self-Attention for Model-Free Distributed Nonlinear Multi-Agent Games

    Sep 22, 2025Eduardo Sebastián, Maitrayee Keskar, Eeman Iqbal +3Multi-Agent Reinforcement LearningPolicy Gradient

  27. Sequential Cohort Selection under Uncertainty

    Aug 22, 2025Hortence Yiepnou, Christos DimitrakakisAlgorithmic FairnessCovariate Shift

  28. Instance-Conditioned Adaptation for Large-scale Generalization of Neural Routing Solver

    May 3, 2024Changliang Zhou, Xi Lin, Zhenkun Wang +3Vehicle Routing ProblemIntelligent Transportation