cs.ROOct 5, 2026

Visual Swarm Navigation via Deep Reinforcement Learning and Evolutionary Hybrid Design

Authors: Álvaro Díez, Fidel Aznar

Organizations: Department of Computer Science and Artificial Intelligence, University of Alicante

Abstract

Swarm robotics presents a robust and cost-effective paradigm for advanced automation in complex, dynamic environments, such as those encountered in search and rescue or environmental monitoring. A fundamental challenge for this field is the data-driven design of decentralized controllers capable of generating emergent collective behaviors. This paper proposes a novel, AI-driven hybrid methodology for the automatic synthesis of swarm robotic controllers for autonomous visual navigation. This approach synergistically combines multi-agent reinforcement learning with neuro-evolutionary strategies, specifically leveraging implementations of the cross-entropy method and the covariance matrix adaptation evolution strategy to optimize a pre-trained individual navigation policy. The underlying deep architecture is engineered for low-cost, resource-constrained platforms, utilizing a compact neural network that relies exclusively on monocular camera imagery. This vision-based design emphasizes computational and energy efficiency, a critical requirement for practical swarm deployments. Experiments, performed in a high-fidelity physics simulator, demonstrate that the resulting controllers enable robust and scalable collective exploration of diverse indoor environments. The controller trained using our cross-entropy method achieves superior exploration coverage, visiting 36.20% more regions compared to the covariance matrix adaptation evolution strategy. Critically, our best vision-based policy achieves exploration performance statistically comparable to traditional methods relying on more expensive distance sensors, while delivering a significant 31.40% average reduction in energy consumption. These findings validate an effective and economically viable autonomous control system, establishing a path for deploying highly efficient collective intelligence in real-world engineering applications.

Figures & tables

Explore similar work

CardsList
  1. Dual Variational Autoencoders for Efficient Sim-to-Real Transfer in Low-Cost Robotic Navigation

    Oct 5, 2026Álvaro Díez, Fidel AznarSim-To-Real Reinforcement LearningRobot Navigation

  2. SwarmNxt: Open-source Software-Hardware Platform for Fast and Agile Aerial Swarms

    Sep 11, 2026Charbel Toumieh, Niel Mistry, Benjamin Jarvis +4Unmanned Aerial Vehicle SwarmsAutonomous Drones

  3. Decentralized Scalable Exploration via Emergent Adaptive Lévy Walks on Minimal-Sensing Platforms

    Jul 28, 2026Wai Lun Leong, Teo Swee Huat RodneyMulti-UavUnmanned Aerial Vehicles