cs.DCSep 29, 2026

ARGOS: Reinforcement Learning-Driven Multidimensional Elasticity for Service Orchestration in the Computing Continuum

Authors: Javier Mateos-Bravo, Sergio Laso, Juan Luis Herrera, Ilir Murturi, Pantelis Frangoudis, Schahram Dustdar

Organizations: Department of Computer Science and Telematics Engineering, University of Extremadura, Spain · Global Process and Product Improvement S.L., Spain · Department of Mechatronics, University of Prishtina, Kosova · Distributed Systems Group, TU Wien, Austria · ICREA, Barcelona, Spain

Abstract

Data-intensive services in the Computing Continuum must balance analytics quality, resource usage, and cost across heterogeneous nodes with limited and uneven capacity. This balance becomes especially difficult when resource scaling reaches capacity limits, because changes in demand and cluster pressure must then be absorbed without violating client-defined quality ranges. Existing orchestrators mainly adapt resources, placements, or replicas, while analytics requirements such as coverage, sample, and freshness remain fixed. This article presents ARGOS, the Adaptive Reinforcement Learning-Driven Governance for Orchestrated Services, an end-to-end controller that formulates multidimensional elasticity as a per-request Markov decision process over analytics quality and cluster pressure, supported by capacity-aware admission. ARGOS is evaluated under controlled workloads and time-varying multi-tenant arrivals on a heterogeneous cluster. Across the controlled scenarios, the deep reinforcement learning policies consistently outperform the non-learning baselines and approach the independently tuned best-fixed reference. A separate live evaluation reports improvements over the static midpoint under realistic and saturated arrivals, with no recorded CPU or memory violations but remaining coverage violations. These results support deep reinforcement learning as an adaptive mechanism for multidimensional elasticity when resource scaling alone is insufficient.

Figures & tables

Explore similar work

CardsList
  1. When Does Deep RL Beat Calibrated Baselines? A Benchmark Study on Adaptive Resource Control

    May 26, 2026Guilin Zhang, Chuanyi Sun, Kai Zhao +3Rate ScalingReinforcement Learning Control

  2. Incentives and Evidence in Learned Service Orchestration

    Jun 15, 2026Syed Izhan Khilji, Alireza Furutanpey, Schahram DustdarOrchestrationKubernetes

  3. Learning to Remember: Attentive Reinforcement Learning for Edge Serverless Autoscaling

    Mar 21, 2026Faraz Shaikh, Gianluca Reali, Mauro FemminellaEdge PlatformsRate Scaling