cs.CCDec 4, 2025

Hardware-Algorithm Co-Optimization of Early-Exit Neural Networks for Multi-Core Edge Accelerators

Authors: Alaa Zniber, Arne Symons, Ouassim Karrakchou, Marian Verhelst, Mounir Ghogho

Organizations: TICLab, International University of Rabat, Morocco · MICAS, KU Leuven, Belgium

Abstract

The deployment of Early-Exiting Neural Networks (EENNs) on edge accelerators requires optimizing not only the network architecture but also its hardware deployment. Exit configuration, quantization, and hardware workload mapping interact in non-trivial ways, influencing memory traffic, accelerator utilization, and ultimately the energy-latency trade-off. This work presents a hardware-aware co-design framework for EENNs that jointly optimizes exit configuration, quantization-aware training, and multi-core hardware mapping within a unified NAS process. Leveraging analytical design space exploration, the framework identifies efficient workload mappings for each candidate architecture while providing accurate latency and energy estimates during the search. We further formulate EENN deployment as a constrained multi-objective optimization problem balancing predictive accuracy, energy-latency product, exit overhead, and dynamic inference efficiency. Experimental results on CIFAR-10 demonstrate that the proposed framework achieves over a 50% reduction in energy-latency product compared with static baselines under 8-bit quantization. These results demonstrate that jointly optimizing architecture and deployment is essential for realizing the full efficiency potential of dynamic inference on heterogeneous edge accelerators.

Figures & tables

Explore similar work

CardsList
  1. A Comparative Study of CNN Optimization Methods for Edge AI: Exploring the Role of Early Exits

    Apr 16, 2026Nekane Fernandez, Ivan Valdes, Steven Van Vaerenbergh +2Edge DevicesNeural Network Inference

  2. MiCoPro: End-to-End Mixed Precision HW/SW Co-design with HW-aware Proxy Model

    Aug 7, 2026Zijun Jiang, Yangdi LyuMixed-Precision QuantizationArtificial Intelligence Systems

  3. A Reconfigurable Multiplier Architecture for Error-Resilient Applications in RISC-V Core

    May 9, 2026Pragun Jaswal, L. Hemanth Krishna, B. SrinivasuRisc-VEdge Devices